Accelerating vision-language models with LFM2.5-VL-DSpark #AI #MachineLearning #ComputerVision
How to Use NVIDIA Warp and MjWarp to Accelerate Robotics Simulation and Learning Workflows #AI #Robotics #Simulation
**Know Who Spoke When: Build Real-Time, Multi-Speaker AI with NVIDIA Nemotron 3 Diarization** #AI #MachineLearning #OpenSource
How UK AISI and EvalEval Are Making Benchmark Results Reproducible #AI #MachineLearning #OpenSource
Transformers now runs llama.cpp quants
Jun Kim, oMLX creator and maintainer, joins Hugging Face to support the MLX community #MachineLearning #OpenSource #AI
Pruning LLMs Like a Physicist: Block Removal as an Ising Optimization Problem
tokenizers v1: encode, decode and scaling, measured
Your Agent Aced the Task. Will It Do It Again? #MachineLearning #AI #Research
Async GRPO with LoRA across HF Jobs: a bucket, a proxy, and no NCCL #MachineLearning #AI #OpenSource
Rebuilding AUTOMATIC1111 with Gradio Workflow #MachineLearning #AI #OpenSource
IBM releases SOTA Granite Time Series PatchTST-FM-r2 model with commercial-friendly license #MachineLearning #TimeSeries #AI
Safety for Whom? Refusing the Right Subset of a Topic, Not the Whole Topic #AI #MachineLearning #Ethics
*NeoMME*: an efficient Multimodal-native and Multilingual Encoder #AI #MachineLearning #NLP
Training a coding model to paint watercolours with TRL and OpenEnv #MachineLearning #AI #OpenSource
Give Your Coding Agents a Memory You Own #AI #MachineLearning #Coding
Fine-tuning a 350M Model for Better Structured Outputs in 100 GRPO Steps #MachineLearning #AI #OpenSource
Real-Time Intelligence with IBM Time Series Models on Confluent #AI #MachineLearning #TimeSeries
BenchMIRT: What are LLM benchmarks actually measuring? #AI #LLM #Research
Introducing @huggingface/kernels: 200+ WebGPU Kernels for Local AI #MachineLearning #AI #WebGPU
The Open ASR Leaderboard Adds Its First Global South Language #MachineLearning #OpenSource #AI
Training and Finetuning Multi-Vector Embedding Models with Sentence Transformers #MachineLearning #AI #NLP
Granite 4.2 LLMs: How They're Built #AI #LLM #OpenSource
Extremely Fast and Accurate Transcription with Granite Speech 5.0 Turbo CTC #AI #SpeechRecognition #MachineLearning
Quantization-Aware Healing: a compressed, 4-bit model that outperforms its full-precision original #MachineLearning #AI #Quantization
Wire It, Run It, Deploy It: AI Workflows in Gradio #MachineLearning #AI #Gradio
Measuring benchmark optimization in speech recognition #MachineLearning #SpeechRecognition #AI
How Hugging Face Inference Endpoints, Jobs, and Buckets Power Search on Papers with Code #MachineLearning #AI #OpenSource
Up to 3.2x Faster Inference with LFM2.5-DSpark #MachineLearning #AI #Inference
LFM2.5 Q4\_0 Checkpoints from Quantization-Aware Distillation #MachineLearning #AI #OpenSource
How Much Memory Does Your Agent Actually Need? #MachineLearning #AI #DeepLearning
Multi-Vector (Late Interaction) Embedding Models with Sentence Transformers #MachineLearning #AI #NLP
Same Cluster, 33 Points More Utilization: What Changed Was the Order #MachineLearning #AI #OpenSource
State of Open Models: Summer 2026 Observations #OpenSource #AI #MachineLearning
Record, train, and deploy from one place with Strands Agents, LeRobot, and Hugging Face Storage Buckets #MachineLearning #AI #OpenSource
What We Learned by Reproducing 2,200 papers from ICML #MachineLearning #AI #Research
Introducing OlmoEarth embeddings: Custom embedding exports from OlmoEarth Studio for downstream analysis #MachineLearning #Embeddings #AI
LFM2.5-VL-3B for Better and Faster Vision Capabilities for the Edge #AI #MachineLearning #ComputerVision
Thinking of ACE? We Can Do It with Fewer Tokens #MachineLearning #AI #NLP
Build Low-Latency Multilingual Voice Agents: Open Weights & Full Deployment Control with NVIDIA Magpie TTS #AI #VoiceAI #OpenSource