Hugging Face Blog
huggingface.co · EN
- 4posts a week
- 24 Septlast post
- #19in AI
The Hugging Face blog
Subscribe
in the reader on this device — no account. Or use any reader:
Latest posts
Accelerating vision-language models with LFM2.5-VL-DSpark
How to Use NVIDIA Warp and MjWarp to Accelerate Robotics Simulation and Learning Workflows
How UK AISI and EvalEval Are Making Benchmark Results Reproducible
Transformers now runs llama.cpp quants
Jun Kim, oMLX creator and maintainer, joins Hugging Face to support the MLX community
Pruning LLMs Like a Physicist: Block Removal as an Ising Optimization Problem
tokenizers v1: encode, decode and scaling, measured
Your Agent Aced the Task. Will It Do It Again?
Async GRPO with LoRA across HF Jobs: a bucket, a proxy, and no NCCL
Rebuilding AUTOMATIC1111 with Gradio Workflow
NeoMME: an efficient Multimodal-native and Multilingual Encoder
Fine-tuning a 350M Model for Better Structured Outputs in 100 GRPO Steps
Give Your Coding Agents a Memory You Own
Training a coding model to paint watercolours with TRL and OpenEnv
BenchMIRT: What are LLM benchmarks actually measuring?
Introducing @huggingface/kernels: 200+ WebGPU Kernels for Local AI
The Open ASR Leaderboard Adds Its First Global South Language
Training and Finetuning Multi-Vector Embedding Models with Sentence Transformers
Granite 4.2 LLMs: How They're Built
Quantization-Aware Healing: a compressed, 4-bit model that outperforms its full-precision original
In the directory since 25 Sept 2026 · last checked 25 Sept 2026 · RSS