Falcon-Emirati: When an LLM Learns the Dialect, the Culture, and the Nuance
Hugging Face · Hugging Face Blog · Oct 06, 2026
Falcon-Emirati LLM is trained to understand Emirati dialect, culture, and nuance.
The Agent Said It Was Done. The Database Disagreed.
Hugging Face · Hugging Face Blog · Oct 03, 2026
No content provided to analyze for AI signal extraction.
Open-sourcing AstaBrief, the fast report-generation model in Asta
Hugging Face · Hugging Face Blog · Oct 02, 2026
Hugging Face open-sources AstaBrief, its fast report-generation model from Asta
AutoSynthData: Generating Training Data for Enterprise Agents
Hugging Face · Hugging Face Blog · Oct 02, 2026
No content was provided to extract signal from.
Introducing Olmo-core 3: Open, scalable training infrastructure for large MoEs
Hugging Face · Hugging Face Blog · Oct 01, 2026
Hugging Face releases OLMo-core 3, open-source scalable training infrastructure for large MoE models
Open TTS Leaderboard: Scalable Evaluation for Multilingual Text-to-Speech and Voice Cloning
Hugging Face · Hugging Face Blog · Sep 30, 2026
Hugging Face launches open leaderboard for standardized multilingual TTS and voice cloning evaluation
NVIDIA Kumo Tabular Sets a New Accuracy-Efficiency Frontier for Tabular Prediction
Hugging Face · Hugging Face Blog · Sep 29, 2026
NVIDIA Kumo Tabular claims a new accuracy-efficiency frontier for tabular prediction tasks
Getting the Source Right, Not Just the Fact: Source-Aware Verification for MCP Agents
Hugging Face · Hugging Face Blog · Sep 29, 2026
MCP agents require source-aware verification, not just factual accuracy checks
Holo4: powering generalist computer-use agents
Hugging Face · Hugging Face Blog · Sep 28, 2026
Hugging Face announces Holo4, a generalist computer-use agent system
Accelerating vision-language models with LFM2.5-VL-DSpark
Hugging Face · Hugging Face Blog · Sep 24, 2026
Parse error
How to Use NVIDIA Warp and MjWarp to Accelerate Robotics Simulation and Learning Workflows
Hugging Face · Hugging Face Blog · Sep 23, 2026
Parse error
How UK AISI and EvalEval Are Making Benchmark Results Reproducible
Hugging Face · Hugging Face Blog · Sep 22, 2026
Parse error
Transformers now runs llama.cpp quants
Hugging Face · Hugging Face Blog · Sep 22, 2026
Parse error
Jun Kim, oMLX creator and maintainer, joins Hugging Face to support the MLX community
Hugging Face · Hugging Face Blog · Sep 22, 2026
Parse error
Pruning LLMs Like a Physicist: Block Removal as an Ising Optimization Problem
Hugging Face · Hugging Face Blog · Sep 21, 2026
Parse error
tokenizers v1: encode, decode and scaling, measured
Hugging Face · Hugging Face Blog · Sep 21, 2026
Parse error
Your Agent Aced the Task. Will It Do It Again?
Hugging Face · Hugging Face Blog · Sep 15, 2026
Hugging Face demonstrates consistent AI performance across tasks.
“Your agent aced the task.”
Rebuilding AUTOMATIC1111 with Gradio Workflow
Hugging Face · Hugging Face Blog · Sep 10, 2026
Hugging Face announces rebuilding AUTOMATIC1111 using Gradio Workflow.
IBM releases SOTA Granite Time Series PatchTST-FM-r2 model with commercial-friendly license
Hugging Face · Hugging Face Blog · Sep 09, 2026
IBM launches a new commercial-friendly time series model.
Safety for Whom? Refusing the Right Subset of a Topic, Not the Whole Topic
Hugging Face · Hugging Face Blog · Sep 08, 2026
AI safety should address specific subsets rather than avoiding the entire topic.
NeoMME: an efficient Multimodal-native and Multilingual Encoder
Hugging Face · Hugging Face Blog · Sep 03, 2026
Hugging Face introduces NeoMME, a novel multimodal-native and multilingual encoder.
“NeoMME demonstrates a new way to seamlessly integrate multiple modalities in AI.”
Fine-tuning a 350M Model for Better Structured Outputs in 100 GRPO Steps
Hugging Face · Hugging Face Blog · Sep 03, 2026
Hugging Face fine-tuned a 350M model for improved structured outputs.
Give Your Coding Agents a Memory You Own
Hugging Face · Hugging Face Blog · Sep 03, 2026
Hugging Face introduces a memory system for coding agents.
“No quote provided.”
Training a coding model to paint watercolours with TRL and OpenEnv
Hugging Face · Hugging Face Blog · Sep 03, 2026
Hugging Face introduces a coding model capable of painting watercolours.
Real-Time Intelligence with IBM Time Series Models on Confluent
Hugging Face · Hugging Face Blog · Sep 02, 2026
IBM is introducing time series models for real-time intelligence on Confluent.
BenchMIRT: What are LLM benchmarks actually measuring?
Hugging Face · Hugging Face Blog · Sep 01, 2026
No content was provided to extract a signal from.
Introducing @huggingface/kernels: 200+ WebGPU Kernels for Local AI
Hugging Face · Hugging Face Blog · Sep 01, 2026
Hugging Face launches 200+ WebGPU kernels enabling local AI inference in browsers
The Open ASR Leaderboard Adds Its First Global South Language
Hugging Face · Hugging Face Blog · Aug 28, 2026
Hugging Face Open ASR Leaderboard adds its first Global South language benchmark.
Training and Finetuning Multi-Vector Embedding Models with Sentence Transformers
Hugging Face · Hugging Face Blog · Aug 26, 2026
Hugging Face releases guidance on training multi-vector embedding models with Sentence Transformers
Granite 4.2 LLMs: How They're Built
Hugging Face · Hugging Face Blog · Aug 25, 2026
No content was provided for analysis.
Quantization-Aware Healing: a compressed, 4-bit model that outperforms its full-precision original
Hugging Face · Hugging Face Blog · Aug 25, 2026
Hugging Face's 4-bit quantized model outperforms its full-precision original via Quantization-Aware Healing.
Wire It, Run It, Deploy It: AI Workflows in Gradio
Hugging Face · Hugging Face Blog · Aug 25, 2026
Hugging Face announced AI workflow capabilities in the Gradio framework.
Measuring benchmark optimization in speech recognition
Hugging Face · Hugging Face Blog · Aug 21, 2026
Hugging Face published research on benchmark optimization in speech recognition
Up to 3.2x Faster Inference with LFM2.5-DSpark
Hugging Face · Hugging Face Blog · Aug 20, 2026
LFM2.5-DSpark achieves up to 3.2x faster inference speeds over baseline
LFM2.5 Q4\_0 Checkpoints from Quantization-Aware Distillation
Hugging Face · Hugging Face Blog · Aug 19, 2026
Hugging Face released LFM2.5 Q4_0 checkpoints via quantization-aware distillation
How Much Memory Does Your Agent Actually Need?
Hugging Face · Hugging Face Blog · Aug 18, 2026
Hugging Face published guidance on memory requirements for AI agents
Multi-Vector (Late Interaction) Embedding Models with Sentence Transformers
Hugging Face · Hugging Face Blog · Aug 18, 2026
Hugging Face introduces multi-vector late interaction embedding support in Sentence Transformers
Same Cluster, 33 Points More Utilization: What Changed Was the Order
Hugging Face · Hugging Face Blog · Aug 17, 2026
Hugging Face blog post content was not provided for analysis.
State of Open Models: Summer 2026 Observations
Hugging Face · Hugging Face Blog · Aug 14, 2026
No content provided to extract signal from.
Record, train, and deploy from one place with Strands Agents, LeRobot, and Hugging Face Storage Buckets
Hugging Face · Hugging Face Blog · Aug 13, 2026
Hugging Face integrates Strands Agents, LeRobot, and Storage Buckets into a unified robotics pipeline.
What We Learned by Reproducing 2,200 papers from ICML
Hugging Face · Hugging Face Blog · Aug 13, 2026
Content body was empty; no extractable signal beyond the title.
Introducing OlmoEarth embeddings: Custom embedding exports from OlmoEarth Studio for downstream analysis
Hugging Face · Hugging Face Blog · Aug 12, 2026
Hugging Face launches OlmoEarth Studio embedding exports for downstream analysis
LFM2.5-VL-3B for Better and Faster Vision Capabilities for the Edge
Hugging Face · Hugging Face Blog · Aug 12, 2026
LFM2.5-VL-3B delivers improved vision-language capabilities optimized for edge deployment
Thinking of ACE? We Can Do It with Fewer Tokens
Hugging Face · Hugging Face Blog · Aug 11, 2026
Hugging Face claims ACE reasoning tasks can be done with fewer tokens
Build Low-Latency Multilingual Voice Agents: Open Weights & Full Deployment Control with NVIDIA Magpie TTS
Hugging Face · Hugging Face Blog · Aug 10, 2026
NVIDIA Magpie TTS offers open-weights multilingual voice agents with full self-hosted deployment control.
Making Knowledge Distillation Cheap Enough to Run at Scale
Hugging Face · Hugging Face Blog · Aug 10, 2026
Hugging Face claims knowledge distillation can now be run cheaply at scale
TutorMoments: Do AI tutors know when to help and when to hold back?
Hugging Face · Hugging Face Blog · Aug 07, 2026
Hugging Face introduces TutorMoments benchmark for AI tutoring intervention timing
Baseten on Hugging Face Inference Providers 🔥
Hugging Face · Hugging Face Blog · Aug 06, 2026
Baseten joins Hugging Face as an inference provider partner.
Deploy local agents everywhere with LFM2.5-2.6B
Hugging Face · Hugging Face Blog · Aug 04, 2026
Hugging Face releases LFM2.5-2.6B, a compact model designed for local agent deployment.
GPU Management: Why Idle GPUs Are the New Grounded Aircraft
Hugging Face · Hugging Face Blog · Jul 30, 2026
Idle GPUs represent wasted capital analogous to grounded aircraft in fleet economics
The OlmoEarth Platform: Geospatial inference at planetary scale
Hugging Face · Hugging Face Blog · Jul 28, 2026
Hugging Face announces OlmoEarth, a geospatial inference platform at planetary scale.
LFM2.5-Encoders for Fast Long-Context Inference on CPU
Hugging Face · Hugging Face Blog · Jul 28, 2026
Hugging Face releases LFM2.5-Encoders optimized for fast long-context inference on CPU hardware
NVIDIA Cosmos-H-Dreams: Bringing Real-Time Generative Simulation to Surgical Robotics
Hugging Face · Hugging Face Blog · Jul 27, 2026
NVIDIA Cosmos-H-Dreams enables real-time generative simulation for surgical robotics training.
Anatomy of a Frontier Lab Agent Intrusion: A Technical Timeline of the July 2026 Incident
Hugging Face · Hugging Face Blog · Jul 27, 2026
Hugging Face published a technical post-mortem of a July 2026 frontier lab agent intrusion.
Bringing Nunchaku 4-bit Diffusion Inference to Diffusers
Hugging Face · Hugging Face Blog · Jul 23, 2026
Hugging Face integrates Nunchaku 4-bit quantization into Diffusers for faster diffusion inference
The State of Simulation for Physical AI: An Overview
Hugging Face · Hugging Face Blog · Jul 21, 2026
No content was provided to extract a signal from.
Grabette: an open system to record robot-manipulation data
Hugging Face · Hugging Face Blog · Jul 21, 2026
Hugging Face releases Grabette, an open system for recording robot manipulation data
Introducing Cosmos 3 Edge
Hugging Face · Hugging Face Blog · Jul 20, 2026
Hugging Face announced Cosmos 3 Edge but content body is empty.
Fine-tune video and image models at scale with NVIDIA NeMo Automodel and 🤗 Diffusers
Hugging Face · Hugging Face Blog · Jul 17, 2026
NVIDIA NeMo Automodel integrates with Hugging Face Diffusers for scaled fine-tuning of video and image models
NVIDIA Nemotron 3 Embed Ranks #1 Overall on RTEB, Advancing Agentic Retrieval
Hugging Face · Hugging Face Blog · Jul 16, 2026
Parse error
Newer Models, Same Advantage
Hugging Face · Hugging Face Blog · Jul 16, 2026
Parse error
Security incident disclosure — July 2026
Hugging Face · Hugging Face Blog · Jul 16, 2026
Parse error
What building Shippy taught us about building agents
Hugging Face · Hugging Face Blog · Jul 15, 2026
Parse error
Model Routing Is Simple. Until It Isn’t.
Hugging Face · Hugging Face Blog · Jul 15, 2026
Parse error
Welcome Inkling by Thinking Machines
Hugging Face · Hugging Face Blog · Jul 15, 2026
Parse error
Introducing Real World VoiceEQ: Measuring the human quality of voice AI
Hugging Face · Hugging Face Blog · Jul 15, 2026
Parse error
Profiling in PyTorch (Part 3): Attention is all you profile
Hugging Face · Hugging Face Blog · Jul 10, 2026
Parse error
Data for Agents
Hugging Face · Hugging Face Blog · Jul 08, 2026
Parse error
Native-speed vLLM transformers modeling backend
Hugging Face · Hugging Face Blog · Jul 08, 2026
Parse error
From Hugging Face to Amazon SageMaker Studio in one click
Hugging Face · Hugging Face Blog · Jul 07, 2026
Parse error
Hugging Face Models on Foundry Managed Compute
Hugging Face · Hugging Face Blog · Jul 07, 2026
Parse error
LeRobot v0.6.0: Imagine, Evaluate, Improve
Hugging Face · Hugging Face Blog · Jul 07, 2026
Parse error
Run AI workloads on any cloud, store on Hugging Face: zero-egress storage with SkyPilot
Hugging Face · Hugging Face Blog · Jul 07, 2026
Parse error
PRX Part 4: Our Data Strategy
Hugging Face · Hugging Face Blog · Jul 06, 2026
Parse error
🤗 Kernels: Major Updates
Hugging Face · Hugging Face Blog · Jul 06, 2026
Parse error
Hugging Face and Cerebras bring Gemma 4 to real-time voice AI
Hugging Face · Hugging Face Blog · Jul 01, 2026
Parse error
ScarfBench: Benchmarking AI Agents for Enterprise Java Framework Migration
Hugging Face · Hugging Face Blog · Jun 30, 2026
Parse error
Why Specialization Is Inevitable
Hugging Face · Hugging Face Blog · Jun 30, 2026
Parse error
Featuring Every Eval Ever Results on Hugging Face Model Pages
Hugging Face · Hugging Face Blog · Jun 30, 2026
Parse error
DiScoFormer: One transformer for density and score, across distributions
Hugging Face · Hugging Face Blog · Jun 29, 2026
DiScoFormer introduces a single transformer that models both density and score across multiple distributions.
Run a vLLM Server on HF Jobs in One Command
Hugging Face · Hugging Face Blog · Jun 26, 2026
Hugging Face enables running a vLLM inference server on HF Jobs with a single command.
Which tokens does a hybrid model predict better?
Hugging Face · Hugging Face Blog · Jun 25, 2026
Hugging Face analyzes which token types hybrid attention-SSM models predict better than transformers.
Accelerating Transformers Fine-Tuning with NVIDIA NeMo AutoModel
Hugging Face · Hugging Face Blog · Jun 24, 2026
NVIDIA NeMo AutoModel accelerates fine-tuning of Hugging Face Transformers models.
Introducing the FFASR Leaderboard: Benchmarking ASR in the Real World
Hugging Face · Hugging Face Blog · Jun 24, 2026
Hugging Face launches the FFASR Leaderboard to benchmark automatic speech recognition on real-world audio.
Build real agentic apps using CUGA: two dozen working examples on a lightweight harness
Hugging Face · Hugging Face Blog · Jun 23, 2026
Hugging Face showcases CUGA, a lightweight harness with two dozen working examples for building real agentic apps.
Shipping huggingface_hub every week with AI, open tools, and a human in the loop
Hugging Face · Hugging Face Blog · Jun 23, 2026
Hugging Face ships huggingface_hub weekly using AI agents with a human reviewer in the loop.
Experimenting with the proposed Cross-Origin Storage API in Transformers.js
Hugging Face · Hugging Face Blog · Jun 23, 2026
Hugging Face is experimenting with the proposed Cross-Origin Storage API in Transformers.js.
PP-OCRv6 on Hugging Face: 50-Language OCR from 1.5M to 34.5M Parameters
Hugging Face · Hugging Face Blog · Jun 22, 2026
Hugging Face released PP-OCRv6, a 50-language OCR model family ranging from 1.5M to 34.5M parameters.
MosaicLeaks: Can your research agent keep a secret?
Hugging Face · Hugging Face Blog · Jun 18, 2026
Hugging Face demonstrates research agents can leak confidential data via prompt-injection attacks.
Beyond LoRA: Can you beat the most popular fine-tuning technique?
Hugging Face · Hugging Face Blog · Jun 18, 2026
Hugging Face explores fine-tuning methods that may outperform LoRA, the most popular technique.
Is it agentic enough? Benchmarking open models on your own tooling
Hugging Face · Hugging Face Blog · Jun 18, 2026
Hugging Face explores benchmarking open models for agentic capability against your own tooling.
MolmoMotion: Language-guided 3D motion forecasting
Hugging Face · Hugging Face Blog · Jun 17, 2026
Hugging Face introduces MolmoMotion, a model for language-guided 3D motion forecasting.
From the Hugging Face Hub to robot hardware with Strands Agents and LeRobot
Hugging Face · Hugging Face Blog · Jun 17, 2026
Hugging Face shows deploying models from the Hub to robot hardware using Strands Agents and LeRobot.
GLM-5.2: Built for Long-Horizon Tasks
Hugging Face · Hugging Face Blog · Jun 17, 2026
GLM-5.2 is a new model from Hugging Face optimized for long-horizon, multi-step agentic tasks.
Agentic Resource Discovery: Let agents search
Hugging Face · Hugging Face Blog · Jun 17, 2026
Hugging Face promotes agentic resource discovery, letting AI agents search for resources autonomously.
olmo-eval: An evaluation workbench for the model development loop
Hugging Face · Hugging Face Blog · Jun 12, 2026
Hugging Face introduces olmo-eval, an evaluation workbench for the model development loop.
Profiling in PyTorch (Part 2): From nn.Linear to a Fused MLP
Hugging Face · Hugging Face Blog · Jun 11, 2026
Hugging Face tutorial demonstrates profiling and fusing a PyTorch nn.Linear into an optimized fused MLP.
Can Voice Agents Handle Bilingual Customers? Benchmarking Frontier ASR on Code-Switched Speech
Hugging Face · Hugging Face Blog · Jun 09, 2026
Hugging Face benchmarks frontier ASR systems on code-switched bilingual speech for voice agents.
Introducing North Mini Code: Cohere’s First Model For Developers
Hugging Face · Hugging Face Blog · Jun 09, 2026
Cohere launched North Mini Code, its first model built specifically for developers.
How an Agent Built a 3D Paris Gallery by Chaining Two Hugging Face Spaces
Hugging Face · Hugging Face Blog · Jun 09, 2026
An AI agent autonomously built a 3D Paris gallery by chaining two Hugging Face Spaces together.
The Open Source Community is backing OpenEnv for Agentic RL
Hugging Face · Hugging Face Blog · Jun 08, 2026
The open source community, led by Hugging Face, is rallying behind OpenEnv as a standard for agentic reinforcement learning environments.
Amazing Digital Dentures (a failed project)
Hugging Face · Hugging Face Blog · Jun 07, 2026
Hugging Face shares a post-mortem on Amazing Digital Dentures, a machine learning project that failed.
Five labs, five minds: building a multi-model finance drama on small models
Hugging Face · Hugging Face Blog · Jun 06, 2026
Hugging Face built a multi-model finance drama simulation running five distinct AI personas on small models.
Thousand Token Wood: shipping a multi-agent economy on a 3B model
Hugging Face · Hugging Face Blog · Jun 05, 2026
Hugging Face shipped a working multi-agent economy running on a single 3B-parameter model.
Nemotron 3.5 Content Safety: Customizable Multimodal Safety for Global Enterprise AI
Hugging Face · Hugging Face Blog · Jun 04, 2026
Nemotron 3.5 Content Safety launches as a customizable multimodal safety model for global enterprise AI.
EVA-Bench Data 2.0: 3 Domains, 121 Tools, 213 Scenarios
Hugging Face · Hugging Face Blog · Jun 04, 2026
Hugging Face expands EVA-Bench to 3 domains, 121 tools, and 213 scenarios for agent evaluation.
Direct Preference Optimization Beyond Chatbots
Hugging Face · Hugging Face Blog · Jun 03, 2026
No article content was provided; only a title about applying Direct Preference Optimization beyond chatbots.
Adding MCP Tools to Reachy Mini
Hugging Face · Hugging Face Blog · Jun 03, 2026
Hugging Face shows how to add MCP tools to its open-source Reachy Mini robot.
Holo3.1: Fast & Local Computer Use Agents
Hugging Face · Hugging Face Blog · Jun 02, 2026
Hugging Face released Holo3.1, computer-use agents that run fast and locally.
Introducing Mellum2: A 12B Mixture-of-Experts Model by JetBrains
Hugging Face · Hugging Face Blog · Jun 01, 2026
JetBrains released Mellum2, a 12B mixture-of-experts model, on Hugging Face.
Beyond LLMs: Why Scalable Enterprise AI Adoption Depends on Agent Logic
Hugging Face · Hugging Face Blog · Jun 01, 2026
Hugging Face argues scalable enterprise AI adoption depends on agent logic, not just LLMs.
Welcome NVIDIA Cosmos 3: The First Open Omni-model for Physical AI Reasoning and Action
Hugging Face · Hugging Face Blog · Jun 01, 2026
NVIDIA releases Cosmos 3, billed as the first open omni-model for physical AI reasoning and action.
Profiling in PyTorch (Part 1): A Beginner's Guide to torch.profiler
Hugging Face · Hugging Face Blog · May 29, 2026
Hugging Face publishes beginner guide to PyTorch profiling with torch.profiler
ITBench-AA: Frontier Models Score Below 50% on the First Benchmark for Agentic Enterprise IT Tasks — by Artificial Analysis and IBM
Hugging Face · Hugging Face Blog · May 27, 2026
Frontier AI models score below 50% on new agentic enterprise IT benchmark ITBench-AA
Reachy Mini goes fully local
Hugging Face · Hugging Face Blog · May 27, 2026
Hugging Face's Reachy Mini robot now runs AI inference fully on-device locally
Shipping a Trillion Parameters With a Hub Bucket: Delta Weight Sync in TRL
Hugging Face · Hugging Face Blog · May 27, 2026
Hugging Face TRL ships delta weight sync for trillion-parameter model training efficiency.
Harness, Scaffold, and the AI Agent Terms Worth Getting Right
Hugging Face · Hugging Face Blog · May 25, 2026
Hugging Face publishes a terminology guide clarifying AI agent concepts like harness and scaffold.
Towards Speed-of-Light Text Generation with Nemotron-Labs Diffusion Language Models
Hugging Face · Hugging Face Blog · May 23, 2026
Nemotron-Labs diffusion language models promise near-instant text generation speeds
Specialization Beats Scale: A Strategic Variable Most AI Procurement Decisions Overlook
Hugging Face · Hugging Face Blog · May 22, 2026
Specialized AI models outperform large-scale general models for enterprise procurement decisions.