151

huggingface_hub v1.0: Five Years of Building the Foundation of Open Machine Learning

huggingface_hub has reached v1.0 after five years, providing core access to over 2 million models, 500k datasets, and 1M Spaces while introducing breaking changes like httpx, hf_xet, and a redesigned CLI to support the next decade of open machine learning.

152

Hugging Face Streaming Datasets Update

Hugging Face has optimized the streaming capabilities of the datasets library, reducing startup requests by 100x and increasing data resolution speed by 10x to enable efficient training on multi-terabyte datasets without local downloads.

153

LeRobot v0.4.0 release notes / what's new

Hugging Face has released LeRobot v0.4.0, introducing scalable Datasets v3.0, VLA models like PI0.5 and GR00T N1.5, and a plugin system for streamlined hardware integration.

154

Hugging Face and Meta introduce OpenEnv for agentic environments

Hugging Face and Meta have launched OpenEnv and the OpenEnv Hub, a standardized, open community hub for secure, sandboxed agentic environments used in training and deployment.

155

Hugging Face and VirusTotal Security Collaboration

Hugging Face has partnered with VirusTotal to continuously scan over 2.2 million public model and dataset repositories on the Hugging Face Hub to protect the ML community from malicious assets.

156

Sentence Transformers joins Hugging Face

The Sentence Transformers library is transitioning from the UKP Lab at TU Darmstadt to Hugging Face to leverage better infrastructure for continued open-source development.

157

Hugging Face AI Sheets Vision Update

Hugging Face AI Sheets is an open-source tool that now supports vision capabilities, allowing users to extract structured data from images, generate visuals from text, and edit images within a spreadsheet interface.

158

Open-Source OCR Models Guide – Choosing, Running, and Extending Modern Vision-Language OCR

Hugging Face announced a comprehensive guide to open-weight OCR models, detailing their capabilities, benchmark performance, cost efficiency, and how to run them locally or via hosted endpoints.

159

AI for Food Allergies: Open Dataset Collection for Biomedical Research

Hugging Face has launched the AI for Food Allergies project, introducing the first open, community-driven collection of curated datasets to accelerate AI-driven allergen prediction, drug discovery, and food safety.

160

Google Cloud C4 and Intel Xeon 6 Performance for GPT OSS

Google Cloud C4 VMs powered by Intel Xeon 6 processors deliver up to a 1.7x improvement in Total Cost of Ownership (TCO) for GPT OSS MoE model inference compared to C3 VMs.

161

Run SmolVLM on Intel CPUs with OpenVINO in 3 Steps

Hugging Face shows how to deploy the small Vision‑Language Model SmolVLM on Intel CPUs using Optimum‑Intel and OpenVINO, achieving up to 65× higher throughput than PyTorch.

162

Nemotron-Personas-India: Synthesized Data for Sovereign AI

NVIDIA has released Nemotron-Personas-India, a CC BY 4.0 licensed synthetic dataset of 21 million Indic personas designed to bridge the data gap for Sovereign AI in India's multilingual and culturally diverse environment.

163

Arm at the PyTorch Conference 2025

Arm is participating in the PyTorch Conference on October 22-23 to showcase AI deployment tools including ExecuTorch and vLLM, and to gather developer feedback on scaling AI across cloud, edge, and mobile platforms.

164

BigCodeArena: Judging code generations end to end with code executions

Hugging Face has launched BigCodeArena, a human-in-the-loop platform that evaluates code generation models by allowing users to execute and interact with generated code in real-time.

165

SOTA OCR with Core ML and dots.ocr

Hugging Face demonstrates the conversion of the 3B parameter dots.ocr model to run on-device using Core ML and MLX, achieving performance that surpasses Gemini 2.5 Pro on OmniDocBench.

166

Hugging Face Introduces RTEB for Retrieval Embedding Evaluation

Hugging Face has launched the beta version of the Retrieval Embedding Benchmark (RTEB), a new standard designed to measure the true generalization and retrieval accuracy of embedding models using a hybrid of open and private datasets.

167

VibeGame: A High-Level Declarative Engine for AI-Assisted Game Development

Hugging Face researcher Dylan Ebert introduces VibeGame, a declarative game engine designed to enable 'vibe coding' by combining high-level abstractions with a web-based stack optimized for AI proficiency.

168

Qwen3-8B acceleration on Intel Core Ultra with depth‑pruned draft models

Accelerated Qwen3-8B on Intel Core Ultra using speculative decoding and a depth‑pruned Qwen3‑0.6B draft achieves ~1.4× speedup, enabling fast local AI agents via Hugging Face smolagents.

169

Nemotron-Personas-Japan Release

NVIDIA has released Nemotron-Personas-Japan, an open synthetic dataset of 6 million Japanese personas designed to support the development of Sovereign AI and culturally aware LLMs.

170

Swift Transformers 1.0 release notes / what's new

Hugging Face has released Swift Transformers 1.0, a stable library providing tokenizers, hub access, and model wrappers to simplify local LLM integration for Apple Silicon developers.

171

Smol2Operator release: turning a small VLM into an open‑source agentic GUI coder

Hugging Face released Smol2Operator, a post‑training recipe that converts the 2.2 B‑parameter SmolVLM2‑2.2B‑Instruct model into an open‑source agentic GUI coder, with all code, datasets, and the resulting model publicly available.

172

SyGra: A Low-Code Framework for LLM and SLM Data Generation

SyGra is a low-code/no-code framework developed by ServiceNow AI to simplify the creation, transformation, and alignment of datasets for Large and Small Language Models.

173

Hugging Face Gaia2 and Meta Agents Research Environments (ARE) Release

Hugging Face has released Gaia2, a read-and-write agentic benchmark, and the Meta Agents Research Environments (ARE) framework to evaluate and debug AI agents in complex, real-world simulated conditions.

174

Scaleway added as Hugging Face Inference Provider – capabilities, usage, and billing

Scaleway is now an official Inference Provider on the Hugging Face Hub, offering low‑latency, European‑hosted serverless access to frontier models with flexible billing options.

175

Hugging Face RiskRubric.ai Announcement

Hugging Face has announced RiskRubric.ai, a standardized risk assessment platform that evaluates AI models across six pillars of transparency, reliability, security, privacy, safety, and reputation to provide comparable risk scores and letter grades.

176

Hugging Face Integrates Public AI as an Inference Provider

Hugging Face has added Public AI as a supported Inference Provider, enabling serverless access to sovereign models from institutions like the Swiss AI Initiative and AI Singapore.

177

LeRobotDataset v3.0 release notes / what's new

Hugging Face has released LeRobotDataset v3.0, a standardized robotics dataset format that enables large-scale data storage and streaming to support millions of episodes.

178

Visible Watermarking with Gradio

Hugging Face has introduced a simple way to add visible watermarks to AI-generated images, videos, and text using the Gradio library to improve synthetic content transparency.

179

Writer Palmyra-mini Family Release

Writer has released the Palmyra-mini family, a set of lightweight open models (1.5B to 1.7B parameters) featuring a base model and two specialized reasoning variants optimized for logic and mathematics.

180

Transformers 4.40 performance upgrades for OpenAI GPT‑OSS: zero‑build kernels, MXFP4 quantization, parallelism, and faster loading

Hugging Face added zero-build kernels, MXFP4 4-bit quantization, tensor and expert parallelism, dynamic sliding-window cache, continuous batching, and faster model loading to the transformers library for OpenAI's GPT‑OSS models, dramatically improving efficiency and scalability.

181

Fine-tuning Hugging Face Hub LLMs with Together AI

Together AI and Hugging Face have integrated their platforms, allowing developers to fine-tune any compatible LLM from the Hugging Face Hub using Together AI's infrastructure.

182

Jupyter Agents: Training LLMs to Reason with Notebooks

Hugging Face introduces a pipeline for generating high-quality synthetic data and optimized scaffolding to transform small LLMs like Qwen3-4B into state-of-the-art data science agents.

183

mmBERT: ModernBERT goes Multilingual

Hugging Face introduces mmBERT, a massively multilingual encoder model trained on 3T+ tokens across 1,800+ languages that outperforms XLM-R in both performance and inference speed.

184

EmbeddingGemma 300M release notes

Google released EmbeddingGemma, a 308M‑parameter multilingual embedding model that tops the MTEB benchmark for sub‑500M models and is optimized for on‑device use.

185

SandboxAQ SAIR Dataset Release

SandboxAQ has released SAIR, the largest dataset of co-folded 3D protein-ligand structures paired with experimental IC50 labels, providing over 5 million AI-generated structures to accelerate AI-powered drug discovery.

186

Hugging Face ZeroGPU Ahead-of-Time Compilation Guide

Hugging Face introduces ahead-of-time (AoT) compilation for ZeroGPU Spaces, enabling speedups of 1.3x to 1.8x for generative models like Flux, Wan, and LTX by eliminating just-in-time compilation overhead.

187

NVIDIA Nemotron Post-Training Dataset v2 and Nemotron Nano 2 9B Release

NVIDIA has released a 6-million sample multilingual reasoning dataset and the Nemotron Nano 2 9B model, which utilizes a hybrid Transformer-Mamba architecture to optimize reasoning costs and throughput.

188

Generate Images with Claude and Hugging Face

Hugging Face enables image generation within Claude by connecting the AI to Hugging Face Spaces via the Model Context Protocol (MCP) server.

189

Hugging Face MCP for Research: Connecting AI to Research Tools

Hugging Face introduces the Research Tracker MCP, enabling AI agents to automate research discovery by integrating arXiv, GitHub, and Hugging Face through the Model Context Protocol.

190

Hugging Face kernel-builder: A Guide to Building and Scaling Production-Ready CUDA Kernels

Hugging Face introduces the kernel-builder library to simplify the development, multi-architecture compilation, and distribution of production-ready CUDA kernels via the Hugging Face Hub.

191

Kimina-Prover-RL Release

Hugging Face has released Kimina-Prover-RL, an open-source RL training pipeline and two SOTA models (0.6B and 1.7B) for formal theorem proving in Lean 4.

192

Arm and ExecuTorch 0.7: Expanding Generative AI to Billions of Devices

Arm and the ExecuTorch 0.7 beta enable automatic AI acceleration via KleidiAI, leveraging the SDOT instruction to bring LLMs like Llama 3.2 to billions of existing Arm-based devices.

193

Arm Neural Super Sampling (NSS) Release

Arm has released Neural Super Sampling (NSS), an AI-powered upscaling solution designed to reduce GPU workloads and enable high-resolution rendering on mobile devices.

194

FilBench: Evaluating LLM Capabilities in Philippine Languages

Hugging Face has introduced FilBench, a comprehensive evaluation suite designed to assess the fluency, linguistic abilities, and cultural knowledge of LLMs in Tagalog, Filipino, and Cebuano.

195

TextQuests: Evaluating LLM Agentic Reasoning in Text-Based Video Games

Hugging Face introduces TextQuests, a benchmark using 25 classic Infocom interactive fiction games to evaluate the long-context reasoning and exploratory capabilities of LLM agents.

196

Hugging Face AI Sheets Release

Hugging Face has released AI Sheets, an open-source no-code tool for building, transforming, and enriching datasets using open AI models.

197

Accelerate ND-Parallel: Efficient Multi-GPU Training Guide

Hugging Face has integrated ND-Parallelism into Accelerate and Axolotl, allowing users to combine Data, Fully Sharded Data, Tensor, and Context parallelism strategies to optimize multi-GPU training for massive models.

198

Vision Language Model Alignment in TRL

Hugging Face has expanded the TRL library to support advanced alignment methods for Vision Language Models, including MPO, GRPO, and GSPO, alongside native SFT support and vLLM integration.

199

NVIDIA AI-Q Blueprint: Top-Ranking Open Deep Research Agent on DeepResearch Bench

NVIDIA's AI-Q Blueprint achieves the top spot for open-licensed stacks on the Hugging Face DeepResearch Bench, utilizing a combination of Llama 3.3-70B Instruct and Llama-3.3-Nemotron-Super-49B-v1.5.

200

3LM: A Benchmark for Arabic LLMs in STEM and Code

Hugging Face and TII UAE introduce 3LM, the first comprehensive benchmark designed to evaluate Arabic Large Language Models on STEM subjects and code generation.