6301

NVIDIA Nemotron Post-Training Dataset v2 and Nemotron Nano 2 9B Release

NVIDIA has released a 6-million sample multilingual reasoning dataset and the Nemotron Nano 2 9B model, which utilizes a hybrid Transformer-Mamba architecture to optimize reasoning costs and throughput.

6302

MIXI ChatGPT Enterprise Deployment

MIXI deployed ChatGPT Enterprise across its organization in 45 days, reducing work hours by over 90% in some projects and enabling employees to create over 1,600 custom GPTs.

6303

Generate Images with Claude and Hugging Face

Hugging Face enables image generation within Claude by connecting the AI to Hugging Face Spaces via the Model Context Protocol (MCP) server.

6304

Qwen-Image-Edit Release Notes

Qwen-Image-Edit is a 20B parameter image editing model that enables precise semantic and appearance editing, including bilingual text modification, by leveraging Qwen2.5-VL and a VAE Encoder.

6305

Hugging Face MCP for Research: Connecting AI to Research Tools

Hugging Face introduces the Research Tracker MCP, enabling AI agents to automate research discovery by integrating arXiv, GitHub, and Hugging Face through the Model Context Protocol.

6306

Hugging Face kernel-builder: A Guide to Building and Scaling Production-Ready CUDA Kernels

Hugging Face introduces the kernel-builder library to simplify the development, multi-architecture compilation, and distribution of production-ready CUDA kernels via the Hugging Face Hub.

6307

DoorDash AI Implementation Strategy

DoorDash is leveraging AI to democratize technical creation for non-engineers and personalize employee development and performance management.

6308

Kimina-Prover-RL Release

Hugging Face has released Kimina-Prover-RL, an open-source RL training pipeline and two SOTA models (0.6B and 1.7B) for formal theorem proving in Lean 4.

6309

Arm and ExecuTorch 0.7: Expanding Generative AI to Billions of Devices

Arm and the ExecuTorch 0.7 beta enable automatic AI acceleration via KleidiAI, leveraging the SDOT instruction to bring LLMs like Llama 3.2 to billions of existing Arm-based devices.

6310

Arm Neural Super Sampling (NSS) Release

Arm has released Neural Super Sampling (NSS), an AI-powered upscaling solution designed to reduce GPU workloads and enable high-resolution rendering on mobile devices.

6311

FilBench: Evaluating LLM Capabilities in Philippine Languages

Hugging Face has introduced FilBench, a comprehensive evaluation suite designed to assess the fluency, linguistic abilities, and cultural knowledge of LLMs in Tagalog, Filipino, and Cebuano.

6312

TextQuests: Evaluating LLM Agentic Reasoning in Text-Based Video Games

Hugging Face introduces TextQuests, a benchmark using 25 classic Infocom interactive fiction games to evaluate the long-context reasoning and exploratory capabilities of LLM agents.

6313

Basis Scales Accounting Automation with OpenAI GPT-5 and o3

Basis uses a multi-agent architecture powered by OpenAI's GPT-5, GPT-4.1, and o3 models to automate complex accounting workflows, achieving up to 30% time savings for accounting firms.

6314

OpenAI Proposal for Harmonized AI Regulation in California

OpenAI has urged Governor Gavin Newsom to align California's AI regulations with national and global standards to prevent a patchwork of state rules that could hinder innovation and US competitiveness.

6315

Hugging Face AI Sheets Release

Hugging Face has released AI Sheets, an open-source no-code tool for building, transforming, and enriching datasets using open AI models.

6316

Accelerate ND-Parallel: Efficient Multi-GPU Training Guide

Hugging Face has integrated ND-Parallelism into Accelerate and Axolotl, allowing users to combine Data, Fully Sharded Data, Tensor, and Context parallelism strategies to optimize multi-GPU training for massive models.

6317

OpenAI GPT-5 Release Notes

OpenAI has released GPT-5, a unified model that integrates reasoning, agents, and advanced math capabilities to improve accuracy, speed, and problem-solving for business operations.

6318

GPT-5 for Developers Release Notes

OpenAI has released GPT-5 in three sizes (gpt-5, gpt-5-mini, and gpt-5-nano), delivering state-of-the-art performance in coding, agentic tool-calling, and long-context retrieval.

6319

GPT-5 Coding and Design Capabilities

OpenAI has announced GPT-5 with a focus on enhanced capabilities for coding and design, though specific technical details were not provided in the announcement.

6320

OpenAI GPT-5 Creative Writing Capabilities

OpenAI has announced new creative writing capabilities for GPT-5, expanding the model's ability to generate high-quality narrative and artistic text.

6321

Medical Research with GPT-5

OpenAI has announced the application of GPT-5 to medical research, expanding the model's utility in specialized scientific domains.

6322

Vision Language Model Alignment in TRL

Hugging Face has expanded the TRL library to support advanced alignment methods for Vision Language Models, including MPO, GRPO, and GSPO, alongside native SFT support and vLLM integration.

6323

GPT-5 Integration in Cursor

OpenAI has announced the integration of GPT-5 into the Cursor code editor, enhancing AI-powered software development capabilities.

6324

OpenAI GPT-5 First Look

OpenAI has provided a first look at GPT-5, marking a new generation of their large language model series.

6325

How Amgen uses GPT-5

Amgen leverages GPT-5 via the OpenAI API to accelerate biotechnology research and drug discovery processes.

6326

OpenAI GPT-5 System Card

OpenAI has released GPT-5, a unified system featuring a real-time router that dynamically switches between fast, high-throughput models and deeper reasoning models based on query complexity.

6327

GPT-5 Safe Completions: Transitioning from Refusal-Based to Output-Centric Safety Training

OpenAI has introduced safe-completions for GPT-5, a safety-training approach that maximizes helpfulness while penalizing unsafe outputs, reducing the binary comply-or-refuse trade-off for dual-use prompts.

6328

OpenAI GPT-5 Release Notes

OpenAI has released GPT-5, a unified AI system featuring a real-time router that switches between a fast model and a deeper reasoning model to provide expert-level intelligence across coding, math, and health.

6329

OpenAI and GSA Partnership: ChatGPT Enterprise for U.S. Federal Workforce

OpenAI is providing ChatGPT Enterprise to the entire U.S. federal executive branch workforce for a nominal fee of $1 per agency for one year to reduce administrative burden and improve public service delivery.

6330

OpenAI Estimating Worst Case Frontier Risks of Open Weight LLMs

OpenAI evaluated the risks of releasing gpt-oss by using malicious fine-tuning (MFT) to attempt to elicit maximum capabilities in biology and cybersecurity, finding that the model did not substantially advance the frontier of risk compared to existing open-weight models.

6331

OpenAI gpt-oss-120b and gpt-oss-20b Release

OpenAI has released gpt-oss-120b and gpt-oss-20b, two open-weight reasoning models under the Apache 2.0 license designed for agentic workflows and customizable reasoning effort.

6332

OpenAI gpt-oss Release Notes

OpenAI has released gpt-oss-120b and gpt-oss-20b, two open-weight reasoning models licensed under Apache 2.0 that deliver frontier-level reasoning and tool-use capabilities optimized for consumer hardware.

6333

OpenAI Open Weights Release

OpenAI has released its most capable open-weights reasoning models to democratize AI access and promote US-led democratic AI standards globally.

6334

NVIDIA AI-Q Blueprint: Top-Ranking Open Deep Research Agent on DeepResearch Bench

NVIDIA's AI-Q Blueprint achieves the top spot for open-licensed stacks on the Hugging Face DeepResearch Bench, utilizing a combination of Llama 3.3-70B Instruct and Llama-3.3-Nemotron-Super-49B-v1.5.

6335

Qwen-Image Release: Native Text Rendering and Precise Image Editing

Qwen-Image is a 20B MMDiT image foundation model that provides state-of-the-art complex text rendering and precise image editing capabilities.

6336

OpenAI Optimizing ChatGPT for User Utility and Wellbeing

OpenAI is shifting ChatGPT's optimization goals from engagement metrics like time spent to real-world utility and user wellbeing, introducing break reminders and improved handling of high-stakes personal decisions.

6337

3LM: A Benchmark for Arabic LLMs in STEM and Code

Hugging Face and TII UAE introduce 3LM, the first comprehensive benchmark designed to evaluate Arabic Large Language Models on STEM subjects and code generation.

6338

Figma AI Integration and Product Evolution

Figma is integrating AI across its platform to automate routine design tasks and introduce prompt-to-app capabilities via Figma Make, shifting the designer's role from implementation to high-level problem solving.

6339

Implementing MCP Servers in Python with Gradio

Hugging Face introduces a method for Python developers to quickly build Model Context Protocol (MCP) servers using Gradio, enabling LLMs to integrate with thousands of AI models and Spaces on the Hugging Face Hub.

6340

OpenAI Introduces Stargate Norway AI Data Center

OpenAI has launched Stargate Norway, a renewable-powered AI data center initiative in Narvik, Norway, aiming to deliver 100,000 NVIDIA GPUs by the end of 2026.

6341

Intercom's Strategy for Sustainable AI Advantage

Intercom achieved a sustainable AI advantage by prioritizing early model fluency, rigorous evaluation frameworks, and a modular architecture that allows for rapid model swapping and cost optimization.

6342

ChatGPT Study Mode Release

OpenAI has introduced study mode in ChatGPT, a learning experience designed to guide students through problems using Socratic questioning and scaffolded responses rather than providing direct answers.

6343

Hugging Face Trackio Release

Hugging Face has released Trackio, a free, open-source, lightweight experiment tracking library that serves as a drop-in replacement for wandb with native integration for Hugging Face Spaces and Datasets.

6344

Qwen GSPO: Scalable Reinforcement Learning for Language Models

Qwen introduces Group Sequence Policy Optimization (GSPO), a sequence-level RL algorithm that improves training stability and efficiency over GRPO, particularly for Mixture-of-Experts (MoE) models.

6345

Hugging Face CLI Update: Transition to hf Command

Hugging Face has renamed the huggingface-cli to hf, introducing a more ergonomic resource-action command structure and a new hf jobs service for running scripts on HF infrastructure.

6346

Parquet Content-Defined Chunking for Efficient Data Deduplication

Hugging Face introduces Parquet Content-Defined Chunking (CDC), now available in PyArrow and Pandas, to enable efficient deduplication of Parquet files on the Xet storage layer, significantly reducing upload and download times.

6347

Qwen-MT Turbo Release Notes

Qwen has released Qwen-MT (qwen-mt-turbo), a lightweight MoE-based translation model supporting 92 languages with high customizability and low API costs.

6348

Outtake Cybersecurity Agents powered by OpenAI

Outtake uses GPT-4.1 and OpenAI o3 to automate the detection and resolution of digital threats, reducing takedown timelines from 60 days to a few hours.

6349

Fast LoRA Inference for Flux with Diffusers and PEFT

Hugging Face introduces an optimization recipe for Flux.1-Dev that achieves up to 2.23x speedup in LoRA inference by combining hotswapping, torch.compile, and FP8 quantization.

6350

TimeScope: A New Benchmark for Long-Video Large Multimodal Model Understanding

Hugging Face introduces TimeScope, an open-source benchmark that evaluates the temporal comprehension of vision-language models using video needles inserted into content ranging from 1 minute to 8 hours.