NVIDIA Nemotron Post-Training Dataset v2 and Nemotron Nano 2 9B Release
NVIDIA has released a 6-million sample multilingual reasoning dataset and the Nemotron Nano 2 9B model, which utilizes a hybrid Transformer-Mamba architecture to optimize reasoning costs and throughput.
MIXI ChatGPT Enterprise Deployment
MIXI deployed ChatGPT Enterprise across its organization in 45 days, reducing work hours by over 90% in some projects and enabling employees to create over 1,600 custom GPTs.
Generate Images with Claude and Hugging Face
Hugging Face enables image generation within Claude by connecting the AI to Hugging Face Spaces via the Model Context Protocol (MCP) server.
Qwen-Image-Edit Release Notes
Qwen-Image-Edit is a 20B parameter image editing model that enables precise semantic and appearance editing, including bilingual text modification, by leveraging Qwen2.5-VL and a VAE Encoder.
Hugging Face MCP for Research: Connecting AI to Research Tools
Hugging Face introduces the Research Tracker MCP, enabling AI agents to automate research discovery by integrating arXiv, GitHub, and Hugging Face through the Model Context Protocol.
Hugging Face kernel-builder: A Guide to Building and Scaling Production-Ready CUDA Kernels
Hugging Face introduces the kernel-builder library to simplify the development, multi-architecture compilation, and distribution of production-ready CUDA kernels via the Hugging Face Hub.
DoorDash AI Implementation Strategy
DoorDash is leveraging AI to democratize technical creation for non-engineers and personalize employee development and performance management.
Kimina-Prover-RL Release
Hugging Face has released Kimina-Prover-RL, an open-source RL training pipeline and two SOTA models (0.6B and 1.7B) for formal theorem proving in Lean 4.
Arm and ExecuTorch 0.7: Expanding Generative AI to Billions of Devices
Arm and the ExecuTorch 0.7 beta enable automatic AI acceleration via KleidiAI, leveraging the SDOT instruction to bring LLMs like Llama 3.2 to billions of existing Arm-based devices.
Arm Neural Super Sampling (NSS) Release
Arm has released Neural Super Sampling (NSS), an AI-powered upscaling solution designed to reduce GPU workloads and enable high-resolution rendering on mobile devices.
FilBench: Evaluating LLM Capabilities in Philippine Languages
Hugging Face has introduced FilBench, a comprehensive evaluation suite designed to assess the fluency, linguistic abilities, and cultural knowledge of LLMs in Tagalog, Filipino, and Cebuano.
TextQuests: Evaluating LLM Agentic Reasoning in Text-Based Video Games
Hugging Face introduces TextQuests, a benchmark using 25 classic Infocom interactive fiction games to evaluate the long-context reasoning and exploratory capabilities of LLM agents.
Basis Scales Accounting Automation with OpenAI GPT-5 and o3
Basis uses a multi-agent architecture powered by OpenAI's GPT-5, GPT-4.1, and o3 models to automate complex accounting workflows, achieving up to 30% time savings for accounting firms.
OpenAI Proposal for Harmonized AI Regulation in California
OpenAI has urged Governor Gavin Newsom to align California's AI regulations with national and global standards to prevent a patchwork of state rules that could hinder innovation and US competitiveness.
Hugging Face AI Sheets Release
Hugging Face has released AI Sheets, an open-source no-code tool for building, transforming, and enriching datasets using open AI models.
Accelerate ND-Parallel: Efficient Multi-GPU Training Guide
Hugging Face has integrated ND-Parallelism into Accelerate and Axolotl, allowing users to combine Data, Fully Sharded Data, Tensor, and Context parallelism strategies to optimize multi-GPU training for massive models.
OpenAI GPT-5 Release Notes
OpenAI has released GPT-5, a unified model that integrates reasoning, agents, and advanced math capabilities to improve accuracy, speed, and problem-solving for business operations.
GPT-5 for Developers Release Notes
OpenAI has released GPT-5 in three sizes (gpt-5, gpt-5-mini, and gpt-5-nano), delivering state-of-the-art performance in coding, agentic tool-calling, and long-context retrieval.
GPT-5 Coding and Design Capabilities
OpenAI has announced GPT-5 with a focus on enhanced capabilities for coding and design, though specific technical details were not provided in the announcement.
OpenAI GPT-5 Creative Writing Capabilities
OpenAI has announced new creative writing capabilities for GPT-5, expanding the model's ability to generate high-quality narrative and artistic text.
Medical Research with GPT-5
OpenAI has announced the application of GPT-5 to medical research, expanding the model's utility in specialized scientific domains.
Vision Language Model Alignment in TRL
Hugging Face has expanded the TRL library to support advanced alignment methods for Vision Language Models, including MPO, GRPO, and GSPO, alongside native SFT support and vLLM integration.
GPT-5 Integration in Cursor
OpenAI has announced the integration of GPT-5 into the Cursor code editor, enhancing AI-powered software development capabilities.
OpenAI GPT-5 First Look
OpenAI has provided a first look at GPT-5, marking a new generation of their large language model series.
How Amgen uses GPT-5
Amgen leverages GPT-5 via the OpenAI API to accelerate biotechnology research and drug discovery processes.
OpenAI GPT-5 System Card
OpenAI has released GPT-5, a unified system featuring a real-time router that dynamically switches between fast, high-throughput models and deeper reasoning models based on query complexity.
GPT-5 Safe Completions: Transitioning from Refusal-Based to Output-Centric Safety Training
OpenAI has introduced safe-completions for GPT-5, a safety-training approach that maximizes helpfulness while penalizing unsafe outputs, reducing the binary comply-or-refuse trade-off for dual-use prompts.
OpenAI GPT-5 Release Notes
OpenAI has released GPT-5, a unified AI system featuring a real-time router that switches between a fast model and a deeper reasoning model to provide expert-level intelligence across coding, math, and health.
OpenAI and GSA Partnership: ChatGPT Enterprise for U.S. Federal Workforce
OpenAI is providing ChatGPT Enterprise to the entire U.S. federal executive branch workforce for a nominal fee of $1 per agency for one year to reduce administrative burden and improve public service delivery.
OpenAI Estimating Worst Case Frontier Risks of Open Weight LLMs
OpenAI evaluated the risks of releasing gpt-oss by using malicious fine-tuning (MFT) to attempt to elicit maximum capabilities in biology and cybersecurity, finding that the model did not substantially advance the frontier of risk compared to existing open-weight models.
OpenAI gpt-oss-120b and gpt-oss-20b Release
OpenAI has released gpt-oss-120b and gpt-oss-20b, two open-weight reasoning models under the Apache 2.0 license designed for agentic workflows and customizable reasoning effort.
OpenAI gpt-oss Release Notes
OpenAI has released gpt-oss-120b and gpt-oss-20b, two open-weight reasoning models licensed under Apache 2.0 that deliver frontier-level reasoning and tool-use capabilities optimized for consumer hardware.
OpenAI Open Weights Release
OpenAI has released its most capable open-weights reasoning models to democratize AI access and promote US-led democratic AI standards globally.
NVIDIA AI-Q Blueprint: Top-Ranking Open Deep Research Agent on DeepResearch Bench
NVIDIA's AI-Q Blueprint achieves the top spot for open-licensed stacks on the Hugging Face DeepResearch Bench, utilizing a combination of Llama 3.3-70B Instruct and Llama-3.3-Nemotron-Super-49B-v1.5.
Qwen-Image Release: Native Text Rendering and Precise Image Editing
Qwen-Image is a 20B MMDiT image foundation model that provides state-of-the-art complex text rendering and precise image editing capabilities.
OpenAI Optimizing ChatGPT for User Utility and Wellbeing
OpenAI is shifting ChatGPT's optimization goals from engagement metrics like time spent to real-world utility and user wellbeing, introducing break reminders and improved handling of high-stakes personal decisions.
3LM: A Benchmark for Arabic LLMs in STEM and Code
Hugging Face and TII UAE introduce 3LM, the first comprehensive benchmark designed to evaluate Arabic Large Language Models on STEM subjects and code generation.
Figma AI Integration and Product Evolution
Figma is integrating AI across its platform to automate routine design tasks and introduce prompt-to-app capabilities via Figma Make, shifting the designer's role from implementation to high-level problem solving.
Implementing MCP Servers in Python with Gradio
Hugging Face introduces a method for Python developers to quickly build Model Context Protocol (MCP) servers using Gradio, enabling LLMs to integrate with thousands of AI models and Spaces on the Hugging Face Hub.
OpenAI Introduces Stargate Norway AI Data Center
OpenAI has launched Stargate Norway, a renewable-powered AI data center initiative in Narvik, Norway, aiming to deliver 100,000 NVIDIA GPUs by the end of 2026.
Intercom's Strategy for Sustainable AI Advantage
Intercom achieved a sustainable AI advantage by prioritizing early model fluency, rigorous evaluation frameworks, and a modular architecture that allows for rapid model swapping and cost optimization.
ChatGPT Study Mode Release
OpenAI has introduced study mode in ChatGPT, a learning experience designed to guide students through problems using Socratic questioning and scaffolded responses rather than providing direct answers.
Hugging Face Trackio Release
Hugging Face has released Trackio, a free, open-source, lightweight experiment tracking library that serves as a drop-in replacement for wandb with native integration for Hugging Face Spaces and Datasets.
Qwen GSPO: Scalable Reinforcement Learning for Language Models
Qwen introduces Group Sequence Policy Optimization (GSPO), a sequence-level RL algorithm that improves training stability and efficiency over GRPO, particularly for Mixture-of-Experts (MoE) models.
Hugging Face CLI Update: Transition to hf Command
Hugging Face has renamed the huggingface-cli to hf, introducing a more ergonomic resource-action command structure and a new hf jobs service for running scripts on HF infrastructure.
Parquet Content-Defined Chunking for Efficient Data Deduplication
Hugging Face introduces Parquet Content-Defined Chunking (CDC), now available in PyArrow and Pandas, to enable efficient deduplication of Parquet files on the Xet storage layer, significantly reducing upload and download times.
Qwen-MT Turbo Release Notes
Qwen has released Qwen-MT (qwen-mt-turbo), a lightweight MoE-based translation model supporting 92 languages with high customizability and low API costs.
Outtake Cybersecurity Agents powered by OpenAI
Outtake uses GPT-4.1 and OpenAI o3 to automate the detection and resolution of digital threats, reducing takedown timelines from 60 days to a few hours.
Fast LoRA Inference for Flux with Diffusers and PEFT
Hugging Face introduces an optimization recipe for Flux.1-Dev that achieves up to 2.23x speedup in LoRA inference by combining hotswapping, torch.compile, and FP8 quantization.
TimeScope: A New Benchmark for Long-Video Large Multimodal Model Understanding
Hugging Face introduces TimeScope, an open-source benchmark that evaluates the temporal comprehension of vision-language models using video needles inserted into content ranging from 1 minute to 8 hours.