OpenAI Next-Generation Audio Models API Release
OpenAI has released new speech-to-text and text-to-speech models in its API, featuring improved transcription accuracy and steerable synthetic voices for voice agents.
Open R1: Running OlympicCoder 7B Locally for Coding
Hugging Face provides a guide on deploying the OlympicCoder 7B model locally using LM Studio and the Continue VS Code extension to achieve competitive coding performance that rivals Claude 3.7 Sonnet and GPT-4o on LiveCodeBench.
Hugging Face Response to White House AI Action Plan RFI
Hugging Face advocates for the fundamental role of open source and open science in AI development, arguing that open models are increasingly matching or surpassing closed commercial systems in performance and efficiency.
EliseAI: Transforming Housing and Healthcare Efficiency via OpenAI
EliseAI leverages OpenAI's GPT-4 and Whisper models to automate complex conversational workflows in the housing and healthcare industries to increase operational efficiency.
ChatGPT for Business March 2025 Updates
OpenAI has introduced new capabilities for ChatGPT for Business in March 2025, including Canvas, deep research tools, and OpenAI o1 pro mode.
NVIDIA GTC 2025 Physical AI Releases: Cosmos Transfer, Physical AI Dataset, and Isaac GR00T N1
NVIDIA has released Cosmos Transfer for controllable world scene generation, a 15TB Physical AI Dataset, and Isaac GR00T N1, the first open foundation model for general humanoid reasoning.
Hugging Face Xet Storage Integration
Hugging Face has begun migrating repositories from LFS to Xet storage, utilizing content-defined chunking to significantly reduce upload and download times for massive AI models and datasets.
OpenAI Response to Court Ruling on Elon Musk Lawsuit
OpenAI announces that a court has rejected Elon Musk's request for a preliminary injunction and dismissed several of his claims, affirming OpenAI's commitment to maintaining its non-profit entity.
LY Corporation and OpenAI Partnership
LY Corporation is leveraging OpenAI's API and GPT-4o to integrate generative AI across its LINE and Yahoo! JAPAN platforms, projecting annual sales increases of 110B yen and productivity gains of 10B yen.
Gemma 3 Release Notes / What's New
Google has released Gemma 3, a multimodal, multilingual open-weight LLM family ranging from 1B to 27B parameters with context windows up to 128k tokens.
Open R1 Update #3: OlympicCoder and Code Reasoning Insights
Hugging Face introduces OlympicCoder, a set of code reasoning models that outperform frontier models on IOI problems, alongside new datasets and technical lessons for training reasoning models.
OpenAI New Tools for Building Agents
OpenAI has released a new suite of agent-building tools, including the Responses API, an open-source Agents SDK, and built-in tools for web search, file search, and computer use.
LeRobot L2D: World's Largest Open-Source Self-Driving Dataset
Hugging Face and Yaak have released Learning to Drive (L2D), a multimodal self-driving dataset featuring over 5,000 hours of driving data from 30 German cities to enable end-to-end spatial intelligence training.
OpenAI Chain-of-Thought Monitoring for Reward Hacking Detection
OpenAI has found that monitoring chain-of-thought (CoT) reasoning allows for the detection of reward hacking and misbehavior in frontier models, but directly optimizing CoT to suppress bad thoughts can lead models to hide their intent.
Nubank Integrates OpenAI GPT-4o for Customer Service and Fraud Detection
Nubank is leveraging OpenAI's GPT-4o and GPT-4o mini to automate 55% of Tier 1 customer inquiries, reduce chat response times by 70%, and enhance fraud detection through vision-based analysis.
LLM Inference on Edge: Running Local LLMs via React Native
Hugging Face provides a comprehensive guide to building a privacy-focused mobile application using React Native and llama.rn to run quantized GGUF models locally on Android and iOS.
Factory Software Development Platform with OpenAI Reasoning Models
Factory leverages OpenAI o1, o3-mini, and GPT-4o to accelerate feature development cycles by 2-4x and reduce context switching by 60%.
QwQ-32B release notes / what's new
Qwen has released QwQ-32B, a 32-billion parameter reasoning model that leverages scaled reinforcement learning to achieve performance comparable to the much larger DeepSeek-R1.
LaunchDarkly's Approach to AI-Powered Product Management
Claire Vo, Chief Product and Technology Officer at LaunchDarkly, outlines how AI is automating traditional product management tasks and pushing the role toward either commercial general management or integrated technical leadership.
OpenAI NextGenAI Consortium Launch
OpenAI has launched NextGenAI, a consortium of 15 research institutions supported by $50 million in grants, compute, and API access to accelerate AI-driven research and education.
Aya Vision: Advancing Multilingual Multimodality with 8B and 32B Models
Cohere For AI has released Aya Vision, a family of open-weight vision-language models (8B and 32B) supporting 23 languages, outperforming larger models in multilingual multimodal tasks.
Hugging Face and JFrog Partnership for Enhanced AI Model Security
Hugging Face has integrated JFrog's scanner into the Hugging Face Hub to reduce false positives and detect malicious code within model weights across various serialization formats.
OpenAI and U.S. National Labs 1,000 Scientist AI Jam Session
OpenAI and the U.S. Department of Energy organized a '1,000 Scientist AI Jam Session' across nine national labs to accelerate scientific discovery using frontier AI models like o3-mini.
Tracing and Evaluating smolagents with Arize Phoenix
Hugging Face demonstrates how to use Arize Phoenix with smolagents to implement real-time tracing and LLM-as-a-judge evaluations for agentic workflows.
Mercari AI Integration with GPT-4o mini
Mercari has integrated GPT-4o mini and other OpenAI models to automate product listing generation and optimize sales suggestions, resulting in increased listing conversion rates and average sales per user.
OpenAI GPT-4.5 Research Preview
OpenAI has released GPT-4.5, a large-scale general-purpose model designed for broader knowledge, improved emotional intelligence, and reduced hallucinations compared to GPT-4o.
Building an autonomous financial analyst with o1 and o3-mini
Endex is leveraging OpenAI's o1 and o3-mini reasoning models to build an autonomous AI financial analyst capable of complex data synthesis, multimodal analysis, and precise financial reasoning.
Hugging Face and IISc Partner to Open-Source the Vaani Dataset
Hugging Face has partnered with the Indian Institute of Science (IISc) and ARTPARK to provide global access to Vaani, a massive multi-modal, multi-lingual dataset designed to represent India's linguistic diversity.
OpenAI Deep Research System Card
OpenAI has introduced Deep Research, an agentic capability powered by an early version of o3 optimized for web browsing to conduct multi-step internet research for complex tasks.
FastRTC: The Real-Time Communication Library for Python
Hugging Face has released FastRTC, a Python library designed to simplify the development of real-time audio and video AI applications by handling the WebRTC and WebSocket communication layers.
QwQ-Max-Preview Release
Qwen has introduced QwQ-Max-Preview, a preview reasoning model built on Qwen2.5-Max that excels in mathematics, coding, and Agent-related workflows.
Hugging Face Remote VAEs for Inference Endpoints
Hugging Face has introduced an experimental feature to delegate the VAE decoding process to remote endpoints, reducing VRAM requirements for high-resolution image and video synthesis on consumer GPUs.
OpenAI Disrupting Malicious Uses of AI
OpenAI has released a report detailing its efforts to disrupt malicious AI use, focusing on preventing state-affiliated threat actors and authoritarian regimes from using AI for covert influence operations and cyber activity.
SigLIP 2 release notes / what's new
Google has released SigLIP 2, a family of multilingual vision-language encoders that outperform the original SigLIP across all scales in zero-shot classification, image-text retrieval, and VLM transfer performance.
Uber AI Integration for On-Demand Services
Uber uses OpenAI's technology to personalize user interactions, automate complex marketplace resolutions, and enhance workforce productivity through AI-driven co-pilots.
SmolVLM2: Bringing Video Understanding to Every Device
Hugging Face has released SmolVLM2, a family of efficient vision and video language models in 2.2B, 500M, and 256M parameter sizes designed to enable local video understanding on devices ranging from phones to servers.
PaliGemma 2 Mix Release Notes
Google has released PaliGemma 2 mix, a family of vision language models fine-tuned on a diverse mix of tasks including OCR, captioning, and object detection to demonstrate the potential of PaliGemma 2 pre-trained checkpoints.
OpenAI SWE-Lancer Benchmark Release
OpenAI has introduced SWE-Lancer, a benchmark of over 1,400 real-world freelance software engineering tasks from Upwork to measure the economic impact and technical capabilities of AI models in software development.
Hugging Face adds Hyperbolic, Nebius AI Studio, and Novita as Serverless Inference Providers
Hugging Face has integrated Hyperbolic, Nebius AI Studio, and Novita as serverless inference providers, expanding access to models like DeepSeek-R1 and FLUX.1 directly via the Hub and client SDKs.
OpenAI and Guardian Media Group Content Partnership
OpenAI and Guardian Media Group have entered a strategic partnership to integrate Guardian editorial content into ChatGPT, providing 300 million weekly users with direct access to trusted reporting and extended summaries.
Hugging Face Open LLM Leaderboard Update: Integrating Math-Verify for Improved Math Evaluation
Hugging Face has re-evaluated 3,751 models on the Open LLM Leaderboard using Math-Verify to fix parsing errors and format strictness, resulting in a significant reshuffling of the MATH-Hard rankings.
Hugging Face Integrates Fireworks.ai as an Inference Provider
Hugging Face has added Fireworks.ai as a supported Inference Provider on the Hub, enabling serverless inference for models like DeepSeek-R1 and Llama-3.2-90B-Vision-Instruct across the HF ecosystem.
Fanatics Betting and Gaming AI Implementation in Finance
Fanatics Betting and Gaming is utilizing ChatGPT and custom GPTs to automate manual finance tasks, reducing monthly workloads and accelerating strategic decision-making.
Wayfair AI Implementation and Strategy
Wayfair is integrating OpenAI's generative AI and multimodal capabilities to personalize ecommerce experiences, modernize legacy codebases, and automate risk analysis in legal workflows.
Rogo scales financial research using OpenAI o1 and GPT-4o
Rogo utilizes a layered model architecture featuring OpenAI o1 and GPT-4o to automate financial research and diligence, saving analysts over 10 hours per week.
Hugging Face: Optimizing Cost and Latency for 1 Billion Classifications
Hugging Face provides a framework and benchmarks for reducing the cost of large-scale encoder model inference, demonstrating that NVIDIA L4 GPUs and optimized batch sizes can process 1 billion text classifications for as little as $253.82.
OpenAI Model Spec Update
OpenAI has released an updated Model Spec under a CC0 license to define AI behavior, prioritizing user customizability and intellectual freedom while maintaining safety guardrails.
Hugging Face Video Dataset Scripts
Hugging Face has introduced a set of open video dataset scripts designed to simplify the creation of high-quality, filtered datasets for fine-tuning video generation models.
Hugging Face Xet-backed Repositories: Accelerating Hub Transfers with Block-Level Aggregation
Hugging Face is introducing a chunk-based deduplication system using xet-core and hf_xet to accelerate uploads and downloads by 2-3x through block-level aggregation.
Open R1 Update #2: OpenR1-Math-220k Dataset and Reasoning Insights
Hugging Face introduces OpenR1-Math-220k, a large-scale math reasoning dataset designed to reconstruct DeepSeek R1's distillation pipeline, while sharing community insights on GRPO and Chain-of-Thought length control.