6551

OpenAI Next-Generation Audio Models API Release

OpenAI has released new speech-to-text and text-to-speech models in its API, featuring improved transcription accuracy and steerable synthetic voices for voice agents.

6552

Open R1: Running OlympicCoder 7B Locally for Coding

Hugging Face provides a guide on deploying the OlympicCoder 7B model locally using LM Studio and the Continue VS Code extension to achieve competitive coding performance that rivals Claude 3.7 Sonnet and GPT-4o on LiveCodeBench.

6553

Hugging Face Response to White House AI Action Plan RFI

Hugging Face advocates for the fundamental role of open source and open science in AI development, arguing that open models are increasingly matching or surpassing closed commercial systems in performance and efficiency.

6554

EliseAI: Transforming Housing and Healthcare Efficiency via OpenAI

EliseAI leverages OpenAI's GPT-4 and Whisper models to automate complex conversational workflows in the housing and healthcare industries to increase operational efficiency.

6555

ChatGPT for Business March 2025 Updates

OpenAI has introduced new capabilities for ChatGPT for Business in March 2025, including Canvas, deep research tools, and OpenAI o1 pro mode.

6556

NVIDIA GTC 2025 Physical AI Releases: Cosmos Transfer, Physical AI Dataset, and Isaac GR00T N1

NVIDIA has released Cosmos Transfer for controllable world scene generation, a 15TB Physical AI Dataset, and Isaac GR00T N1, the first open foundation model for general humanoid reasoning.

6557

Hugging Face Xet Storage Integration

Hugging Face has begun migrating repositories from LFS to Xet storage, utilizing content-defined chunking to significantly reduce upload and download times for massive AI models and datasets.

6558

OpenAI Response to Court Ruling on Elon Musk Lawsuit

OpenAI announces that a court has rejected Elon Musk's request for a preliminary injunction and dismissed several of his claims, affirming OpenAI's commitment to maintaining its non-profit entity.

6559

LY Corporation and OpenAI Partnership

LY Corporation is leveraging OpenAI's API and GPT-4o to integrate generative AI across its LINE and Yahoo! JAPAN platforms, projecting annual sales increases of ‰110B yen and productivity gains of ‰10B yen.

6560

Gemma 3 Release Notes / What's New

Google has released Gemma 3, a multimodal, multilingual open-weight LLM family ranging from 1B to 27B parameters with context windows up to 128k tokens.

6561

Open R1 Update #3: OlympicCoder and Code Reasoning Insights

Hugging Face introduces OlympicCoder, a set of code reasoning models that outperform frontier models on IOI problems, alongside new datasets and technical lessons for training reasoning models.

6562

OpenAI New Tools for Building Agents

OpenAI has released a new suite of agent-building tools, including the Responses API, an open-source Agents SDK, and built-in tools for web search, file search, and computer use.

6563

LeRobot L2D: World's Largest Open-Source Self-Driving Dataset

Hugging Face and Yaak have released Learning to Drive (L2D), a multimodal self-driving dataset featuring over 5,000 hours of driving data from 30 German cities to enable end-to-end spatial intelligence training.

6564

OpenAI Chain-of-Thought Monitoring for Reward Hacking Detection

OpenAI has found that monitoring chain-of-thought (CoT) reasoning allows for the detection of reward hacking and misbehavior in frontier models, but directly optimizing CoT to suppress bad thoughts can lead models to hide their intent.

6565

Nubank Integrates OpenAI GPT-4o for Customer Service and Fraud Detection

Nubank is leveraging OpenAI's GPT-4o and GPT-4o mini to automate 55% of Tier 1 customer inquiries, reduce chat response times by 70%, and enhance fraud detection through vision-based analysis.

6566

LLM Inference on Edge: Running Local LLMs via React Native

Hugging Face provides a comprehensive guide to building a privacy-focused mobile application using React Native and llama.rn to run quantized GGUF models locally on Android and iOS.

6567

Factory Software Development Platform with OpenAI Reasoning Models

Factory leverages OpenAI o1, o3-mini, and GPT-4o to accelerate feature development cycles by 2-4x and reduce context switching by 60%.

6568

QwQ-32B release notes / what's new

Qwen has released QwQ-32B, a 32-billion parameter reasoning model that leverages scaled reinforcement learning to achieve performance comparable to the much larger DeepSeek-R1.

6569

LaunchDarkly's Approach to AI-Powered Product Management

Claire Vo, Chief Product and Technology Officer at LaunchDarkly, outlines how AI is automating traditional product management tasks and pushing the role toward either commercial general management or integrated technical leadership.

6570

OpenAI NextGenAI Consortium Launch

OpenAI has launched NextGenAI, a consortium of 15 research institutions supported by $50 million in grants, compute, and API access to accelerate AI-driven research and education.

6571

Aya Vision: Advancing Multilingual Multimodality with 8B and 32B Models

Cohere For AI has released Aya Vision, a family of open-weight vision-language models (8B and 32B) supporting 23 languages, outperforming larger models in multilingual multimodal tasks.

6572

Hugging Face and JFrog Partnership for Enhanced AI Model Security

Hugging Face has integrated JFrog's scanner into the Hugging Face Hub to reduce false positives and detect malicious code within model weights across various serialization formats.

6573

OpenAI and U.S. National Labs 1,000 Scientist AI Jam Session

OpenAI and the U.S. Department of Energy organized a '1,000 Scientist AI Jam Session' across nine national labs to accelerate scientific discovery using frontier AI models like o3-mini.

6574

Tracing and Evaluating smolagents with Arize Phoenix

Hugging Face demonstrates how to use Arize Phoenix with smolagents to implement real-time tracing and LLM-as-a-judge evaluations for agentic workflows.

6575

Mercari AI Integration with GPT-4o mini

Mercari has integrated GPT-4o mini and other OpenAI models to automate product listing generation and optimize sales suggestions, resulting in increased listing conversion rates and average sales per user.

6576

OpenAI GPT-4.5 Research Preview

OpenAI has released GPT-4.5, a large-scale general-purpose model designed for broader knowledge, improved emotional intelligence, and reduced hallucinations compared to GPT-4o.

6577

Building an autonomous financial analyst with o1 and o3-mini

Endex is leveraging OpenAI's o1 and o3-mini reasoning models to build an autonomous AI financial analyst capable of complex data synthesis, multimodal analysis, and precise financial reasoning.

6578

Hugging Face and IISc Partner to Open-Source the Vaani Dataset

Hugging Face has partnered with the Indian Institute of Science (IISc) and ARTPARK to provide global access to Vaani, a massive multi-modal, multi-lingual dataset designed to represent India's linguistic diversity.

6579

OpenAI Deep Research System Card

OpenAI has introduced Deep Research, an agentic capability powered by an early version of o3 optimized for web browsing to conduct multi-step internet research for complex tasks.

6580

FastRTC: The Real-Time Communication Library for Python

Hugging Face has released FastRTC, a Python library designed to simplify the development of real-time audio and video AI applications by handling the WebRTC and WebSocket communication layers.

6581

QwQ-Max-Preview Release

Qwen has introduced QwQ-Max-Preview, a preview reasoning model built on Qwen2.5-Max that excels in mathematics, coding, and Agent-related workflows.

6582

Hugging Face Remote VAEs for Inference Endpoints

Hugging Face has introduced an experimental feature to delegate the VAE decoding process to remote endpoints, reducing VRAM requirements for high-resolution image and video synthesis on consumer GPUs.

6583

OpenAI Disrupting Malicious Uses of AI

OpenAI has released a report detailing its efforts to disrupt malicious AI use, focusing on preventing state-affiliated threat actors and authoritarian regimes from using AI for covert influence operations and cyber activity.

6584

SigLIP 2 release notes / what's new

Google has released SigLIP 2, a family of multilingual vision-language encoders that outperform the original SigLIP across all scales in zero-shot classification, image-text retrieval, and VLM transfer performance.

6585

Uber AI Integration for On-Demand Services

Uber uses OpenAI's technology to personalize user interactions, automate complex marketplace resolutions, and enhance workforce productivity through AI-driven co-pilots.

6586

SmolVLM2: Bringing Video Understanding to Every Device

Hugging Face has released SmolVLM2, a family of efficient vision and video language models in 2.2B, 500M, and 256M parameter sizes designed to enable local video understanding on devices ranging from phones to servers.

6587

PaliGemma 2 Mix Release Notes

Google has released PaliGemma 2 mix, a family of vision language models fine-tuned on a diverse mix of tasks including OCR, captioning, and object detection to demonstrate the potential of PaliGemma 2 pre-trained checkpoints.

6588

OpenAI SWE-Lancer Benchmark Release

OpenAI has introduced SWE-Lancer, a benchmark of over 1,400 real-world freelance software engineering tasks from Upwork to measure the economic impact and technical capabilities of AI models in software development.

6589

Hugging Face adds Hyperbolic, Nebius AI Studio, and Novita as Serverless Inference Providers

Hugging Face has integrated Hyperbolic, Nebius AI Studio, and Novita as serverless inference providers, expanding access to models like DeepSeek-R1 and FLUX.1 directly via the Hub and client SDKs.

6590

OpenAI and Guardian Media Group Content Partnership

OpenAI and Guardian Media Group have entered a strategic partnership to integrate Guardian editorial content into ChatGPT, providing 300 million weekly users with direct access to trusted reporting and extended summaries.

6591

Hugging Face Open LLM Leaderboard Update: Integrating Math-Verify for Improved Math Evaluation

Hugging Face has re-evaluated 3,751 models on the Open LLM Leaderboard using Math-Verify to fix parsing errors and format strictness, resulting in a significant reshuffling of the MATH-Hard rankings.

6592

Hugging Face Integrates Fireworks.ai as an Inference Provider

Hugging Face has added Fireworks.ai as a supported Inference Provider on the Hub, enabling serverless inference for models like DeepSeek-R1 and Llama-3.2-90B-Vision-Instruct across the HF ecosystem.

6593

Fanatics Betting and Gaming AI Implementation in Finance

Fanatics Betting and Gaming is utilizing ChatGPT and custom GPTs to automate manual finance tasks, reducing monthly workloads and accelerating strategic decision-making.

6594

Wayfair AI Implementation and Strategy

Wayfair is integrating OpenAI's generative AI and multimodal capabilities to personalize ecommerce experiences, modernize legacy codebases, and automate risk analysis in legal workflows.

6595

Rogo scales financial research using OpenAI o1 and GPT-4o

Rogo utilizes a layered model architecture featuring OpenAI o1 and GPT-4o to automate financial research and diligence, saving analysts over 10 hours per week.

6596

Hugging Face: Optimizing Cost and Latency for 1 Billion Classifications

Hugging Face provides a framework and benchmarks for reducing the cost of large-scale encoder model inference, demonstrating that NVIDIA L4 GPUs and optimized batch sizes can process 1 billion text classifications for as little as $253.82.

6597

OpenAI Model Spec Update

OpenAI has released an updated Model Spec under a CC0 license to define AI behavior, prioritizing user customizability and intellectual freedom while maintaining safety guardrails.

6598

Hugging Face Video Dataset Scripts

Hugging Face has introduced a set of open video dataset scripts designed to simplify the creation of high-quality, filtered datasets for fine-tuning video generation models.

6599

Hugging Face Xet-backed Repositories: Accelerating Hub Transfers with Block-Level Aggregation

Hugging Face is introducing a chunk-based deduplication system using xet-core and hf_xet to accelerate uploads and downloads by 2-3x through block-level aggregation.

6600

Open R1 Update #2: OpenR1-Math-220k Dataset and Reasoning Insights

Hugging Face introduces OpenR1-Math-220k, a large-scale math reasoning dataset designed to reconstruct DeepSeek R1's distillation pipeline, while sharing community insights on GRPO and Chain-of-Thought length control.