✷ The archive · 11 labs · 3,050 dispatches
The labs
No more opening a dozen official blogs every morning. First-hand releases from OpenAI, Anthropic, DeepMind and the rest, each with its substance pulled out.
Serving Agentic Workloads at Scale with vLLM x Mooncake
vLLM integrates Mooncake's distributed KV cache store to boost agentic LLM serving, delivering 3.8× higher throughput, 46× lower TTFT, and 8.6× lower end‑to‑end latency on realistic traces while scaling to 60 GB200 GPUs.
Hugging Face Open ASR Leaderboard: Private Datasets to Combat Benchmaxxing
Hugging Face has introduced private evaluation datasets from Appen Inc. and DataoceanAI to the Open ASR Leaderboard to prevent test-set contamination and provide a more robust measure of real-world ASR performance.
OpenAI B2B Signals: How Frontier Firms are Pulling Ahead
OpenAI introduces B2B Signals to reveal that frontier firms—those in the 95th percentile of usage—now use 3.5x more intelligence per worker than typical firms, driven by complex, agentic workflows.
OpenAI Introduces ChatGPT Futures: Class of 2026
OpenAI has launched ChatGPT Futures, a program providing grants and model access to students in the Class of 2026 who use AI to build tools, advance research, and create social impact.
Singular Bank Integrates ChatGPT and Codex to Launch Singularity Assistant
Singular Bank has developed Singularity, an internal AI assistant powered by ChatGPT and Codex, reducing banker productivity losses by 60 to 90 minutes per day through automated portfolio analysis and communication.
Uber and OpenAI Integration: AI-Powered Driver Guidance and Voice Booking
Uber is integrating OpenAI frontier models to launch Uber Assistant for driver earnings optimization and voice-driven ride booking to reduce friction for riders.
Grok Imagine Quality Mode API Release
xAI has launched Quality Mode for the Grok Imagine API, providing enterprise developers with enhanced realism, improved multilingual text rendering, and superior creative control for image generation and editing.
xAI Grok Connectors Release
xAI has launched Connectors, deep integrations for Grok on web, iOS, and Android that allow the AI to read and write data across third-party applications like Google Workspace, GitHub, and Notion.
xAI and Anthropic Compute Partnership
xAI has partnered with Anthropic to provide access to the Colossus 1 supercomputer, with additional plans to explore orbital AI compute capacity.
Anthropic Expands Compute Capacity via SpaceX Partnership and Increases Claude Usage Limits
Anthropic has partnered with SpaceX to access the Colossus 1 data center, enabling higher usage limits for Claude Code and the Claude API.
GPT-5.5 Instant System Card
OpenAI has released the GPT-5.5 Instant system card, marking the first time an Instant model is classified as High capability in Cybersecurity and Biological & Chemical Preparedness categories.
GPT-5.5 Instant release notes / what's new
OpenAI has released GPT-5.5 Instant, a new default model for ChatGPT that improves factuality, reduces verbosity, and enhances personalization through deeper integration with user context.
OpenAI MRC Multipath Reliable Connection Protocol
OpenAI has released the Multipath Reliable Connection (MRC) protocol via the Open Compute Project to improve GPU networking resilience and performance in large-scale AI training clusters.
OpenAI European Youth Safety Blueprint and EMEA Youth & Wellbeing Grant
OpenAI announced its European Youth Safety Blueprint and the first recipients of the EMEA Youth & Wellbeing Grant to promote age‑appropriate AI use and protect young people in Europe, the Middle East, and Africa.
OpenAI expands ChatGPT ads pilot with self-serve Ads Manager and CPC bidding
OpenAI is expanding its ChatGPT ads pilot by introducing a beta self-serve Ads Manager, cost-per-click (CPC) bidding, and expanded measurement tools to make advertising more accessible to businesses of all sizes.
OpenAI and PwC Collaboration for AI Agents in Finance
OpenAI and PwC are collaborating to deploy AI agents that automate finance workflows, using OpenAI's own finance organization as a testing ground to modernize the office of the CFO.
OpenAI Low-Latency Voice AI Infrastructure
OpenAI rearchitected its WebRTC stack using a split relay-plus-transceiver architecture to deliver low-latency, scalable voice AI to over 900 million weekly active users.
Anthropic Forms New Enterprise AI Services Company with Blackstone, Hellman & Friedman, and Goldman Sachs
Anthropic has partnered with Blackstone, Hellman & Friedman, and Goldman Sachs to launch a new AI services company focused on deploying Claude for mid-sized organizations.
Google DeepMind AI Co-Clinician Research Initiative
Google DeepMind has introduced the AI co-clinician research initiative, a multimodal AI system designed to support healthcare providers through evidence synthesis and real-time patient interaction under clinical supervision.
Qwen-Scope Interpretability Toolkit Release
Qwen has released Qwen-Scope, an interpretability toolkit using Sparse Autoencoders (SAEs) to decompose hidden representations of Qwen3 and Qwen3.5 models into interpretable features for model optimization.
OpenAI Advanced Account Security Release
OpenAI has launched Advanced Account Security, an opt-in setting for ChatGPT and Codex accounts that mandates phishing-resistant authentication and disables traditional recovery methods to protect high-risk users.
Anthropic Claude personal guidance study reveals usage patterns and reduces sycophancy in Opus 4.7 and Mythos Preview
Anthropic announced that about 6% of Claude conversations are personal guidance requests and that the new Opus 4.7 and Mythos Preview models cut sycophantic responses in half, especially for relationship advice.
xAI Custom Voices Release
xAI has introduced Custom Voices, allowing users to clone their own voice from a short recording for use across Grok Text to Speech and Voice Agent APIs.
OpenAI Analysis of Model Behavior: The 'Goblin' Lexical Tic
OpenAI identified that a reward signal for the 'Nerdy' personality feature caused models from GPT-5.1 to GPT-5.5 to develop an unintended habit of using goblin and gremlin metaphors, which then generalized across other model behaviors.
IBM Granite 4.1 LLMs release notes / technical overview
IBM has released Granite 4.1, a family of dense, decoder-only LLMs (3B, 8B, and 30B) that achieve high performance through rigorous data curation and a multi-stage reinforcement learning pipeline.
OpenAI Stargate Compute Infrastructure Update
OpenAI has surpassed its initial 10GW AI infrastructure goal for the United States and trained GPT-5.5 at its flagship Stargate site in Abilene, Texas.
OpenAI Cybersecurity Action Plan for the Intelligence Age
OpenAI has introduced a five-pillar Action Plan to democratize AI-powered cyber defense and strengthen resilience against AI-driven threats.
DeepInfra Integration with Hugging Face Inference Providers
Hugging Face has added DeepInfra as a supported Inference Provider, enabling serverless access to over 100 models, including DeepSeek V4 and Kimi-K2.6, directly through the Hub and client SDKs.
Evaluating Claude's Bioinformatics Research Capabilities with BioMysteryBench
Anthropic introduced BioMysteryBench, a new bioinformatics benchmark where latest Claude models perform on par with human experts and solve several problems that human experts could not.
NVIDIA Nemotron 3 Nano Omni release notes / what's new
NVIDIA has released Nemotron 3 Nano Omni, an omni-modal model capable of long-context reasoning across text, images, video, and audio, delivering best-in-class accuracy on document intelligence and video understanding benchmarks.
FlashQLA: CP-/Bwd-Friendly Fused Linear Attention Kernels for GDN
Qwen has open-sourced FlashQLA, a high-performance linear attention kernel library built on TileLang that achieves 2-3x forward and 2x backward speedups for Gated Delta Network (GDN) layers on NVIDIA Hopper GPUs.
NVIDIA Nemotron 3 Nano Omni Support in vLLM
vLLM now supports NVIDIA Nemotron 3 Nano Omni, a highly efficient 30B MoE multimodal model that unifies vision, audio, and language reasoning in a single loop to power agentic AI.
OpenAI Community Safety and Violence Mitigation Framework
OpenAI has detailed its multi-layered approach to preventing the use of ChatGPT for planning or executing violence, combining model training, automated detection, and human review.
OpenAI GPT-5.5, Codex, and Managed Agents on AWS release notes
On April 28 2026, OpenAI announced that its frontier models including GPT-5.5, Codex coding agent, and Amazon Bedrock Managed Agents powered by OpenAI are now available in limited preview on AWS, enabling enterprises to use these capabilities within their existing AWS security, compliance, and procurement workflows.
OpenAI achieves FedRAMP Moderate authorization
OpenAI has achieved FedRAMP 20x Moderate authorization for ChatGPT Enterprise and its API Platform, enabling U.S. government agencies to access frontier AI models including GPT-5.5 with federal security and privacy standards.
Mistral AI Workflows Public Preview
Mistral AI has released Workflows in public preview, providing an orchestration layer for enterprise AI that ensures durability, observability, and fault tolerance for production AI processes.
Google DeepMind Partnership with the Republic of Korea
Google DeepMind has partnered with South Korea's Ministry of Science and ICT to establish an AI Campus in Seoul and deploy frontier AI models to accelerate scientific research in life sciences, energy, and climate.
The Next Phase of the Microsoft OpenAI Partnership
OpenAI and Microsoft have amended their partnership agreement to provide greater flexibility, allowing OpenAI to serve products across any cloud provider while maintaining Microsoft as its primary cloud partner.
Choco automates food distribution with AI agents
Choco announced AI-powered OrderAgent and VoiceAgent built with OpenAI APIs to automate food distribution order processing.
OpenAI Privacy Filter: Building Scalable PII Detection Web Apps
OpenAI has released Privacy Filter, an open-source 1.5B-parameter PII detector capable of labeling eight categories of sensitive data across a 128k context window.
Symphony: Open‑Source Spec for Codex Orchestration
OpenAI released Symphony, an open‑source specification that turns issue trackers like Linear into always‑on orchestrator for Codex coding agents, enabling teams to automate routine implementation work and increase PR throughput.
OpenAI Operating Principles for AGI Development
OpenAI has outlined five core principles—Democratization, Empowerment, Resilience, Universal Prosperity, and Adaptability—to ensure that artificial general intelligence (AGI) benefits all of humanity.
OpenAI AI Jobs Transition Framework
OpenAI has introduced the AI Jobs Transition Framework to analyze how AI affects employment across 921 occupations, categorizing jobs into four paths based on automation risk, reorganization, growth potential, and stability.
vLLM DeepSeek V4 Support: Efficient Long-context Attention
vLLM now supports DeepSeek V4-Pro and V4-Flash, implementing a new attention mechanism that enables context lengths up to one million tokens with significant KV cache memory savings.
DeepSeek-V4 Release Notes: Efficient 1M-Token Context for AI Agents
DeepSeek-V4 introduces a 1M-token context window powered by a hybrid CSA/HCA attention mechanism, specifically optimized for long-running agentic workloads and tool-use trajectories.
DeepSeek-V4 Preview Release
DeepSeek has released the open-source DeepSeek-V4 preview, introducing two model variants—Pro and Flash—featuring a 1M token context window and state-of-the-art agentic coding capabilities.
Anthropic and NEC Partnership for AI-Native Engineering in Japan
Anthropic and NEC have partnered to deploy Claude to 30,000 NEC employees and develop secure, industry-specific AI products for the Japanese finance, manufacturing, and government sectors.
Anthropic Election Safeguards Update
Anthropic has implemented a comprehensive suite of safeguards, including neutrality training and automated classifiers, to ensure Claude remains impartial and secure during the 2026 US midterms and other global elections.
OpenAI GPT-5.5 Release Notes
OpenAI has released GPT-5.5 and GPT-5.5 Pro, introducing significant advancements in agentic coding, computer use, and scientific research while maintaining the latency of GPT-5.4.
OpenAI GPT-5.5 System Card
OpenAI has released GPT-5.5, a model optimized for complex real-world tasks and tool use, featuring enhanced task understanding and a more robust safety framework.