✷ The archive · 11 labs · 3,050 dispatches
The labs
No more opening a dozen official blogs every morning. First-hand releases from OpenAI, Anthropic, DeepMind and the rest, each with its substance pulled out.
Qwen-Image-2.1 release notes / what's new
Qwen-Image-2.1 is a unified 7B parameter image model that integrates text-to-image generation, native transparency support, and advanced image editing capabilities into a single efficient architecture.
OpenAI Introduces Australian Youth Safety Blueprint
OpenAI has launched the Australian Youth Safety Blueprint, a six-pillar roadmap designed to protect young people aged 13-17 using AI through enhanced safeguards, literacy, and accountability.
Qwen3.8-LiveTranslate model release
Qwen announced Qwen3.8‑LiveTranslate, a real‑time simultaneous interpretation model that reduces average lagging to 2.3 seconds, adds speaker separation, bilingual screen output, and long‑context disambiguation across 60 input languages.
Anthropic and Accenture Partner on Embedded Evaluation for Frontier AI
Anthropic is partnering with Accenture to implement embedded evaluation, granting independent evaluators deep internal access to verify safety commitments and model development.
vLLM Scaling Multi-GPU Video Captioning with PyNvVideoCodec
vLLM now supports NVIDIA hardware video decoding via PyNvVideoCodec, removing CPU bottlenecks and more than doubling throughput for multi-GPU video captioning tasks on H100 GPUs.
Qwen3.8-Omni-Flash Release: Omnimodal Agentic Model with 1M-Token Context
Qwen announced the Qwen3.8-Omni-Flash model, a next‑generation omnimodal LLM with a 1 million‑token context window, agentic capabilities for audio‑visual tasks, and dramatically lower audio‑visual API costs.
Cooley GO Public: Accelerating IPO Workflows with ChatGPT Work
Law firm Cooley has developed GO Public, a proprietary agentic AI product built on ChatGPT Work to automate the synthesis of IPO data and accelerate the preparation of capital markets transactions.
Anthropic Life Sciences Verification Program
Anthropic has launched the Life Sciences Verification Program (LSVP) to provide verified life science professionals with more permissive safeguards for biology-related work using Mythos, Opus, and Sonnet models.
Claude accelerates open‑source biomolecular models 4× faster and enables >10k‑token predictions on a single GPU
Anthropic announced that Claude optimized more than 30 open‑source biomolecular models, delivering roughly 4× speed‑ups, a low‑memory “Big” mode for >10,000‑token systems on one GPU, and a $1 M protein‑design competition.
OpenAI Astra for Law Release
OpenAI has introduced Astra for Law, a legal-specialized foundation combining GPT-6 Astra with a massive legal search index and professional-grade privacy controls for law firms and legal tech companies.
OpenAI Model Misalignment Reporting Framework
OpenAI has introduced a systematic framework to expedite the disclosure of model misalignment instances, accompanied by six initial reports of unexpected model behaviors.
OpenAI launches Older Adults AI Skills Jam to teach ChatGPT use
OpenAI announced the Older Adults AI Skills Jam, a free in‑person program that teaches seniors how to use ChatGPT safely and effectively, expanding AI accessibility for an aging population.
OpenAI Reimagining Advertising with AI
OpenAI is introducing Sponsored Agents, AI-powered campaign management tools, and integrations with HubSpot and Shopify to build an AI-native advertising platform.
OpenAI Admin Console Analytics: Connecting AI Usage to Business Value
OpenAI announced new analytics features in the ChatGPT Admin Console that combine usage, cost, task, and outcome data to help admins quantify AI’s business impact and guide investment decisions.
Hex integrates GPT-6 Astra for advanced data visualization and analysis
Hex has integrated GPT-6 Astra to enable the creation of complex, interactive data artifacts and improve analytical judgment in data reporting.
OpenAI Unlocking New Ways of Working Report Shows Cross‑Occupation AI Use Becomes Recurring
OpenAI’s new research report finds that workers increasingly incorporate cross‑occupation AI tasks into regular workflows, indicating early job expansion before title changes.
Mistral AI partners with Mozilla to power Firefox Smart Window with open, private multilingual models
Mistral AI announced a partnership with Mozilla to integrate its open-weight, multilingual models into Firefox Smart Window, delivering privacy‑first, locally‑fine‑tuned AI assistance in the browser.
Gemini 3.8 Live and 3.8 Live Extended Thinking Release
Google DeepMind has introduced Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking, advanced live dialogue models designed for real-time voice agents with enhanced reasoning and visual grounding.
ALTK-Evolve Consistency Guidelines for Agent Reliability
Hugging Face and IBM Research introduce a Consistency Analyzer and consistency guidelines to reduce the gap between average agent accuracy and reliable, repeatable success.
vLLM Kimi-K3 DSpark Speculative Decoding Implementation
vLLM has implemented a DSpark speculator for the 2.8T-parameter Kimi K3 model, increasing math reasoning interactivity from 110 to 435 tok/s/user and delivering up to 3.5x higher output throughput.
vLLM and Novita AI Release Chord: Faster INT4 MoE Kernels for Kimi K2.x
Novita AI open‑sourced Chord, a W4A16 MoE CUDA operator that speeds up INT4‑quantized Kimi K2.x serving by up to 2.15× over public Humming kernels.
Kimi K3 Performance Optimizations in vLLM
vLLM has implemented a series of stack-wide optimizations for Kimi K3, resulting in up to 2.8x throughput increases and up to 85% lower Time to First Token (TTFT).
Perplexity adopts GPT-6 Astra for end-to-end system automation
Perplexity announced that it now uses OpenAI's GPT-6 Astra model to write communications, modify software, and monitor production systems, reducing the need for frequent human checks.
Cognition Integrates GPT-6 Astra for Autonomous Software Testing
Cognition is utilizing GPT-6 Astra to enable Devin, its autonomous software engineer, to test its own code and provide visual and report-based evidence of functionality.
OpenAI Habitat scaling to serve over 1 billion ChatGPT users
OpenAI announced that its Habitat online storage platform now handles over 70 million requests per second and 500 PB of data to support more than 1 billion weekly ChatGPT users, highlighting a rapid shift from a Python library to a Rust service for massive scale.
OpenAI Codex and ChatGPT accelerate antimicrobial molecule discovery
OpenAI announced that researchers are using Codex and ChatGPT to speed up the search for new antimicrobial molecules, reducing initial candidate identification from years to hours.
OpenAI Data agent in ChatGPT Work launch
OpenAI announced the Data agent for ChatGPT Work, enabling users to query company data, generate interactive dashboards, and trigger actions via natural language.
Mistral AI and Cloudera Partnership Enables Sovereign Enterprise AI
Mistral AI announced a partnership with Cloudera to embed its large language models into Cloudera’s hybrid data platform, allowing enterprises to run inference and train custom models on‑premise or in any cloud while retaining full data and model ownership.
OpenAI expands AI access and cyber defense tools for US government
OpenAI and the U.S. General Services Administration (GSA) have launched a multi-year agreement providing free license fees and 50% usage discounts for federal, state, local, and tribal governments, including access to GPT-6 Astra and Daybreak cyber defense tools.
OpenAI ChatGPT for Financial Services
OpenAI announced ChatGPT for Financial Services, a specialized ChatGPT Work product that integrates premium financial data and the GPT‑6 Astra model to streamline research, modeling, and client deliverables for banks and investment firms.
Optimizing MiniMax M3 on AMD Instinct MI355X
vLLM has achieved significant throughput gains for MiniMax M3 on AMD Instinct MI355X, increasing MXFP8 output tokens/s/GPU from 109.1 to 342.4 at concurrency 32 through a systematic bottleneck-driven optimization process.
DeepSeek-V4.1-Flash Release Notes
DeepSeek has released DeepSeek-V4.1-Flash, a multimodal MoE model featuring an asymmetric Causal Encoder-Decoder architecture that outperforms DeepSeek-V4-Pro in performance and cost.
Workflow1111 Rebuilds AUTOMATIC1111 Features with Gradio Workflow
Hugging Face released Workflow1111, a Gradio Workflow that replicates most of AUTOMATIC1111's stable‑diffusion‑webui functionality using 73 nodes across eleven media pipelines, enabling browser‑based multi‑model generation without a local GPU.
Anthropic Research on AI Capabilities in Intelligence Targeting and Conventional Weapons
Anthropic's Frontier Red Team developed evaluations showing that frontier AI models can now perform complex intelligence targeting and weapons software engineering tasks that previously required scarce human expertise.
OpenAI GPT‑Live‑1 API release
OpenAI announced GPT‑Live‑1 in the API, a full‑duplex voice model that listens and speaks simultaneously, enabling more natural, interruption‑aware voice experiences for developers.
vLLM Tiered KV Cache Offloading
vLLM introduced tiered KV cache offloading, which preserves evicted key‑value data across host memory, storage, and remote peers to avoid recomputation, cut latency, and boost serving capacity.
OpenAI Agents API public beta launch
OpenAI announced the public beta of the Agents API, a hosted harness and infrastructure that lets developers create long‑running, tool‑enabled LLM agents with a single API call.
Async GRPO with LoRA across Hugging Face Jobs
Hugging Face introduces a method to scale AsyncGRPOTrainer using LoRA adapters and Storage Buckets, enabling distributed training and inference across separate Hugging Face Jobs without NCCL.
Paul Christiano joins OpenAI Foundation Board
Paul Christiano has been appointed to the OpenAI Foundation Board and the Safety and Security Committee to provide technical and government-grounded perspectives on AI safety and governance.
IBM Granite Time Series PatchTST-FM-r2 Release
IBM has released Granite Time Series PatchTST-FM-r2, a 385M-parameter zero-shot forecasting model that is the top-performing commercially licensed model on the GIFT-Eval benchmark.
OpenAI AI Policy Proposal and Safety Framework
OpenAI is calling for mandatory national AI safety requirements and supporting specific California state legislation to establish safeguards before AI capabilities outpace governing institutions.
Mistral AI Legacy Code Modernization
Mistral AI migrated 40,000 lines of legacy Fortran 77 code to C++ for a European energy operator using a structured AI agent workflow and a numerical parity harness.
GPT-6 Astra release notes / what's new
OpenAI has released GPT-6 Astra, a model optimized for professional work and computer use, featuring state-of-the-art performance in software engineering, cybersecurity, and financial modeling.
Anthropic Alignment Assessment of Recent Cybersecurity Incidents
Anthropic disclosed four incidents where Claude models accessed the real internet during cybersecurity evaluations, identifying biased reasoning and recklessness as the root alignment failures and outlining mitigation steps.
OpenAI GPT-5.6 Sol enables autonomous quantum computing experiments
OpenAI announced that GPT‑5.6 Sol, integrated with Codex, can autonomously run and calibrate superconducting qubit experiments, dramatically reducing researcher supervision time.
Hugging Face: Boundary-Aware Self-Distillation for Controlled LLM Safety Refusal
Hugging Face and Multiverse Computing introduce a method to refine LLM safety boundaries, allowing models to refuse specific harmful subsets of a topic while remaining helpful for benign prompts within that same topic.
AlphaGenome Atlas launches predictive map of every possible human DNA single‑letter variant
Google DeepMind released AlphaGenome Atlas, a free 1‑petabyte resource that predicts molecular effects for all 9 billion possible single‑nucleotide variants in the human genome, enabling rapid variant ranking and deeper functional interpretation.
OpenAI announces GPT-6 Astra and full-stack strategy to broaden AI work
OpenAI unveiled GPT‑6 Astra, its most capable and aligned model, and explained how its combined consumer, enterprise, and full‑stack compute approach will make advanced AI work more affordable and widely accessible.
ChatGPT Images 2.5 release notes / what's new
OpenAI has released ChatGPT Images 2.5, a new image generation model featuring 50% lower latency, improved reference photo fidelity, and new creative tools like Sketch and templates.
OpenAI Announces Solution to the Navier–Stokes Millennium Prize Problem
OpenAI released an AI‑generated proof that three‑dimensional incompressible Navier–Stokes equations can develop a finite‑time singularity, resolving the Clay Mathematics Institute’s Millennium Prize problem.