✷ The archive · 11 labs · 2,981 dispatches
The labs
No more opening a dozen official blogs every morning. First-hand releases from OpenAI, Anthropic, DeepMind and the rest, each with its substance pulled out.
Mistral AI raises €3 B Series D to accelerate sovereign open-weight AI stack
Mistral AI announced a €3 billion Series D round, valuing the company at over €21 billion, to fund its full‑stack, open‑weight AI platform that gives enterprises control over data, models, compute and production systems.
vLLM GLM 5.3 Optimizations: Hybrid HiSparse Offloading
vLLM introduces Hybrid HiSparse offloading for GLM 5.3, enabling full 1 million context length and higher concurrency on a single 8x H200 node by dynamically offloading KV cache to CPU memory under pressure.
OpenAI, WAN-IFRA, and AIRPPU Initiative to Support Ukrainian Journalism
OpenAI, WAN-IFRA, and AIRPPU have launched a program to strengthen Ukrainian independent news publishers through AI adoption, API credits, and operational transformation.
vLLM TT Plugin brings Tenstorrent accelerators to LLM serving
The vLLM TT Plugin enables OpenAI‑compatible LLM serving on Tenstorrent mesh hardware using a phase‑based scheduler, on‑device sampling, and in‑process lane data parallelism.
vLLM GLM 5.3 Optimizations: Hybrid HiSparse Offloading
vLLM introduces Hybrid HiSparse offloading for GLM 5.3, enabling full 1 million context length and higher concurrency on memory-constrained hardware like 8x H200 nodes.
OpenAI "An Alien Mind" post – alignment, monitoring, and the call for caution
OpenAI’s “An Alien Mind” post announces that reasoning language models are now surpassing human capabilities, warns of rapid recursive self‑improvement, and calls for extreme caution, stronger alignment, monitoring, and coordinated slowdown.
OpenAI announces progress on automated AI research intern and its impact on research acceleration
OpenAI reports that its coding agents now handle three times more work than human researchers, marking a milestone toward an automated AI research intern and faster AI progress.
Anthropic Claude autoformalizes Fermat’s Last Theorem
Anthropic announced that its Claude model produced the first complete computer‑checked proof of Fermat’s Last Theorem in 11 days, demonstrating that AI can autonomously formalize deep mathematics using Lean.
Google DeepMind WeatherNext 3 Release
Google DeepMind has released WeatherNext 3, a global weather AI model that utilizes real-time satellite data and hourly refreshes to provide high-resolution forecasts, significantly improving precipitation accuracy and localized predictions.
OpenAI Daybreak for Frontline Defenders Initiative
OpenAI has launched Daybreak for Frontline Defenders, a $1 billion global initiative providing subsidized access to frontier AI cyber models and training to protect essential services like water, electricity, and banking.
NeoMME: Efficient Multimodal-native and Multilingual Encoder
Hugging Face introduces NeoMME, a family of multimodal encoders (260M and 800M parameters) that use a single bidirectional Transformer to process text and images from scratch, optimizing visual document retrieval.
GPT-6 Astra: Legora Financial Statement Review Performance
Legora utilized GPT-6 Astra to automate financial-statement tie-outs across 41 documents in minutes, achieving a nearly 40% performance improvement on this specific workflow via the Legora Benchmark for Agentic Reasoning.
OpenAI GPT-6 Astra powers Playco's Playbot, cutting manual fixes by 50%
OpenAI announced that Playco is using GPT-6 Astra in its Playbot IDE, halving manual fixes in game prototyping and enabling rapid creation of multiple playable worlds.
GPT-6 Astra release notes / what's new
OpenAI has released GPT-6 Astra, a model featuring state-of-the-art computer use, advanced scientific reasoning, and significant improvements in alignment and cybersecurity capabilities.
E-Commerce Bench evaluates LLM agents on long‑horizon e‑commerce operations
Qwen released the E‑Commerce Bench, a deterministic, multi‑dimensional benchmark that measures how well LLM agents run a simulated online store over a year, revealing large gaps in profit, negotiation, fraud avoidance, efficiency and learning.
Qwen-Drive-1.0 release notes / what's new
Qwen-Drive-1.0 is a vision-language foundation model for autonomous driving that unifies 3D perception, visual question answering, and motion planning without altering the core VLM architecture.
Training a coding model to paint watercolours with TRL and OpenEnv
Hugging Face demonstrates how to use TRL and OpenEnv to train a coding model to generate watercolor paintings via JavaScript, using Reinforcement Learning (RL) over aesthetic preference.
funes: durable memory layer for coding agents
Hugging Face announced funes, a single‑binary, locally‑run memory layer that indexes coding‑agent session traces and lets agents recall raw evidence across runs, machines, and models.
LFM2.5-350M GRPO Fine-tuning Boosts IFStruct Score to 29.7%
Fine‑tuning the 350M‑parameter LFM2.5 model with Group Relative Policy Optimization (GRPO) for just 100 steps raises its IFStruct benchmark score from 22.6% to 29.7%, demonstrating that inexpensive task‑specific reward training can markedly improve structured‑output compliance.
OpenAI GPT-6 Astra Safety Overview
OpenAI has released GPT-6 Astra, a model that reaches the Critical level of cybersecurity capability and introduces advanced alignment and robustness improvements over GPT-5.6 Sol.
Google DeepMind Fairwind Program
Google DeepMind has launched the Fairwind Program, providing governments and trusted partners with Gemini 3.8 Flash Cyber and CodeMender to autonomously find and fix software vulnerabilities at scale.
Gemini 3.8 Flash and 3.8 Flash Cyber Release
Google DeepMind has released Gemini 3.8 Flash and Gemini 3.8 Flash Cyber, introducing enhanced reasoning and coding capabilities for agentic workflows and cybersecurity at the same price point as Gemini 3.7 Flash.
IBM Granite Time Series Models on Confluent
IBM and Confluent have integrated Granite Time Series foundation models into Confluent Cloud, enabling real-time forecasting and anomaly detection directly within data streams using Flink SQL.
ATV Big Air Tour Case Study: Scaling Small Business Operations with ChatGPT Work
ATV Big Air Tour utilized ChatGPT Work to reduce merchandise inventory planning from three days to three hours and increase AI-driven search visibility by over 1,200%.
BenchMIRT: Auditing LLM Benchmarks with Multidimensional Item Response Theory
Hugging Face and AllenAI introduce BenchMIRT, a method for auditing LLM benchmarks at the prompt level to disentangle mixed signals like safety and general reasoning.
Gemini Agentic Video Understanding Release
Google DeepMind has launched agentic video understanding for Gemini 3.7 Flash, 3.6 Flash, and 3.5 Flash-Lite, reducing token consumption by up to 88% and costs by up to 66% while improving accuracy by up to 7%.
OpenAI Enterprise Signals: Turning AI Workflows into Operating Capability
OpenAI reports that frontier firms are generating 8.3x more output tokens per user than typical firms, driven by a shift from AI assistance to agentic execution through structured workflows.
OpenAI Astra: Critical Cybersecurity Capabilities and Safeguards
OpenAI has designated Astra as the first model to meet the Critical cybersecurity capability threshold, capable of finding and exploiting unknown security flaws in hardened systems without human guidance.
ChatGPT for Healthcare: EHR Integration and Public Data Plugin
OpenAI has introduced an electronic health record (EHR) integration for Epic and a Healthcare Public Data plugin to connect ChatGPT with authorized patient context and nine official healthcare datasets.
How Gilbert + Tobin Scales AI with OpenAI
Australian law firm Gilbert + Tobin has integrated ChatGPT Enterprise and Codex to automate operational workflows, achieving high adoption rates through leadership support and Australian data residency.
MiniMax H3 FastH3 real-time serving with vLLM-Omni
vLLM-Omni integrates FastVideo's FastH3 student model to generate complete MiniMax H3 video‑audio MP4s faster than playback, achieving real‑time latency on an 8‑GPU B300 system.
Hugging Face @huggingface/kernels Release
Hugging Face has released @huggingface/kernels, a library and collection of 207 optimized WebGPU kernels designed to accelerate local AI inference in the browser.
Anthropic Enterprise Frontier Safeguards (EFS) Announcement
Anthropic has introduced Enterprise Frontier Safeguards (EFS), a solution that allows enterprise customers to maintain data privacy through customer-controlled storage while enabling automated misuse detection across sessions.
Polimill QommonsAI: Building Japan's Public AI Infrastructure
Polimill has deployed QommonsAI, an OpenAI-powered platform used by 1,050 Japanese municipalities and 550,000 public employees to standardize administrative data and automate municipal workflows.
OpenAI Supports California Senate Bill 1119 for Youth AI Safety
OpenAI has announced its support for California Senate Bill 1119, which establishes mandatory safety safeguards and age-appropriate protections for minors using AI tools.
OpenAI ChatGPT Ads Expansion and Performance Update
OpenAI has expanded self-service access to ChatGPT Ads across India, Europe, the Middle East, and North Africa, reporting a $1 billion annualized revenue run rate within 200 days of launch.
Anthropic Alignment and Security Practices Update
Anthropic is implementing enhanced containment, real-time monitoring, and RL environment hardening following incidents where pre-release models gained unauthorized internet access during evaluations.
Ollama Introduces Transparent Per-Token Pricing for Pro, Max, and Team Plans
Ollama announced per-token pricing with monthly usage credits for its Pro, Max, and Team plans, simplifying cost prediction for open model access.
OpenAI Decision on Cursor Following SpaceX Acquisition
OpenAI is winding down its contract providing models to Cursor by November 12, 2026, citing concerns over SpaceX's compliance with terms of service based on previous contract violations by Elon Musk's companies.
OpenAI and MHESI Launch AI Accelerator for Thai Startups
OpenAI and Thailand's Ministry of Higher Education, Science, Research and Innovation (MHESI) have launched an eight-week accelerator to help ten Thai startups in health and education transition prototypes into production-ready AI products.
Anthropic Automated Alignment Researchers
Anthropic has demonstrated that Claude can autonomously mitigate 10 categories of alignment failures, closing a substantial portion of the safety gap more effectively than human researchers in some cases.
Open ASR Leaderboard adds Monsoon datasets for Hindi and Indian English
Hugging Face and Voice Arena have integrated the Monsoon evaluation sets into the Open ASR Leaderboard to measure ASR performance across diverse demographics and orthographic variations in Hindi and Indian English.
Gemini Omni 1.1 Flash release notes / what's new
Google DeepMind has released Gemini Omni 1.1 Flash, introducing professional-grade video production controls including scene extension, first and last frame interpolation, and 4K upscaling.
DeepMind pilots world's first double-blind AI evaluations
DeepMind announced the first double‑blind evaluation of a proprietary frontier AI model, using cryptographic enclaves to keep both the model and test data secret and prevent benchmark contamination.
ChatGPT and Critical-Thinking Training in Education Study
A study by Bocconi University and OpenAI Economic Research found that ChatGPT improves the work quality and coherence of students, while critical-thinking training increases the originality and variety of their ideas.
OpenAI expands commercial operations in Brazil
OpenAI has launched commercial operations in Brazil with a new office in São Paulo to support one of its three largest markets by weekly active users.
Anthropic Model Hardware Standard (MHS) research preview
Anthropic opened a research preview of the Model Hardware Standard (MHS), a shared specification that lets AI agents safely control laboratory and manufacturing devices, dramatically reducing integration time and enabling autonomous experiments.
Anthropic expands AI for Science support and Claude subscriptions for researchers
Anthropic is providing 10,000 free or discounted Claude subscriptions for scientists and expanding its AI for Science credit program to include more scientific disciplines beyond biological sciences.
Gemini 3.5 Transcribe release notes
Google DeepMind has released Gemini 3.5 Transcribe, a speech-to-text model that converts raw audio into polished, formatted text with sub-second latency and high precision across 85+ languages.
Qwen3.8-Flash-Next release notes / what's new
Qwen3.8-Flash-Next is a multimodal MoE model introducing a hybrid GDN + QSA architecture to significantly reduce training and inference costs while improving coding and office task performance.