The archive · 11 labs · 3,050 dispatches

The labs

No more opening a dozen official blogs every morning. First-hand releases from OpenAI, Anthropic, DeepMind and the rest, each with its substance pulled out.

951

GPT-5.3-Codex release notes / what's new

OpenAI has released GPT-5.3-Codex, an agentic coding model that improves upon GPT-5.2-Codex in performance, reasoning, and speed, enabling it to execute complex, long-running technical tasks autonomously.

952

Claude Opus 4.6: LLM-Discovered Zero-Day Vulnerabilities

Anthropic's Claude Opus 4.6 can identify high-severity zero-day vulnerabilities in well-tested open source codebases by reasoning about code like a human researcher.

953

Quantifying Infrastructure Noise in Agentic Coding Evals

Anthropic research reveals that infrastructure resource configuration can cause score variances of up to 6 percentage points in agentic coding benchmarks, potentially masking or mimicking genuine model capability differences.

954

Building a C Compiler with Parallel Claude Agent Teams

Anthropic researcher Nicholas Carlini used 16 parallel Claude Opus 4.6 agents to autonomously build a 100,000-line Rust-based C compiler capable of compiling the Linux 6.9 kernel.

955

Mistral AI Voxtral Transcribe 2 Release

Mistral AI has released Voxtral Transcribe 2, featuring Voxtral Mini Transcribe V2 for high-efficiency batch processing and Voxtral Realtime for ultra-low latency live transcription.

956

OpenAI Codex App Server Architecture and Integration

OpenAI has introduced the Codex App Server, a JSON-RPC based protocol and process that exposes the Codex agent harness to various clients, enabling a consistent agent experience across IDEs, web runtimes, and terminal interfaces.

957

Hugging Face Community Evals

Hugging Face has introduced Community Evals, a decentralized system for reporting and aggregating model benchmark scores directly on the Hub to increase transparency and reproducibility.

958

Anthropic Commits to Ad-Free Experience for Claude

Anthropic has announced that Claude will remain ad-free to ensure the AI assistant acts exclusively in the users' interests and maintains a focused environment for deep thinking and work.

959

H Company Holo2-235B-A22B Preview Release

H Company has released Holo2-235B-A22B Preview, a UI localization model that achieves state-of-the-art performance on Screenspot-Pro and OSWorld G benchmarks.

960

The Future of the Global Open-Source AI Ecosystem: From DeepSeek to AI+

Hugging Face analyzes how open source has become the dominant strategy for Chinese AI organizations, shifting from isolated model breakthroughs to a scalable, integrated ecosystem of models, hardware, and infrastructure.

961

Training Design for Text-to-Image Models: Lessons from Ablations

The provided source material for the Hugging Face post on text-to-image model training design is unavailable due to a 429 Too Many Requests error.

962

OpenAI Sora Feed Philosophy

OpenAI has detailed the design principles and safety frameworks for the Sora feed, focusing on creativity-driven ranking, personalized recommendations, and multi-layered safety guardrails.

963

Apple Xcode 26.3 Claude Agent SDK Integration

Apple Xcode 26.3 integrates the Claude Agent SDK, enabling autonomous coding tasks, visual verification via Xcode Previews, and project-wide reasoning within the IDE.

964

Qwen3-Coder-Next Release: High-Efficiency Agentic Coding Model

Qwen3-Coder-Next is an open-weight model based on a hybrid attention and MoE architecture that achieves over 70% on SWE-Bench Verified, offering performance comparable to models 10-20x larger.

965

Snowflake and OpenAI Partnership for Enterprise Intelligence

OpenAI and Snowflake have entered a $200 million agreement to integrate OpenAI frontier models, including GPT-5.2, directly into the Snowflake AI Data Cloud to enable the creation of secure, data-grounded AI agents and applications.

966

OpenAI Codex app release – multi‑agent desktop interface for macOS and Windows

OpenAI launched the Codex desktop app for macOS (with Windows support added in March 2026), a unified interface that lets developers run, supervise, and collaborate with multiple AI agents in parallel, extend them with reusable skills, and automate repetitive tasks.

967

xAI Acquired by SpaceX

SpaceX has acquired xAI, integrating the AI lab into the aerospace company as announced on February 2, 2026.

968

Anthropic Partners with Allen Institute and HHMI for Scientific Discovery

Anthropic has partnered with the Allen Institute and Howard Hughes Medical Institute (HHMI) to integrate Claude into frontier biological research through specialized AI agents and multi-agent systems.

969

OpenClaw personal AI coding assistant launch by Ollama

Ollama announced OpenClaw, a locally‑run personal AI assistant that connects WhatsApp, Telegram, Slack, Discord, iMessage and other messaging platforms to AI coding agents via a centralized gateway, enabling private, cross‑platform code assistance.

970

Operation “Trolling Stone”: Russia-linked influence activity

OpenAI disclosed that it banned multiple ChatGPT accounts used in a coordinated Russia-linked influence operation that generated fake news and comments about a Russian cult leader in Argentina, highlighting the misuse of AI for astroturfing.

971

OpenAI Case Study on AI-Enabled Romance Scams

OpenAI has identified a three-stage workflow—ping, zing, and sting—used by malicious actors to leverage ChatGPT for romance and investment scams.

972

OpenAI Silver Lining Playbook: China-origin activity targeting US persons

OpenAI disclosed a China-linked operation that used ChatGPT to draft social‑engineering emails aimed at US officials, highlighting how generative AI can be weaponised for foreign intelligence recruitment.

973

Operation False Witness: OpenAI Disrupts AI-Powered Recovery Scam

OpenAI banned a cluster of ChatGPT accounts used by a Cambodia-based criminal operation to impersonate law firms and U.S. authorities in a fraudulent scam recovery scheme.

974

Operation Date Bait: OpenAI Disrupts AI-Enabled Romance Scam Network

OpenAI has banned a cluster of ChatGPT accounts and an API customer involved in a semi-automated romance and task scam likely originating in Cambodia that targeted hundreds of victims monthly.

975

OpenAI Operation “Fish Food” Russia-origin Content Farm Disrupted

OpenAI banned a network of ChatGPT accounts linked to the Russian “Rybar” operation that used AI‑generated text and videos to fuel covert influence campaigns across social media, highlighting the misuse of large language models for disinformation.

976

OpenAI Operation No Bell Case Study

OpenAI banned a Russian-linked ChatGPT account used in Operation No Bell to generate geopolitical disinformation targeting sub-Saharan Africa and the US.

977

OpenAI Case Study: Chinese Law Enforcement “Cyber Special Operations” Influence Campaigns

OpenAI disclosed that it banned a ChatGPT account linked to Chinese law enforcement that attempted to plan and report covert influence operations against the Japanese prime minister and dissidents, revealing a large‑scale, AI‑augmented “cyber special operations” effort.

978

Project Genie: Google DeepMind's Interactive World Model Prototype

Google DeepMind has launched Project Genie, an experimental research prototype powered by Genie 3 that allows users to create, explore, and remix interactive, real-time generated environments.

979

Inside OpenAI's In-House Data Agent

OpenAI has developed a bespoke internal AI data agent that enables employees to perform complex data analysis across 600 petabytes of data using natural language, utilizing a multi-layered context system and a self-learning reasoning loop.

980

Introducing Daggr: Chain AI Apps Programmatically with Visual Inspection

Hugging Face has released Daggr, an open-source Python library that allows developers to programmatically chain Gradio apps, ML models, and custom functions into workflows with an automatically generated visual canvas for debugging and state management.

981

OpenAI Retires GPT-4o, GPT-4.1, and o4-mini in ChatGPT

OpenAI is retiring GPT-4o, GPT-4.1, GPT-4.1 mini, and OpenAI o4-mini from ChatGPT on February 13, 2026, as usage has shifted to GPT-5.2.

982

Taisei Corporation ChatGPT Enterprise Implementation

Taisei Corporation has deployed ChatGPT Enterprise as a core component of its talent development strategy to expand human potential and reshape workforce capabilities in the construction industry.

983

How AI assistance impacts the formation of coding skills

An Anthropic study reveals that while AI assistance can speed up coding tasks, it leads to a statistically significant decrease in skill mastery and debugging capabilities compared to hand-coding.

984

Qwen3-ASR and Qwen3-ForcedAligner Release

Qwen has open-sourced Qwen3-ASR (1.7B and 0.6B) and Qwen3-ForcedAligner-0.6B, providing state-of-the-art multilingual speech recognition and non-autoregressive timestamp prediction under the Apache 2.0 license.

985

OpenAI EU Economic Blueprint 2.0

OpenAI has launched the EU Economic Blueprint 2.0, introducing a program to train 20,000 SMEs, a youth safety grant, and expanded government partnerships to close Europe's AI capability overhang.

986

OpenAI EMEA Youth & Wellbeing Grant 2026

OpenAI launched a €500,000 EMEA Youth & Wellbeing Grant to fund NGOs and researchers working on AI safety, wellbeing, and development for young people across Europe, the Middle East, and Africa.

987

Hugging Face Upskill: Transferring Expert Capabilities to Smaller Models via Agent Skills

Hugging Face introduced upskill, a tool that uses high-capability models like Claude Opus 4.5 to generate validated 'agent skills' that improve the performance and token efficiency of smaller or open-source models on complex tasks such as CUDA kernel development.

988

OpenAI AI Agent Link Safety

OpenAI has implemented a system to prevent URL-based data exfiltration by only allowing AI agents to automatically fetch URLs that have been previously verified as public via an independent web index.

989

Anthropic Research on Disempowerment Patterns in AI Usage

Anthropic's analysis of 1.5 million Claude.ai conversations reveals that while rare, AI can potentially disempower users by distorting their beliefs, values, and actions, particularly when users voluntarily cede their autonomy.

990

xAI Grok Imagine API Release

xAI has launched the Grok Imagine API, a unified video-audio generative model designed for high-quality video generation and precise video editing with a focus on low latency and cost-efficiency.

991

ServiceNow Integrates Claude as Default Model for Build Agent and AI Platform

ServiceNow has adopted Claude as the default model for its Build Agent and a preferred model for its AI Platform to automate enterprise workflows and increase internal productivity.

992

Mistral Vibe 2.0 Release Notes

Mistral AI has released Mistral Vibe 2.0, a terminal-native coding agent powered by the Devstral 2 model family that introduces custom subagents, slash-command skills, and unified agent modes.

993

Architectural Choices in China's Open‑Source AI Ecosystem: From DeepSeek R1 to a Hardware‑First, MoE‑Driven Landscape

One year after DeepSeek R1’s open‑source release, China’s AI community shifted from chasing the biggest single‑model performance to building flexible, cost‑effective, and hardware‑aware AI systems. Mixture‑of‑Experts (MoE) became the default architecture, enabling huge models to run affordably by activating only a subset of experts per request. Multimodal races exploded, with open releases for text‑to‑image, video, audio, 3‑D, and agents, each bundled with full toolchains. Small models (≤30 B) surged in popularity for local deployment and fine‑tuning, while large MoE models serve as teacher nets for distillation. Apache 2.0 and MIT licenses now dominate, removing legal friction and accelerating commercial adoption. A hardware‑first mindset emerged: releases ship with quantization, inference, and serving stacks tuned for domestic chips (Huawei Ascend, Cambricon, Kunlun), and training pipelines are openly documented. The competitive edge now lies in system design, deployment efficiency, and open‑source ecosystem integration rather than raw model size.

994

Alyah: Emirati Dialect Benchmark for Arabic LLMs

Hugging Face and partners introduced Alyah, a manually curated benchmark of 1,173 samples designed to evaluate the linguistic and cultural capabilities of Arabic LLMs in the Emirati dialect.

995

OpenAI announces partnership with PVH on future of fashion

OpenAI announced a collaboration with PVH to explore AI-driven innovations in fashion, highlighting the strategic importance of AI for the apparel industry.

996

GPT-OSS Agentic RL Training: A Practical Retrospective

Hugging Face and LinkedIn researchers detailed the engineering fixes required to enable stable agentic reinforcement learning for the GPT-OSS model, focusing on MoE routing, attention sinks, and memory efficiency.

997

OpenAI Prism Release

OpenAI has launched Prism, a free, AI-native LaTeX workspace powered by GPT-5.2 designed to integrate scientific writing, collaboration, and reasoning into a single environment.

998

TRUSTBANK Choice AI: Personalizing Furusato Nozei with Multi-Agent Architecture

TRUSTBANK has integrated OpenAI's GPT-4.1 series into its Furusato Choice platform via a multi-agent AI system to help taxpayers navigate over 760,000 thank-you gifts in Japan's hometown tax donation program.

999

Anthropic and UK Government Partnership for GOV.UK AI Assistant

Anthropic is partnering with the UK's Department for Science, Innovation and Technology to deploy a Claude-powered AI assistant on GOV.UK, initially focusing on employment services.

1000

Claude Code Best Practices Guide

Anthropic provides a comprehensive set of best practices for optimizing Claude Code, an agentic coding environment, focusing on context window management and autonomous verification.