✷ The archive · 11 labs · 3,050 dispatches
The labs
No more opening a dozen official blogs every morning. First-hand releases from OpenAI, Anthropic, DeepMind and the rest, each with its substance pulled out.
GPT-5.3-Codex release notes / what's new
OpenAI has released GPT-5.3-Codex, an agentic coding model that improves upon GPT-5.2-Codex in performance, reasoning, and speed, enabling it to execute complex, long-running technical tasks autonomously.
Claude Opus 4.6: LLM-Discovered Zero-Day Vulnerabilities
Anthropic's Claude Opus 4.6 can identify high-severity zero-day vulnerabilities in well-tested open source codebases by reasoning about code like a human researcher.
Quantifying Infrastructure Noise in Agentic Coding Evals
Anthropic research reveals that infrastructure resource configuration can cause score variances of up to 6 percentage points in agentic coding benchmarks, potentially masking or mimicking genuine model capability differences.
Building a C Compiler with Parallel Claude Agent Teams
Anthropic researcher Nicholas Carlini used 16 parallel Claude Opus 4.6 agents to autonomously build a 100,000-line Rust-based C compiler capable of compiling the Linux 6.9 kernel.
Mistral AI Voxtral Transcribe 2 Release
Mistral AI has released Voxtral Transcribe 2, featuring Voxtral Mini Transcribe V2 for high-efficiency batch processing and Voxtral Realtime for ultra-low latency live transcription.
OpenAI Codex App Server Architecture and Integration
OpenAI has introduced the Codex App Server, a JSON-RPC based protocol and process that exposes the Codex agent harness to various clients, enabling a consistent agent experience across IDEs, web runtimes, and terminal interfaces.
Hugging Face Community Evals
Hugging Face has introduced Community Evals, a decentralized system for reporting and aggregating model benchmark scores directly on the Hub to increase transparency and reproducibility.
Anthropic Commits to Ad-Free Experience for Claude
Anthropic has announced that Claude will remain ad-free to ensure the AI assistant acts exclusively in the users' interests and maintains a focused environment for deep thinking and work.
H Company Holo2-235B-A22B Preview Release
H Company has released Holo2-235B-A22B Preview, a UI localization model that achieves state-of-the-art performance on Screenspot-Pro and OSWorld G benchmarks.
The Future of the Global Open-Source AI Ecosystem: From DeepSeek to AI+
Hugging Face analyzes how open source has become the dominant strategy for Chinese AI organizations, shifting from isolated model breakthroughs to a scalable, integrated ecosystem of models, hardware, and infrastructure.
Training Design for Text-to-Image Models: Lessons from Ablations
The provided source material for the Hugging Face post on text-to-image model training design is unavailable due to a 429 Too Many Requests error.
OpenAI Sora Feed Philosophy
OpenAI has detailed the design principles and safety frameworks for the Sora feed, focusing on creativity-driven ranking, personalized recommendations, and multi-layered safety guardrails.
Apple Xcode 26.3 Claude Agent SDK Integration
Apple Xcode 26.3 integrates the Claude Agent SDK, enabling autonomous coding tasks, visual verification via Xcode Previews, and project-wide reasoning within the IDE.
Qwen3-Coder-Next Release: High-Efficiency Agentic Coding Model
Qwen3-Coder-Next is an open-weight model based on a hybrid attention and MoE architecture that achieves over 70% on SWE-Bench Verified, offering performance comparable to models 10-20x larger.
Snowflake and OpenAI Partnership for Enterprise Intelligence
OpenAI and Snowflake have entered a $200 million agreement to integrate OpenAI frontier models, including GPT-5.2, directly into the Snowflake AI Data Cloud to enable the creation of secure, data-grounded AI agents and applications.
OpenAI Codex app release – multi‑agent desktop interface for macOS and Windows
OpenAI launched the Codex desktop app for macOS (with Windows support added in March 2026), a unified interface that lets developers run, supervise, and collaborate with multiple AI agents in parallel, extend them with reusable skills, and automate repetitive tasks.
xAI Acquired by SpaceX
SpaceX has acquired xAI, integrating the AI lab into the aerospace company as announced on February 2, 2026.
Anthropic Partners with Allen Institute and HHMI for Scientific Discovery
Anthropic has partnered with the Allen Institute and Howard Hughes Medical Institute (HHMI) to integrate Claude into frontier biological research through specialized AI agents and multi-agent systems.
OpenClaw personal AI coding assistant launch by Ollama
Ollama announced OpenClaw, a locally‑run personal AI assistant that connects WhatsApp, Telegram, Slack, Discord, iMessage and other messaging platforms to AI coding agents via a centralized gateway, enabling private, cross‑platform code assistance.
Operation “Trolling Stone”: Russia-linked influence activity
OpenAI disclosed that it banned multiple ChatGPT accounts used in a coordinated Russia-linked influence operation that generated fake news and comments about a Russian cult leader in Argentina, highlighting the misuse of AI for astroturfing.
OpenAI Case Study on AI-Enabled Romance Scams
OpenAI has identified a three-stage workflow—ping, zing, and sting—used by malicious actors to leverage ChatGPT for romance and investment scams.
OpenAI Silver Lining Playbook: China-origin activity targeting US persons
OpenAI disclosed a China-linked operation that used ChatGPT to draft social‑engineering emails aimed at US officials, highlighting how generative AI can be weaponised for foreign intelligence recruitment.
Operation False Witness: OpenAI Disrupts AI-Powered Recovery Scam
OpenAI banned a cluster of ChatGPT accounts used by a Cambodia-based criminal operation to impersonate law firms and U.S. authorities in a fraudulent scam recovery scheme.
Operation Date Bait: OpenAI Disrupts AI-Enabled Romance Scam Network
OpenAI has banned a cluster of ChatGPT accounts and an API customer involved in a semi-automated romance and task scam likely originating in Cambodia that targeted hundreds of victims monthly.
OpenAI Operation “Fish Food” Russia-origin Content Farm Disrupted
OpenAI banned a network of ChatGPT accounts linked to the Russian “Rybar” operation that used AI‑generated text and videos to fuel covert influence campaigns across social media, highlighting the misuse of large language models for disinformation.
OpenAI Operation No Bell Case Study
OpenAI banned a Russian-linked ChatGPT account used in Operation No Bell to generate geopolitical disinformation targeting sub-Saharan Africa and the US.
OpenAI Case Study: Chinese Law Enforcement “Cyber Special Operations” Influence Campaigns
OpenAI disclosed that it banned a ChatGPT account linked to Chinese law enforcement that attempted to plan and report covert influence operations against the Japanese prime minister and dissidents, revealing a large‑scale, AI‑augmented “cyber special operations” effort.
Project Genie: Google DeepMind's Interactive World Model Prototype
Google DeepMind has launched Project Genie, an experimental research prototype powered by Genie 3 that allows users to create, explore, and remix interactive, real-time generated environments.
Inside OpenAI's In-House Data Agent
OpenAI has developed a bespoke internal AI data agent that enables employees to perform complex data analysis across 600 petabytes of data using natural language, utilizing a multi-layered context system and a self-learning reasoning loop.
Introducing Daggr: Chain AI Apps Programmatically with Visual Inspection
Hugging Face has released Daggr, an open-source Python library that allows developers to programmatically chain Gradio apps, ML models, and custom functions into workflows with an automatically generated visual canvas for debugging and state management.
OpenAI Retires GPT-4o, GPT-4.1, and o4-mini in ChatGPT
OpenAI is retiring GPT-4o, GPT-4.1, GPT-4.1 mini, and OpenAI o4-mini from ChatGPT on February 13, 2026, as usage has shifted to GPT-5.2.
Taisei Corporation ChatGPT Enterprise Implementation
Taisei Corporation has deployed ChatGPT Enterprise as a core component of its talent development strategy to expand human potential and reshape workforce capabilities in the construction industry.
How AI assistance impacts the formation of coding skills
An Anthropic study reveals that while AI assistance can speed up coding tasks, it leads to a statistically significant decrease in skill mastery and debugging capabilities compared to hand-coding.
Qwen3-ASR and Qwen3-ForcedAligner Release
Qwen has open-sourced Qwen3-ASR (1.7B and 0.6B) and Qwen3-ForcedAligner-0.6B, providing state-of-the-art multilingual speech recognition and non-autoregressive timestamp prediction under the Apache 2.0 license.
OpenAI EU Economic Blueprint 2.0
OpenAI has launched the EU Economic Blueprint 2.0, introducing a program to train 20,000 SMEs, a youth safety grant, and expanded government partnerships to close Europe's AI capability overhang.
OpenAI EMEA Youth & Wellbeing Grant 2026
OpenAI launched a €500,000 EMEA Youth & Wellbeing Grant to fund NGOs and researchers working on AI safety, wellbeing, and development for young people across Europe, the Middle East, and Africa.
Hugging Face Upskill: Transferring Expert Capabilities to Smaller Models via Agent Skills
Hugging Face introduced upskill, a tool that uses high-capability models like Claude Opus 4.5 to generate validated 'agent skills' that improve the performance and token efficiency of smaller or open-source models on complex tasks such as CUDA kernel development.
OpenAI AI Agent Link Safety
OpenAI has implemented a system to prevent URL-based data exfiltration by only allowing AI agents to automatically fetch URLs that have been previously verified as public via an independent web index.
Anthropic Research on Disempowerment Patterns in AI Usage
Anthropic's analysis of 1.5 million Claude.ai conversations reveals that while rare, AI can potentially disempower users by distorting their beliefs, values, and actions, particularly when users voluntarily cede their autonomy.
xAI Grok Imagine API Release
xAI has launched the Grok Imagine API, a unified video-audio generative model designed for high-quality video generation and precise video editing with a focus on low latency and cost-efficiency.
ServiceNow Integrates Claude as Default Model for Build Agent and AI Platform
ServiceNow has adopted Claude as the default model for its Build Agent and a preferred model for its AI Platform to automate enterprise workflows and increase internal productivity.
Mistral Vibe 2.0 Release Notes
Mistral AI has released Mistral Vibe 2.0, a terminal-native coding agent powered by the Devstral 2 model family that introduces custom subagents, slash-command skills, and unified agent modes.
Architectural Choices in China's Open‑Source AI Ecosystem: From DeepSeek R1 to a Hardware‑First, MoE‑Driven Landscape
One year after DeepSeek R1’s open‑source release, China’s AI community shifted from chasing the biggest single‑model performance to building flexible, cost‑effective, and hardware‑aware AI systems. Mixture‑of‑Experts (MoE) became the default architecture, enabling huge models to run affordably by activating only a subset of experts per request. Multimodal races exploded, with open releases for text‑to‑image, video, audio, 3‑D, and agents, each bundled with full toolchains. Small models (≤30 B) surged in popularity for local deployment and fine‑tuning, while large MoE models serve as teacher nets for distillation. Apache 2.0 and MIT licenses now dominate, removing legal friction and accelerating commercial adoption. A hardware‑first mindset emerged: releases ship with quantization, inference, and serving stacks tuned for domestic chips (Huawei Ascend, Cambricon, Kunlun), and training pipelines are openly documented. The competitive edge now lies in system design, deployment efficiency, and open‑source ecosystem integration rather than raw model size.
Alyah: Emirati Dialect Benchmark for Arabic LLMs
Hugging Face and partners introduced Alyah, a manually curated benchmark of 1,173 samples designed to evaluate the linguistic and cultural capabilities of Arabic LLMs in the Emirati dialect.
OpenAI announces partnership with PVH on future of fashion
OpenAI announced a collaboration with PVH to explore AI-driven innovations in fashion, highlighting the strategic importance of AI for the apparel industry.
GPT-OSS Agentic RL Training: A Practical Retrospective
Hugging Face and LinkedIn researchers detailed the engineering fixes required to enable stable agentic reinforcement learning for the GPT-OSS model, focusing on MoE routing, attention sinks, and memory efficiency.
OpenAI Prism Release
OpenAI has launched Prism, a free, AI-native LaTeX workspace powered by GPT-5.2 designed to integrate scientific writing, collaboration, and reasoning into a single environment.
TRUSTBANK Choice AI: Personalizing Furusato Nozei with Multi-Agent Architecture
TRUSTBANK has integrated OpenAI's GPT-4.1 series into its Furusato Choice platform via a multi-agent AI system to help taxpayers navigate over 760,000 thank-you gifts in Japan's hometown tax donation program.
Anthropic and UK Government Partnership for GOV.UK AI Assistant
Anthropic is partnering with the UK's Department for Science, Innovation and Technology to deploy a Claude-powered AI assistant on GOV.UK, initially focusing on employment services.
Claude Code Best Practices Guide
Anthropic provides a comprehensive set of best practices for optimizing Claude Code, an agentic coding environment, focusing on context window management and autonomous verification.