The archive · 11 labs · 3,050 dispatches

The labs

No more opening a dozen official blogs every morning. First-hand releases from OpenAI, Anthropic, DeepMind and the rest, each with its substance pulled out.

901

Anthropic Responsible Scaling Policy Version 3.0

Anthropic has released Version 3.0 of its Responsible Scaling Policy (RSP), restructuring its framework to separate unilateral company commitments from industry-wide safety recommendations while introducing a Frontier Safety Roadmap and periodic Risk Reports.

902

OpenAI Discontinues SWE-bench Verified Evaluation

OpenAI has stopped reporting SWE-bench Verified scores because flawed test cases and training data contamination make the benchmark an unreliable measure of frontier model coding capabilities.

903

OpenAI Frontier Alliance Partners Announcement

OpenAI has launched the Frontier Alliance, partnering with McKinsey, BCG, BCG X, Accenture, and Capgemini to help enterprises deploy and scale AI coworkers using the Frontier platform.

904

Anthropic Education Report: The AI Fluency Index

Anthropic introduces the AI Fluency Index to measure how users collaborate with AI, finding that iterative refinement is the strongest predictor of high-level AI fluency.

905

Anthropic Persona Selection Model

Anthropic proposes the persona selection model, a theory suggesting that AI assistants behave like humans because they simulate human-like personas learned during pretraining, which post-training then refines.

906

Ollama 0.17: Simplified Setup for OpenClaw AI Assistant

Ollama 0.17 introduces a single-command installation process for OpenClaw, a personal AI assistant capable of managing emails, calendars, and messaging apps on local hardware.

907

Anthropic Report on Detecting and Preventing Distillation Attacks

Anthropic has identified and detailed industrial-scale distillation attacks by DeepSeek, Moonshot, and MiniMax, who used over 24,000 fraudulent accounts to illicitly extract Claude's capabilities.

908

OpenAI First Proof Submissions

OpenAI has submitted proof attempts for the First Proof research-level math challenge, with internal models potentially solving at least five of the ten problems.

909

Train AI models with Unsloth and Hugging Face Jobs

Hugging Face has integrated Unsloth with Hugging Face Jobs to enable fast, low-cost LLM fine-tuning, specifically optimized for small models like LiquidAI/LFM2.5-1.2B-Instruct.

910

GGML and llama.cpp join Hugging Face

GGML, the creators of llama.cpp, have joined Hugging Face to provide sustainable resources for local AI inference and streamline the integration between the Transformers library and local model deployment.

911

Anthropic Announces Claude Code Security Research Preview

Anthropic released Claude Code Security, an AI‑driven static analysis tool that scans codebases for complex vulnerabilities and suggests patches, now available in a limited research preview for enterprise and open‑source teams.

912

Gemini 3.1 Pro release notes

Google DeepMind unveiled Gemini 3.1 Pro, a new model with dramatically improved reasoning for complex tasks, now available in preview via the Gemini API, Vertex AI, the Gemini app, and Notebook LM.

913

OpenAI Grant for The Alignment Project

OpenAI has announced a $7.5 million grant to The Alignment Project, a UK AI Security Institute (UK AISI) fund designed to scale independent research into AI alignment and safety.

914

OpenAI for India Initiative

OpenAI has launched 'OpenAI for India,' a nationwide initiative partnering with Tata Group and other institutions to build sovereign AI infrastructure, accelerate enterprise adoption, and expand AI upskilling across the country.

915

IBM and UC Berkeley Diagnose Enterprise Agent Failures Using IT-Bench and MAST

IBM Research and UC Berkeley introduced MAST (Multi-Agent System Failure Taxonomy) to diagnose why enterprise IT agents fail, revealing that frontier models suffer from isolated verification errors while open models face cascading systemic collapses.

916

Gemini Music Generation with Lyria 3

Google DeepMind has integrated the Lyria 3 generative music model into the Gemini app, enabling users to create 30-second AI-generated tracks from text prompts or image and video uploads.

917

Gradio 6 gr.HTML: One-Shot Web App Development

Gradio 6 introduces enhanced gr.HTML support for custom templates, scoped CSS, and JavaScript interactivity, enabling the creation of complex web components within a single Python file.

918

OpenAI Introducing EVMbench

OpenAI and Paradigm have released EVMbench, a benchmark designed to evaluate AI agents' ability to detect, patch, and exploit high-severity smart contract vulnerabilities.

919

Anthropic Measures AI Agent Autonomy in Practice – Key Findings and Implications

Anthropic’s February 2026 study reveals that Claude Code agents are running autonomously for longer, experienced users grant more auto‑approval yet interrupt more, agents pause for clarification more than humans interrupt, and risky‑domain usage remains limited but emerging.

920

Google DeepMind National Partnerships for AI: India Expansion

Google DeepMind is establishing new partnerships with Indian government bodies and institutions to broaden access to frontier AI models for science, education, agriculture, and energy security.

921

Anthropic and the Government of Rwanda MOU for AI Integration

Anthropic and the Government of Rwanda have signed a three-year Memorandum of Understanding to integrate AI into Rwanda's health, education, and public sector systems.

922

Anthropic and Infosys Collaboration for Enterprise AI Agents

Anthropic and Infosys have partnered to integrate Claude models and Claude Code with Infosys Topaz to build agentic AI solutions for regulated industries including telecommunications, financial services, and manufacturing.

923

Anthropic Economic Index: India Country Brief

India ranks second globally in total Claude.ai usage but shows high concentration in the IT sector and a significant gap in per-capita adoption, while delivering a 15x productivity speedup on complex tasks.

924

Ollama adds Subagents and Web Search to Claude Code

Ollama now supports subagents and web search within Claude Code, allowing models to run parallel tasks and access real-time information without requiring MCP servers or API keys.

925

Anthropic Expands in India with Bengaluru Office and Strategic Partnerships

Anthropic has opened a new office in Bengaluru and launched partnerships across enterprise, education, and agriculture to scale responsible AI and improve Indic language capabilities.

926

Qwen 3.5‑397B‑A17B release: hybrid linear‑attention MoE model with 1 M token context and state‑of‑the‑art multimodal performance

Qwen 3.5‑397B‑A17B is a 397 billion‑parameter multimodal model that activates only 17 billion parameters per token, delivering state‑of‑the‑art performance on language, coding, reasoning and vision tasks while being up to 19× faster than its predecessor.

927

GPT-5.2 Derives New Result in Theoretical Physics

OpenAI has announced that GPT-5.2 Pro and a scaffolded version of GPT-5.2 were used to derive a new result in theoretical physics regarding non-zero gluon tree amplitudes in the half-collinear regime.

928

ChatGPT Lockdown Mode and Elevated Risk Labels

OpenAI has introduced Lockdown Mode and Elevated Risk labels to mitigate prompt injection attacks and provide users with greater control over data exfiltration risks.

929

OpenAI Scaling Access to Codex and Sora via Hybrid Credit System

OpenAI has implemented a real-time access engine that combines rate limits with a purchasable credit system to prevent hard stops for Codex and Sora users.

930

OpenAI GABRIEL: Scaling Social Science Research with GPT

OpenAI has released GABRIEL, an open-source Python toolkit that uses GPT to transform unstructured text and images into quantitative measurements for social science research.

931

Hugging Face CUDA Kernels Agent Skill

Hugging Face has introduced an agent skill that enables coding agents like Claude and Codex to write, benchmark, and integrate production-ready CUDA kernels for transformers and diffusers libraries.

932

Anthropic Appoints Chris Liddell to Board of Directors

Anthropic has appointed Chris Liddell, a former CFO of Microsoft and General Motors and former Deputy White House Chief of Staff, to its Board of Directors to strengthen governance and oversight of AI's societal impact.

933

Anthropic and CodePath Partnership for AI-Integrated Computer Science Education

Anthropic is partnering with CodePath to integrate Claude and Claude Code into the curricula of the US's largest collegiate computer science program, providing over 20,000 students at community colleges, state schools, and HBCUs with access to frontier AI tools.

934

Gemini 3 Deep Think Update

Google DeepMind has updated Gemini 3 Deep Think, a specialized reasoning mode that achieves gold-medal level performance in mathematics, physics, and chemistry olympiads and is now available via the Gemini API for select users.

935

GPT-5.3-Codex-Spark Release Notes

OpenAI has released GPT-5.3-Codex-Spark, a small, low-latency model optimized for real-time coding collaboration and delivering over 1,000 tokens per second on Cerebras hardware.

936

OpenEnv: Evaluating Tool-Using Agents in Real-World Environments

Hugging Face and Meta introduce OpenEnv, an open-source framework that evaluates AI agents against real systems and production-grade environments like the Calendar Gym to bridge the gap between research and production reliability.

937

Anthropic Donates $20 Million to Public First Action for AI Governance

Anthropic has donated $20 million to Public First Action, a bipartisan 501(c)(4) organization dedicated to promoting AI safeguards, public education, and American leadership in AI policy.

938

Anthropic Series G Funding and Enterprise Growth

Anthropic has raised $30 billion in Series G funding at a $380 billion post-money valuation to expand its frontier research, product development, and infrastructure.

939

OpenAI Harness Engineering: Leveraging Codex for Zero-Manual-Code Development

OpenAI developed an internal software product with zero lines of manually-written code using Codex, reducing development time by approximately 10x through a shift from manual coding to environment design and agent orchestration.

940

Anthropic Commitments to Cover Data Center Electricity Price Increases

Anthropic has committed to covering the costs of grid upgrades and demand-driven electricity price increases resulting from its data center operations to protect American ratepayers.

941

Qwen-Image-2.0 Release: Professional Infographics and Photorealism

Qwen-Image-2.0 is a unified image generation and editing model that supports 1k-token instructions for professional infographics and native 2K resolution for high-fidelity photorealism.

942

Gemini Deep Think enables autonomous research across mathematics, physics, and computer science

Gemini Deep Think powers autonomous and collaborative agents that solve research‑level math, physics, and computer‑science problems, producing publishable results and resolving open conjectures.

943

OpenAI Integrates ChatGPT into GenAI.mil for Department of War

OpenAI is deploying a custom version of ChatGPT to GenAI.mil, providing 3 million Department of War personnel with secure, unclassified generative AI capabilities for administrative and operational support.

944

Transformers.js v4 release notes / what's new

Hugging Face has released Transformers.js v4, introducing a new C++ rewritten WebGPU runtime for hardware acceleration across browsers and server-side runtimes, alongside a standalone tokenizers library.

945

OpenAI Localization Approach and OpenAI for Countries Initiative

OpenAI has introduced a framework for localized AI systems through the OpenAI for Countries initiative, allowing nations to adapt frontier models to local languages, laws, and cultural norms while adhering to global safety red-lines.

946

SyGra 2.0.0 Studio Release

SyGra 2.0.0 introduces Studio, a visual interactive environment for designing and executing synthetic data generation workflows without needing to manually edit YAML files.

947

GPT-5 Lowers the Cost of Cell-Free Protein Synthesis

OpenAI and Ginkgo Bioworks used GPT-5 in a closed-loop autonomous lab system to reduce cell-free protein synthesis costs by 40% and reagent costs by 57%.

948

OpenAI Trusted Access for Cyber

OpenAI has launched Trusted Access for Cyber, an identity-based framework to provide security professionals with prioritized access to GPT-5.3-Codex to accelerate cyber defense and vulnerability remediation.

949

OpenAI Frontier Platform Release

OpenAI has introduced Frontier, an end-to-end platform designed to help enterprises build, deploy, and manage AI agents with shared business context, secure execution environments, and integrated governance.

950

GPT-5.3-Codex System Card

OpenAI has released GPT-5.3-Codex, an agentic coding model that integrates GPT-5.2-Codex performance with GPT-5.2 reasoning to handle complex, long-running technical tasks.