✷ The archive · 11 labs · 3,050 dispatches
The labs
No more opening a dozen official blogs every morning. First-hand releases from OpenAI, Anthropic, DeepMind and the rest, each with its substance pulled out.
Anthropic Responsible Scaling Policy Version 3.0
Anthropic has released Version 3.0 of its Responsible Scaling Policy (RSP), restructuring its framework to separate unilateral company commitments from industry-wide safety recommendations while introducing a Frontier Safety Roadmap and periodic Risk Reports.
OpenAI Discontinues SWE-bench Verified Evaluation
OpenAI has stopped reporting SWE-bench Verified scores because flawed test cases and training data contamination make the benchmark an unreliable measure of frontier model coding capabilities.
OpenAI Frontier Alliance Partners Announcement
OpenAI has launched the Frontier Alliance, partnering with McKinsey, BCG, BCG X, Accenture, and Capgemini to help enterprises deploy and scale AI coworkers using the Frontier platform.
Anthropic Education Report: The AI Fluency Index
Anthropic introduces the AI Fluency Index to measure how users collaborate with AI, finding that iterative refinement is the strongest predictor of high-level AI fluency.
Anthropic Persona Selection Model
Anthropic proposes the persona selection model, a theory suggesting that AI assistants behave like humans because they simulate human-like personas learned during pretraining, which post-training then refines.
Ollama 0.17: Simplified Setup for OpenClaw AI Assistant
Ollama 0.17 introduces a single-command installation process for OpenClaw, a personal AI assistant capable of managing emails, calendars, and messaging apps on local hardware.
Anthropic Report on Detecting and Preventing Distillation Attacks
Anthropic has identified and detailed industrial-scale distillation attacks by DeepSeek, Moonshot, and MiniMax, who used over 24,000 fraudulent accounts to illicitly extract Claude's capabilities.
OpenAI First Proof Submissions
OpenAI has submitted proof attempts for the First Proof research-level math challenge, with internal models potentially solving at least five of the ten problems.
Train AI models with Unsloth and Hugging Face Jobs
Hugging Face has integrated Unsloth with Hugging Face Jobs to enable fast, low-cost LLM fine-tuning, specifically optimized for small models like LiquidAI/LFM2.5-1.2B-Instruct.
GGML and llama.cpp join Hugging Face
GGML, the creators of llama.cpp, have joined Hugging Face to provide sustainable resources for local AI inference and streamline the integration between the Transformers library and local model deployment.
Anthropic Announces Claude Code Security Research Preview
Anthropic released Claude Code Security, an AI‑driven static analysis tool that scans codebases for complex vulnerabilities and suggests patches, now available in a limited research preview for enterprise and open‑source teams.
Gemini 3.1 Pro release notes
Google DeepMind unveiled Gemini 3.1 Pro, a new model with dramatically improved reasoning for complex tasks, now available in preview via the Gemini API, Vertex AI, the Gemini app, and Notebook LM.
OpenAI Grant for The Alignment Project
OpenAI has announced a $7.5 million grant to The Alignment Project, a UK AI Security Institute (UK AISI) fund designed to scale independent research into AI alignment and safety.
OpenAI for India Initiative
OpenAI has launched 'OpenAI for India,' a nationwide initiative partnering with Tata Group and other institutions to build sovereign AI infrastructure, accelerate enterprise adoption, and expand AI upskilling across the country.
IBM and UC Berkeley Diagnose Enterprise Agent Failures Using IT-Bench and MAST
IBM Research and UC Berkeley introduced MAST (Multi-Agent System Failure Taxonomy) to diagnose why enterprise IT agents fail, revealing that frontier models suffer from isolated verification errors while open models face cascading systemic collapses.
Gemini Music Generation with Lyria 3
Google DeepMind has integrated the Lyria 3 generative music model into the Gemini app, enabling users to create 30-second AI-generated tracks from text prompts or image and video uploads.
Gradio 6 gr.HTML: One-Shot Web App Development
Gradio 6 introduces enhanced gr.HTML support for custom templates, scoped CSS, and JavaScript interactivity, enabling the creation of complex web components within a single Python file.
OpenAI Introducing EVMbench
OpenAI and Paradigm have released EVMbench, a benchmark designed to evaluate AI agents' ability to detect, patch, and exploit high-severity smart contract vulnerabilities.
Anthropic Measures AI Agent Autonomy in Practice – Key Findings and Implications
Anthropic’s February 2026 study reveals that Claude Code agents are running autonomously for longer, experienced users grant more auto‑approval yet interrupt more, agents pause for clarification more than humans interrupt, and risky‑domain usage remains limited but emerging.
Google DeepMind National Partnerships for AI: India Expansion
Google DeepMind is establishing new partnerships with Indian government bodies and institutions to broaden access to frontier AI models for science, education, agriculture, and energy security.
Anthropic and the Government of Rwanda MOU for AI Integration
Anthropic and the Government of Rwanda have signed a three-year Memorandum of Understanding to integrate AI into Rwanda's health, education, and public sector systems.
Anthropic and Infosys Collaboration for Enterprise AI Agents
Anthropic and Infosys have partnered to integrate Claude models and Claude Code with Infosys Topaz to build agentic AI solutions for regulated industries including telecommunications, financial services, and manufacturing.
Anthropic Economic Index: India Country Brief
India ranks second globally in total Claude.ai usage but shows high concentration in the IT sector and a significant gap in per-capita adoption, while delivering a 15x productivity speedup on complex tasks.
Ollama adds Subagents and Web Search to Claude Code
Ollama now supports subagents and web search within Claude Code, allowing models to run parallel tasks and access real-time information without requiring MCP servers or API keys.
Anthropic Expands in India with Bengaluru Office and Strategic Partnerships
Anthropic has opened a new office in Bengaluru and launched partnerships across enterprise, education, and agriculture to scale responsible AI and improve Indic language capabilities.
Qwen 3.5‑397B‑A17B release: hybrid linear‑attention MoE model with 1 M token context and state‑of‑the‑art multimodal performance
Qwen 3.5‑397B‑A17B is a 397 billion‑parameter multimodal model that activates only 17 billion parameters per token, delivering state‑of‑the‑art performance on language, coding, reasoning and vision tasks while being up to 19× faster than its predecessor.
GPT-5.2 Derives New Result in Theoretical Physics
OpenAI has announced that GPT-5.2 Pro and a scaffolded version of GPT-5.2 were used to derive a new result in theoretical physics regarding non-zero gluon tree amplitudes in the half-collinear regime.
ChatGPT Lockdown Mode and Elevated Risk Labels
OpenAI has introduced Lockdown Mode and Elevated Risk labels to mitigate prompt injection attacks and provide users with greater control over data exfiltration risks.
OpenAI Scaling Access to Codex and Sora via Hybrid Credit System
OpenAI has implemented a real-time access engine that combines rate limits with a purchasable credit system to prevent hard stops for Codex and Sora users.
OpenAI GABRIEL: Scaling Social Science Research with GPT
OpenAI has released GABRIEL, an open-source Python toolkit that uses GPT to transform unstructured text and images into quantitative measurements for social science research.
Hugging Face CUDA Kernels Agent Skill
Hugging Face has introduced an agent skill that enables coding agents like Claude and Codex to write, benchmark, and integrate production-ready CUDA kernels for transformers and diffusers libraries.
Anthropic Appoints Chris Liddell to Board of Directors
Anthropic has appointed Chris Liddell, a former CFO of Microsoft and General Motors and former Deputy White House Chief of Staff, to its Board of Directors to strengthen governance and oversight of AI's societal impact.
Anthropic and CodePath Partnership for AI-Integrated Computer Science Education
Anthropic is partnering with CodePath to integrate Claude and Claude Code into the curricula of the US's largest collegiate computer science program, providing over 20,000 students at community colleges, state schools, and HBCUs with access to frontier AI tools.
Gemini 3 Deep Think Update
Google DeepMind has updated Gemini 3 Deep Think, a specialized reasoning mode that achieves gold-medal level performance in mathematics, physics, and chemistry olympiads and is now available via the Gemini API for select users.
GPT-5.3-Codex-Spark Release Notes
OpenAI has released GPT-5.3-Codex-Spark, a small, low-latency model optimized for real-time coding collaboration and delivering over 1,000 tokens per second on Cerebras hardware.
OpenEnv: Evaluating Tool-Using Agents in Real-World Environments
Hugging Face and Meta introduce OpenEnv, an open-source framework that evaluates AI agents against real systems and production-grade environments like the Calendar Gym to bridge the gap between research and production reliability.
Anthropic Donates $20 Million to Public First Action for AI Governance
Anthropic has donated $20 million to Public First Action, a bipartisan 501(c)(4) organization dedicated to promoting AI safeguards, public education, and American leadership in AI policy.
Anthropic Series G Funding and Enterprise Growth
Anthropic has raised $30 billion in Series G funding at a $380 billion post-money valuation to expand its frontier research, product development, and infrastructure.
OpenAI Harness Engineering: Leveraging Codex for Zero-Manual-Code Development
OpenAI developed an internal software product with zero lines of manually-written code using Codex, reducing development time by approximately 10x through a shift from manual coding to environment design and agent orchestration.
Anthropic Commitments to Cover Data Center Electricity Price Increases
Anthropic has committed to covering the costs of grid upgrades and demand-driven electricity price increases resulting from its data center operations to protect American ratepayers.
Qwen-Image-2.0 Release: Professional Infographics and Photorealism
Qwen-Image-2.0 is a unified image generation and editing model that supports 1k-token instructions for professional infographics and native 2K resolution for high-fidelity photorealism.
Gemini Deep Think enables autonomous research across mathematics, physics, and computer science
Gemini Deep Think powers autonomous and collaborative agents that solve research‑level math, physics, and computer‑science problems, producing publishable results and resolving open conjectures.
OpenAI Integrates ChatGPT into GenAI.mil for Department of War
OpenAI is deploying a custom version of ChatGPT to GenAI.mil, providing 3 million Department of War personnel with secure, unclassified generative AI capabilities for administrative and operational support.
Transformers.js v4 release notes / what's new
Hugging Face has released Transformers.js v4, introducing a new C++ rewritten WebGPU runtime for hardware acceleration across browsers and server-side runtimes, alongside a standalone tokenizers library.
OpenAI Localization Approach and OpenAI for Countries Initiative
OpenAI has introduced a framework for localized AI systems through the OpenAI for Countries initiative, allowing nations to adapt frontier models to local languages, laws, and cultural norms while adhering to global safety red-lines.
SyGra 2.0.0 Studio Release
SyGra 2.0.0 introduces Studio, a visual interactive environment for designing and executing synthetic data generation workflows without needing to manually edit YAML files.
GPT-5 Lowers the Cost of Cell-Free Protein Synthesis
OpenAI and Ginkgo Bioworks used GPT-5 in a closed-loop autonomous lab system to reduce cell-free protein synthesis costs by 40% and reagent costs by 57%.
OpenAI Trusted Access for Cyber
OpenAI has launched Trusted Access for Cyber, an identity-based framework to provide security professionals with prioritized access to GPT-5.3-Codex to accelerate cyber defense and vulnerability remediation.
OpenAI Frontier Platform Release
OpenAI has introduced Frontier, an end-to-end platform designed to help enterprises build, deploy, and manage AI agents with shared business context, secure execution environments, and integrated governance.
GPT-5.3-Codex System Card
OpenAI has released GPT-5.3-Codex, an agentic coding model that integrates GPT-5.2-Codex performance with GPT-5.2 reasoning to handle complex, long-running technical tasks.