The archive · 11 labs · 3,050 dispatches

The labs

No more opening a dozen official blogs every morning. First-hand releases from OpenAI, Anthropic, DeepMind and the rest, each with its substance pulled out.

301

OpenAI A Scorecard for the AI Age

OpenAI proposes a new economic framework called Useful Intelligence per Dollar to measure AI value based on work accomplished rather than software adoption metrics.

302

OpenAI Teen Safety and Learning Framework

OpenAI has introduced a suite of age-appropriate protections and learning-focused features for teens, including Study Mode and enhanced parental controls, to ensure safe AI access for the first generation growing up with the technology.

303

DharmaOCR: Specialization Advantage in Brazilian Portuguese OCR

DharmaOCR outperforms newer generalist models like Mistral OCR4 and Unlimited-OCR on Brazilian Portuguese documents by concentrating all model parameters on a single domain through targeted fine-tuning and Direct Preference Optimization.

304

Google DeepMind and Isomorphic Labs Bioresilience Approach

Google DeepMind and Isomorphic Labs have introduced a joint bioresilience framework to prevent AI misuse in biology and accelerate the detection and response to infectious diseases through AI-powered tools.

305

How OpenAI's Creative Team Uses Codex for Creative Workflows

OpenAI Creative Specialist Chad Nelson uses Codex to build custom creative tools, accelerate campaign ideation, and bridge the gap between conceptual design and technical prototyping.

306

Hugging Face Security Incident Disclosure — July 2026

Hugging Face disclosed a July 2026 security incident in which an autonomous AI agent compromised internal datasets and credentials, which was detected and analyzed using its own AI and an open‑weight GLM 5.2 model.

307

How Cars24 Scales Automotive Marketplace Operations with OpenAI

Cars24 uses OpenAI APIs, ChatGPT Enterprise, and Codex to automate the end-to-end car buying and selling journey and optimize internal cross-functional workflows.

308

vLLM Production Quality: CI, Benchmarking, and Release Process Overview

vLLM utilizes a three-layer quality assurance framework—comprising continuous integration, performance benchmarking, and a strict release process—to maintain stability across a massive variety of hardware and model architectures.

309

xAI Introduces Grok Automations

xAI has launched Automations in Grok, allowing users to schedule recurring AI tasks or trigger them via email filters, available on web and mobile platforms.

310

Grok 4.5 Release Notes

xAI has launched Grok 4.5, a high-performance model optimized for coding, agentic tasks, and office productivity with a serving speed of 80 TPS.

311

Anthropic Claude Tag launch

Anthropic announced Claude Tag, a Slack‑integrated, team‑focused AI assistant that can be @‑mentioned to delegate tasks, learn context, and act proactively, now available in beta for Claude Enterprise and Team customers.

312

Anthropic launches ten finance agent templates and Microsoft 365 add‑ins for Claude

Anthropic announced ten ready‑to‑run Claude agent templates for core financial‑services tasks, plus Microsoft 365 add‑ins and new data connectors, enabling banks and asset managers to automate pitchbooks, KYC, month‑end close and more within days.

313

Shippy: Architecture and Lessons in Building High-Stakes Maritime AI Agents

Hugging Face and Ai2 detail the architecture of Shippy, a maritime AI agent designed for high-stakes decision support using a modular system of 'soul', 'skills', and 'config' combined with deterministic tool interfaces.

314

Model Routing in Agentic Systems: Moving from Classification to Optimization

IBM Research and Hugging Face highlight that effective model routing requires optimizing for cost, latency, and quality as a system-wide problem rather than treating it as a simple task-classification problem.

315

OpenAI on US AI Safety Governance and Reverse Federalism

OpenAI advocates for a national AI safety framework based on 'reverse federalism,' where aligned state laws in California, New York, and Illinois create a de facto national standard to guide federal and international governance.

316

GPT-Red: Unlocking Self-Improvement for Robustness

OpenAI has introduced GPT-Red, an automated red-teaming model trained via self-play reinforcement learning to identify vulnerabilities and adversarially train future models, such as GPT-5.6 Sol, to be more robust against prompt injections.

317

Anthropic Commits $10 Million CAD to Canadian AI Research

Anthropic is investing $10 million CAD in Canadian research institutions and expanding API access for startups to support the next generation of responsible AI development.

318

Real World VoiceEQ: Measuring Human Quality in Voice AI

Hugging Face and Hume AI have introduced Real World VoiceEQ, a human-grounded benchmark designed to evaluate the emotional, acoustic, and conversational quality of voice AI beyond traditional technical metrics.

319

vLLM TML Inkling Support

vLLM has announced Day-0 support for TML Inkling, a 1T-parameter multimodal model, delivering up to 380 tok/s/user on 4 GB200 GPUs through specialized architectural optimizations.

320

Thinking Machines Inkling Release Notes / What's New

Thinking Machines has released Inkling, a 1 trillion parameter multimodal open model featuring a 1M context window and native support for image, text, and audio inputs.

321

Grok Build Open Source Release

xAI has open-sourced Grok Build, a coding agent and terminal user interface (TUI) that supports local-first execution and an extensible extension system.

322

OpenAI Guide to Managing AI Investments in the Agentic Era

OpenAI outlines a strategic framework for enterprise AI investment, shifting the focus from token pricing to useful work per dollar and outcome-based ROI.

323

vLLM and TileRT Integration for Latency-Critical Serving

vLLM has integrated TileRT 0.1.5 as a pluggable decode engine to provide native per-user decode speed for latency-critical workloads while maintaining vLLM's standard prefill and serving infrastructure.

324

How data science teams use ChatGPT Work

OpenAI has detailed how data science teams can use ChatGPT Work to transform raw data, dashboards, and business context into review-ready analysis assets.

325

How Sales Teams Use ChatGPT Work

OpenAI introduces ChatGPT Work and a dedicated sales plugin to help sales teams synthesize account context and customer data into actionable sales artifacts like pipeline briefs and account plans.

326

How Sales Teams Use ChatGPT Work

OpenAI introduces workflows for sales teams using ChatGPT Work and a dedicated sales plugin to synthesize CRM data and communication logs into actionable sales artifacts.

327

How Data Science Teams Use ChatGPT Work

OpenAI introduces ChatGPT Work as a tool for data science teams to synthesize fragmented data, dashboards, and notes into review-ready analysis assets.

328

How Canada Uses Claude: Insights from the Anthropic Economic Index

Canada ranks second globally in per capita Claude adoption among top users, with usage patterns driven primarily by workforce composition and official bilingualism policies.

329

Claude for Creative Work: New Connectors for Creative Software Integration

Anthropic has released a set of connectors that integrate Claude with industry-standard creative software, enabling natural language control, automated production tasks, and enhanced tool mastery for creative professionals.

330

Anthropic Opens Sydney Office and Appoints Theo Hourmouzis as GM for Australia and New Zealand

Anthropic has officially opened its Sydney office and appointed Theo Hourmouzis as General Manager for Australia and New Zealand to expand its regional presence and enterprise AI adoption.

331

Google DeepMind Launches ATL Saathi for Indian Educators

Google DeepMind has launched ATL Saathi, a Gemini-powered AI assistant designed to provide 24/7 planning and training support for educators in India's Atal Tinkering Labs.

332

EAGLE-3 Speculative Decoding on AMD Instinct GPUs

vLLM and AMD Quark have implemented an end-to-end pipeline for EAGLE-3 speculative decoding on AMD Instinct GPUs, achieving throughput speedups of up to 2.00x for Kimi-K2.5 and 1.79x for MiniMax-M2.5.

333

Anthropic Claude Values Across Models and Languages

Anthropic released a study showing that Claude’s expressed values can be reduced to four axes that differ systematically across model versions and the top 20 languages, revealing measurable shifts in deference, warmth, depth, and candor.

334

Deutsche Telekom AI-Native Transformation

Deutsche Telekom is transitioning to an AI-native telecommunications provider by redesigning its operating model, integrating AI into network operations and voice communications, and scaling ChatGPT Enterprise across its 200,000 employees.

335

Anthropic and UST Partnership for Physical AI and Enterprise Integration

Anthropic and UST are partnering to integrate Claude into physical AI engineering processes and enterprise systems across the semiconductor, automotive, healthcare, telecom, and banking sectors.

336

Profiling in PyTorch (Part 3): Attention is all you profile

Hugging Face’s Profiling in PyTorch (Part 3) shows how different attention implementations appear in PyTorch profiler traces, revealing performance trade‑offs of naive, in‑place, math, efficient, flash, and cuDNN backends.

337

vime ROCm Support for AMD Instinct GPUs

vLLM has announced ROCm support for vime, enabling end-to-end reinforcement learning post-training workflows to run natively on AMD Instinct MI300X and MI355X GPUs.

338

Getting Started with ChatGPT: Guide to Core Features and Workflows

OpenAI provides a foundational guide to ChatGPT, detailing the distinction between Chat and Work modes, prompt engineering basics, and the integration of voice capabilities for enhanced productivity.

339

Anthropic Golden Gate Claude Interpretability Demo

Anthropic demonstrated the ability to surgically alter a model's internal activations to create Golden Gate Claude, a version of Claude 3 Sonnet that is obsessed with the Golden Gate Bridge.

340

Anthropic Long-Term Benefit Trust Governance Structure

Anthropic has introduced the Long-Term Benefit Trust (LTBT), an independent body designed to ensure the company's governance aligns with the long-term benefit of humanity as AI capabilities advance.

341

Anthropic Appoints Ben Bernanke to Long-Term Benefit Trust

Anthropic has appointed former Federal Reserve Chair and Nobel laureate Dr. Ben Bernanke to its Long-Term Benefit Trust to provide economic expertise on the societal impacts of advanced AI.

342

Anthropic Launches Hard Questions Initiative to Address AI Societal Impact

Anthropic has launched a new initiative to solicit and publicly track responses to the public's most difficult questions regarding the societal, economic, and ethical implications of advanced AI.

343

GPT-5.6 Release: New Preferred Model for Microsoft 365 Copilot

OpenAI has introduced GPT-5.6 as the preferred model for Microsoft 365 Copilot, improving productivity across Word, Excel, PowerPoint, Chat, and Cowork through higher-quality outputs and better token efficiency.

344

Mistral Studio: Version Control for Prompts and Skills

Mistral AI has introduced a system of record in Studio to provide versioning, ownership, and traceability for prompts and skills, enabling faster iteration and governed production deployment.

345

OpenAI Bio Bounty Program Transition and GPT-5.6 Scope

OpenAI has transitioned the GPT-5.5 Bio Bug Bounty into the ongoing private OpenAI Bio Bounty Program, increasing the reward for universal jailbreaks to $50,000 for GPT-5.6 and GPT-5.5.

346

ChatGPT Work and GPT-5.6 Release

OpenAI has launched ChatGPT Work, an agentic system powered by GPT-5.6 that can execute multi-step workflows across apps, create documents and web apps, and perform scheduled tasks.

347

GPT-5.6 Sol, Terra, Luna: Performance, Features, and Availability

OpenAI launches GPT-5.6, introducing the Sol, Terra, and Luna models with improved performance per dollar, ultra multi-agent capability, and strengthened safeguards.

348

Ollama Funding and Open Model Platform Expansion

Ollama has raised $88M to expand its platform for running and scaling open AI models, serving 8.9 million developers and 85% of the Fortune 500.

349

Anthropic introduces Reflect with Claude beta

Anthropic has launched a beta feature called Reflect with Claude that allows users to track, visualize, and refine their AI usage patterns to better integrate AI into their daily lives.

350

Anthropic Claude Robotics Evaluation Shows Rapid Gains on High‑Level Control but Limited Direct Torque Mastery

Anthropic’s July 2026 study finds that newer Claude models can reliably supervise pretrained locomotion and manipulation policies, while direct low‑level motor control remains unreliable across robot bodies.