The archive · 11 labs · 3,056 dispatches

The labs

No more opening a dozen official blogs every morning. First-hand releases from OpenAI, Anthropic, DeepMind and the rest, each with its substance pulled out.

1451

OpenAI Supporting Nonprofit and Community Innovation Fund

OpenAI is opening applications for a $50 million fund to provide unrestricted grants to U.S.-based 501(c)(3) nonprofits and community organizations to foster AI-driven innovation for the public good.

1452

Grok Code Fast 1 release notes / what's new

xAI has released grok-code-fast-1, a new model architecture optimized for speed, cost-efficiency, and agentic coding workflows across multiple programming languages.

1453

OpenAI Collective Alignment and Model Spec Updates

OpenAI has introduced a collective alignment research effort to integrate public preferences from over 1,000 global participants into its Model Spec to ensure AI default behaviors reflect diverse human values.

1454

OpenAI and Anthropic Joint Safety Evaluation Findings

OpenAI and Anthropic released joint safety evaluation results showing Claude 4 models excel at instruction hierarchy while OpenAI's reasoning models lead on jailbreak resistance, with both labs highlighting the value of cross‑lab collaboration for AI alignment.

1455

Anthropic Education Report: How Educators Use Claude

Anthropic's analysis of 74,000 conversations reveals that higher education professionals primarily use Claude for curriculum development and administrative automation, while increasingly leveraging Artifacts to build custom interactive learning tools.

1456

Anthropic National Security and Public Sector Advisory Council

Anthropic has formed a bipartisan National Security and Public Sector Advisory Council to help the U.S. government and allied democracies maintain technological advantages in AI for national security.

1457

Anthropic Threat Intelligence Report August 2025

Anthropic's August 2025 report details how malicious actors are weaponizing agentic AI to automate cyberattacks, scale fraudulent employment, and develop ransomware with minimal technical skill.

1458

Anthropic Education Report: How Educators Use Claude

Anthropic's analysis of 74,000 conversations reveals that university educators primarily use Claude for curriculum development and administrative automation, while increasingly using Artifacts to build custom interactive learning tools.

1459

OpenAI Mental Health and Crisis Intervention Safeguards

OpenAI has detailed its layered safety stack for managing mental health crises in ChatGPT, including the introduction of GPT-5's improved response accuracy and plans for direct emergency service integration.

1460

OpenAI Learning Accelerator India Launch

OpenAI has launched the Learning Accelerator in India to provide AI research, training, and ChatGPT licenses to millions of educators and students through strategic institutional partnerships.

1461

OpenAI and Retro Biosciences: Accelerating Life Sciences with GPT-4b micro

OpenAI and Retro Biosciences developed GPT-4b micro, a specialized protein engineering model that re-engineered Yamanaka factors to achieve over 50-fold higher stem cell reprogramming expression than wild-type controls.

1462

Blue J Scaling Tax Research with GPT-4.1

Blue J leverages GPT-4.1 and a proprietary RAG system to automate complex tax research across the US, Canada, and the UK, reducing the disagree rate to fewer than 1 in 700 responses.

1463

Anthropic and NNSA Develop AI Nuclear Safeguards Classifier

Anthropic partnered with the U.S. Department of Energy's National Nuclear Security Administration to develop an AI classifier that identifies nuclear proliferation risks with 96% accuracy in preliminary testing.

1464

Anthropic and NNSA Develop AI Nuclear Safeguards Classifier

Anthropic has partnered with the U.S. Department of Energy's National Nuclear Security Administration to develop an AI classifier that identifies nuclear proliferation risks with 96% accuracy.

1465

DeepSeek-V3.1 Release Notes

DeepSeek-V3.1 introduces hybrid inference modes, improved agentic capabilities, and continued pretraining for long context extension, marking a step toward agent-centric AI.

1466

Anthropic Higher Education Initiatives: Advisory Board and AI Fluency Courses

Anthropic has launched a Higher Education Advisory Board and three Creative Commons AI Fluency courses to guide the responsible integration of Claude into teaching, learning, and research.

1467

NVIDIA Nemotron Post-Training Dataset v2 and Nemotron Nano 2 9B Release

NVIDIA has released a 6-million sample multilingual reasoning dataset and the Nemotron Nano 2 9B model, which utilizes a hybrid Transformer-Mamba architecture to optimize reasoning costs and throughput.

1468

MIXI ChatGPT Enterprise Deployment

MIXI deployed ChatGPT Enterprise across its organization in 45 days, reducing work hours by over 90% in some projects and enabling employees to create over 1,600 custom GPTs.

1469

Claude Code and Admin Controls for Team and Enterprise Plans

Anthropic has introduced premium seats for Team and Enterprise plans that bundle Claude and Claude Code, alongside a new Compliance API for programmatic usage monitoring.

1470

Generate Images with Claude and Hugging Face

Hugging Face enables image generation within Claude by connecting the AI to Hugging Face Spaces via the Model Context Protocol (MCP) server.

1471

Qwen-Image-Edit Release Notes

Qwen-Image-Edit is a 20B parameter image editing model that enables precise semantic and appearance editing, including bilingual text modification, by leveraging Qwen2.5-VL and a VAE Encoder.

1472

Hugging Face MCP for Research: Connecting AI to Research Tools

Hugging Face introduces the Research Tracker MCP, enabling AI agents to automate research discovery by integrating arXiv, GitHub, and Hugging Face through the Model Context Protocol.

1473

Hugging Face kernel-builder: A Guide to Building and Scaling Production-Ready CUDA Kernels

Hugging Face introduces the kernel-builder library to simplify the development, multi-architecture compilation, and distribution of production-ready CUDA kernels via the Hugging Face Hub.

1474

DoorDash AI Implementation Strategy

DoorDash is leveraging AI to democratize technical creation for non-engineers and personalize employee development and performance management.

1475

Claude Opus 4 and 4.1 Conversation-Ending Capability

Anthropic has introduced a feature allowing Claude Opus 4 and 4.1 to end conversations in rare cases of persistent abuse or harmful interactions as part of an exploration into AI welfare.

1476

Anthropic Usage Policy Update August 2025

Anthropic is updating its Usage Policy effective September 15, 2025, to refine restrictions on agentic use, political content, and law enforcement applications while clarifying high-risk consumer-facing requirements.

1477

Kimina-Prover-RL Release

Hugging Face has released Kimina-Prover-RL, an open-source RL training pipeline and two SOTA models (0.6B and 1.7B) for formal theorem proving in Lean 4.

1478

Arm and ExecuTorch 0.7: Expanding Generative AI to Billions of Devices

Arm and the ExecuTorch 0.7 beta enable automatic AI acceleration via KleidiAI, leveraging the SDOT instruction to bring LLMs like Llama 3.2 to billions of existing Arm-based devices.

1479

Arm Neural Super Sampling (NSS) Release

Arm has released Neural Super Sampling (NSS), an AI-powered upscaling solution designed to reduce GPU workloads and enable high-resolution rendering on mobile devices.

1480

FilBench: Evaluating LLM Capabilities in Philippine Languages

Hugging Face has introduced FilBench, a comprehensive evaluation suite designed to assess the fluency, linguistic abilities, and cultural knowledge of LLMs in Tagalog, Filipino, and Cebuano.

1481

TextQuests: Evaluating LLM Agentic Reasoning in Text-Based Video Games

Hugging Face introduces TextQuests, a benchmark using 25 classic Infocom interactive fiction games to evaluate the long-context reasoning and exploratory capabilities of LLM agents.

1482

Basis Scales Accounting Automation with OpenAI GPT-5 and o3

Basis uses a multi-agent architecture powered by OpenAI's GPT-5, GPT-4.1, and o3 models to automate complex accounting workflows, achieving up to 30% time savings for accounting firms.

1483

OpenAI Proposal for Harmonized AI Regulation in California

OpenAI has urged Governor Gavin Newsom to align California's AI regulations with national and global standards to prevent a patchwork of state rules that could hinder innovation and US competitiveness.

1484

Anthropic Safeguards Framework for Claude

Anthropic utilizes a multi-layered safeguards approach encompassing policy development, model training, pre-deployment testing, and real-time enforcement to prevent the misuse of Claude.

1485

Anthropic Expands Claude Access to U.S. Government Branches

Anthropic is offering Claude for Enterprise and Claude for Government to all three branches of the U.S. government for $1 to remove cost barriers to AI adoption.

1486

Claude 2025 Cyber Competition Performance and Implications

Anthropic entered Claude in seven 2025 cybersecurity competitions, where it placed in the top 25% overall but lagged behind top human teams on the hardest challenges, highlighting both the offensive potential of LLMs and the need for AI‑enabled defenses.

1487

Hugging Face AI Sheets Release

Hugging Face has released AI Sheets, an open-source no-code tool for building, transforming, and enriching datasets using open AI models.

1488

Accelerate ND-Parallel: Efficient Multi-GPU Training Guide

Hugging Face has integrated ND-Parallelism into Accelerate and Axolotl, allowing users to combine Data, Fully Sharded Data, Tensor, and Context parallelism strategies to optimize multi-GPU training for massive models.

1489

OpenAI GPT-5 Release Notes

OpenAI has released GPT-5, a unified model that integrates reasoning, agents, and advanced math capabilities to improve accuracy, speed, and problem-solving for business operations.

1490

GPT-5 for Developers Release Notes

OpenAI has released GPT-5 in three sizes (gpt-5, gpt-5-mini, and gpt-5-nano), delivering state-of-the-art performance in coding, agentic tool-calling, and long-context retrieval.

1491

GPT-5 Coding and Design Capabilities

OpenAI has announced GPT-5 with a focus on enhanced capabilities for coding and design, though specific technical details were not provided in the announcement.

1492

OpenAI GPT-5 Creative Writing Capabilities

OpenAI has announced new creative writing capabilities for GPT-5, expanding the model's ability to generate high-quality narrative and artistic text.

1493

Medical Research with GPT-5

OpenAI has announced the application of GPT-5 to medical research, expanding the model's utility in specialized scientific domains.

1494

Vision Language Model Alignment in TRL

Hugging Face has expanded the TRL library to support advanced alignment methods for Vision Language Models, including MPO, GRPO, and GSPO, alongside native SFT support and vLLM integration.

1495

GPT-5 Integration in Cursor

OpenAI has announced the integration of GPT-5 into the Cursor code editor, enhancing AI-powered software development capabilities.

1496

OpenAI GPT-5 First Look

OpenAI has provided a first look at GPT-5, marking a new generation of their large language model series.

1497

How Amgen uses GPT-5

Amgen leverages GPT-5 via the OpenAI API to accelerate biotechnology research and drug discovery processes.

1498

OpenAI GPT-5 System Card

OpenAI has released GPT-5, a unified system featuring a real-time router that dynamically switches between fast, high-throughput models and deeper reasoning models based on query complexity.

1499

GPT-5 Safe Completions: Transitioning from Refusal-Based to Output-Centric Safety Training

OpenAI has introduced safe-completions for GPT-5, a safety-training approach that maximizes helpfulness while penalizing unsafe outputs, reducing the binary comply-or-refuse trade-off for dual-use prompts.

1500

OpenAI GPT-5 Release Notes

OpenAI has released GPT-5, a unified AI system featuring a real-time router that switches between a fast model and a deeper reasoning model to provide expert-level intelligence across coding, math, and health.