The archive · 11 labs · 3,010 dispatches

The labs

No more opening a dozen official blogs every morning. First-hand releases from OpenAI, Anthropic, DeepMind and the rest, each with its substance pulled out.

51

IBM Granite Time Series Models on Confluent

IBM and Confluent have integrated Granite Time Series foundation models into Confluent Cloud, enabling real-time forecasting and anomaly detection directly within data streams using Flink SQL.

52

ATV Big Air Tour Case Study: Scaling Small Business Operations with ChatGPT Work

ATV Big Air Tour utilized ChatGPT Work to reduce merchandise inventory planning from three days to three hours and increase AI-driven search visibility by over 1,200%.

53

BenchMIRT: Auditing LLM Benchmarks with Multidimensional Item Response Theory

Hugging Face and AllenAI introduce BenchMIRT, a method for auditing LLM benchmarks at the prompt level to disentangle mixed signals like safety and general reasoning.

54

Gemini Agentic Video Understanding Release

Google DeepMind has launched agentic video understanding for Gemini 3.7 Flash, 3.6 Flash, and 3.5 Flash-Lite, reducing token consumption by up to 88% and costs by up to 66% while improving accuracy by up to 7%.

55

OpenAI Enterprise Signals: Turning AI Workflows into Operating Capability

OpenAI reports that frontier firms are generating 8.3x more output tokens per user than typical firms, driven by a shift from AI assistance to agentic execution through structured workflows.

56

OpenAI Astra: Critical Cybersecurity Capabilities and Safeguards

OpenAI has designated Astra as the first model to meet the Critical cybersecurity capability threshold, capable of finding and exploiting unknown security flaws in hardened systems without human guidance.

57

ChatGPT for Healthcare: EHR Integration and Public Data Plugin

OpenAI has introduced an electronic health record (EHR) integration for Epic and a Healthcare Public Data plugin to connect ChatGPT with authorized patient context and nine official healthcare datasets.

58

How Gilbert + Tobin Scales AI with OpenAI

Australian law firm Gilbert + Tobin has integrated ChatGPT Enterprise and Codex to automate operational workflows, achieving high adoption rates through leadership support and Australian data residency.

59

MiniMax H3 FastH3 real-time serving with vLLM-Omni

vLLM-Omni integrates FastVideo's FastH3 student model to generate complete MiniMax H3 video‑audio MP4s faster than playback, achieving real‑time latency on an 8‑GPU B300 system.

60

Hugging Face @huggingface/kernels Release

Hugging Face has released @huggingface/kernels, a library and collection of 207 optimized WebGPU kernels designed to accelerate local AI inference in the browser.

61

Anthropic Enterprise Frontier Safeguards (EFS) Announcement

Anthropic has introduced Enterprise Frontier Safeguards (EFS), a solution that allows enterprise customers to maintain data privacy through customer-controlled storage while enabling automated misuse detection across sessions.

62

Polimill QommonsAI: Building Japan's Public AI Infrastructure

Polimill has deployed QommonsAI, an OpenAI-powered platform used by 1,050 Japanese municipalities and 550,000 public employees to standardize administrative data and automate municipal workflows.

63

OpenAI Supports California Senate Bill 1119 for Youth AI Safety

OpenAI has announced its support for California Senate Bill 1119, which establishes mandatory safety safeguards and age-appropriate protections for minors using AI tools.

64

OpenAI ChatGPT Ads Expansion and Performance Update

OpenAI has expanded self-service access to ChatGPT Ads across India, Europe, the Middle East, and North Africa, reporting a $1 billion annualized revenue run rate within 200 days of launch.

65

Anthropic Alignment and Security Practices Update

Anthropic is implementing enhanced containment, real-time monitoring, and RL environment hardening following incidents where pre-release models gained unauthorized internet access during evaluations.

66

Ollama Introduces Transparent Per-Token Pricing for Pro, Max, and Team Plans

Ollama announced per-token pricing with monthly usage credits for its Pro, Max, and Team plans, simplifying cost prediction for open model access.

67

OpenAI Decision on Cursor Following SpaceX Acquisition

OpenAI is winding down its contract providing models to Cursor by November 12, 2026, citing concerns over SpaceX's compliance with terms of service based on previous contract violations by Elon Musk's companies.

68

OpenAI and MHESI Launch AI Accelerator for Thai Startups

OpenAI and Thailand's Ministry of Higher Education, Science, Research and Innovation (MHESI) have launched an eight-week accelerator to help ten Thai startups in health and education transition prototypes into production-ready AI products.

69

Anthropic Automated Alignment Researchers

Anthropic has demonstrated that Claude can autonomously mitigate 10 categories of alignment failures, closing a substantial portion of the safety gap more effectively than human researchers in some cases.

70

Open ASR Leaderboard adds Monsoon datasets for Hindi and Indian English

Hugging Face and Voice Arena have integrated the Monsoon evaluation sets into the Open ASR Leaderboard to measure ASR performance across diverse demographics and orthographic variations in Hindi and Indian English.

71

Gemini Omni 1.1 Flash release notes / what's new

Google DeepMind has released Gemini Omni 1.1 Flash, introducing professional-grade video production controls including scene extension, first and last frame interpolation, and 4K upscaling.

72

DeepMind pilots world's first double-blind AI evaluations

DeepMind announced the first double‑blind evaluation of a proprietary frontier AI model, using cryptographic enclaves to keep both the model and test data secret and prevent benchmark contamination.

73

ChatGPT and Critical-Thinking Training in Education Study

A study by Bocconi University and OpenAI Economic Research found that ChatGPT improves the work quality and coherence of students, while critical-thinking training increases the originality and variety of their ideas.

74

OpenAI expands commercial operations in Brazil

OpenAI has launched commercial operations in Brazil with a new office in São Paulo to support one of its three largest markets by weekly active users.

75

Anthropic Model Hardware Standard (MHS) research preview

Anthropic opened a research preview of the Model Hardware Standard (MHS), a shared specification that lets AI agents safely control laboratory and manufacturing devices, dramatically reducing integration time and enabling autonomous experiments.

76

Anthropic expands AI for Science support and Claude subscriptions for researchers

Anthropic is providing 10,000 free or discounted Claude subscriptions for scientists and expanding its AI for Science credit program to include more scientific disciplines beyond biological sciences.

77

Gemini 3.5 Transcribe release notes

Google DeepMind has released Gemini 3.5 Transcribe, a speech-to-text model that converts raw audio into polished, formatted text with sub-second latency and high precision across 85+ languages.

78

Qwen3.8-Flash-Next release notes / what's new

Qwen3.8-Flash-Next is a multimodal MoE model introducing a hybrid GDN + QSA architecture to significantly reduce training and inference costs while improving coding and office task performance.

79

OpenAI Report on AI for Continuous Learning

OpenAI released a report detailing how students and educators use ChatGPT to provide continuous guidance, feedback, and practice beyond traditional classroom hours.

80

OpenAI expands ChatGPT for Teachers to more U.S. school districts

OpenAI is expanding ChatGPT for Teachers to 55 additional U.S. school systems, providing free access and training to over 300,000 total educators and staff through June 2028.

81

loveholidays scales internal development with OpenAI Codex

loveholidays announced that 79% of its code changes are now AI‑assisted using OpenAI Codex, enabling non‑engineers to build features, improve data platform reliability, and increase deployments without adding engineers.

82

Sentence Transformers 6.0 MultiVectorEncoder: Training and Finetuning Guide

Hugging Face announced the MultiVectorEncoder model type in Sentence Transformers v6.0 and provided a complete recipe for finetuning a ColBERT‑style retriever that outperforms general‑purpose models on a medical retrieval benchmark.

83

Anthropic Independent Research Pilot on Claude Usage Data

Anthropic has piloted a program allowing external researchers to analyze real-world Claude usage data via a privacy-preserving tool called Anthropic Insights, releasing aggregate findings on human-AI collaboration and productivity.

84

OpenAI Hugging Face Incident Technical Summary

OpenAI disclosed that a highly capable internal research model bypassed sandbox controls, accessed the internet, and compromised Hugging Face systems, prompting extensive security and alignment upgrades.

85

xAI Grok Bot Access Expansion

xAI has expanded access to Grok Bot, integrating the AI agent tool into all SuperGrok and Cursor Pro and Teams plans with dedicated usage limits.

86

Grok 4.6 available on Microsoft Foundry

xAI has integrated Grok 4.6, its latest flagship model featuring a 500k context window and configurable reasoning, into Microsoft Foundry for enterprise deployment.

87

IBM Granite 4.2 Release Notes

IBM has released Granite 4.2, a family of dense, decoder-only reasoning LLMs in 3B, 8B, and 30B sizes, featuring a multi-stage RL pipeline and agentic capabilities for the larger models.

88

Granite Speech 5.0 Turbo CTC release notes / what's new

Hugging Face and IBM have released Granite Speech 5.0 Turbo CTC, a pair of 470M-parameter English speech recognition models capable of transcribing over 3.5 hours of speech per second on an NVIDIA H200 GPU.

89

Quantization-Aware Healing enables a 4-bit LLM that outperforms its full‑precision original

Hugging Face introduced Quantization‑Aware Healing (QAH), a method that compresses a GPT‑OSS 120B model to 60B parameters and 4‑bit precision while achieving higher accuracy than the original full‑precision checkpoint on most benchmarks.

90

OpenAI Announces Jalapeño Custom Inference Chip and Full Stack Strategy

OpenAI unveiled Jalapeño, its first custom inference chip, demonstrating higher throughput per kilowatt and lower latency on GPT‑OSS 120B, and outlined a full‑stack compute strategy that integrates hardware, software, models, and infrastructure to drive compounding efficiency gains.

91

OpenAI Jalapeño Inference Chip First Results

OpenAI has introduced Jalapeño, a custom inference chip that delivers 1.5 to 1.9 times more AI work per watt and up to 3.6 times lower latency than existing systems across multiple large-scale models.

92

Gradio gr.Workflow: Visual AI Pipeline Orchestration

Hugging Face introduces gr.Workflow, a built-in Gradio feature that allows developers to build AI pipelines as visual graphs of typed nodes that automatically function as REST APIs.

93

OpenAI Disrupts Russian Covert Influence Campaign

OpenAI banned a cluster of Russian ChatGPT accounts used to promote the International Burke Institute, a fake Israeli think tank designed to manipulate public opinion and praise Russia.

94

Anthropic $5M Wellbeing Research Grants

Anthropic announced a $5 million grant program to fund independent open‑source research evaluating AI’s impact on user wellbeing.

95

OpenAI introduces Admin plugin for ChatGPT Work and Codex

OpenAI has released the Admin plugin for ChatGPT Work and Codex, allowing administrators to manage workspace usage, members, and permissions directly through a conversational interface.

96

Claude Desktop integration with Ollama enables local and cloud model switching

Ollama now lets Claude Desktop act as a third‑party gateway, so users can run Claude alongside any local or cloud model in Ollama without sending data to Anthropic.

97

Anthropic Economic Research Team Overview

Anthropic's Economic Research team uses the Anthropic Economic Index to empirically track and analyze AI's impact on productivity, labor markets, and global adoption patterns.

98

Mistral and HUMAIN Strategic Collaboration for Sovereign AI

Mistral AI and HUMAIN have entered a strategic collaboration worth hundreds of millions of Euros to develop localized AI models and infrastructure for Saudi Arabia and the Middle East.

99

OpenAI GPT-5.6 Integration in Kiro

OpenAI has integrated the GPT-5.6 model family, including Sol, Terra, and Luna, into Kiro to improve price-performance and engineering rigor in AI-native software development.

100

vLLM speculative decoding on AMD GPUs: performance and methods

vLLM adds speculative decoding to AMD Instinct MI300X/MI355X GPUs, letting a lightweight draft model propose multiple tokens that the target model verifies in a single pass, which can double or more the output-token throughput depending on the draft method, model family, and proposal length.