5701

OpenMed CodonRoBERTa multi-species mRNA language models release

OpenMed released an end-to-end protein engineering pipeline with CodonRoBERTa-large-v2 (perplexity 4.10, CAI 0.404) and a 25-species codon‑optimization model suite trained in 55 GPU‑hours for $165.

5702

TRL v1.0 release notes / what's new

Hugging Face releases TRL v1.0, a stable post-training library implementing over 75 methods, featuring a dual-track stability model to balance rapid experimental iteration with production-grade reliability.

5703

vLLM Hidden States Extraction System

vLLM v0.18.0 introduces a native hidden states extraction system that enables high-performance retrieval of internal model representations for training speculative decoding draft models.

5704

OpenAI AI Jam for Disaster Management in Asia

OpenAI, in partnership with the Gates Foundation, APDC, and DataKind, hosted an AI Jam in Bangkok to help disaster management professionals from 13 Asian countries develop practical AI workflows for emergency response.

5705

Qwen3.5-Omni release notes

Qwen announced Qwen3.5-Omni, a new omnimodal LLM that handles text, images, audio, and video, supports 256k context, 113-language speech recognition, 36-language synthesis, and adds real-time features like semantic interruption, websearch, voice control, and voice cloning.

5706

Google DeepMind AI-Enabled Pointer

Google DeepMind is developing an AI-enabled pointer powered by Gemini that replaces complex text prompts with intuitive pointing and voice commands to interact with digital content across applications.

5707

STADLER AI Implementation and Productivity Gains

STADLER, a 230-year-old industrial recycling company, achieved 30-40% time savings on knowledge tasks and 2.5x faster drafting by embedding OpenAI's ChatGPT as a company-wide productivity layer.

5708

Hugging Face OpenClaw Migration Guide

Hugging Face provides two methods—Inference Providers and local llama.cpp setup—to migrate OpenClaw agents from restricted Claude models to open-source alternatives.

5709

Gemini 3.1 Flash Live release notes / what's new

Google DeepMind has released Gemini 3.1 Flash Live, a high-quality audio and voice model designed for natural, real-time dialogue with improved precision, lower latency, and expanded global availability.

5710

Google DeepMind Research on AI Harmful Manipulation

Google DeepMind has released a new research study and an empirically validated toolkit to measure and mitigate the risk of AI being used for harmful manipulation of human thought and behavior.

5711

Lyria 3 Pro release: longer, structurally aware music generation across Google products

DeepMind announced Lyria 3 Pro, a music‑generation model that creates up to three‑minute tracks with structural awareness and is now integrated across Google products like Vertex AI, AI Studio, Google Vids, Gemini, and ProducerAI.

5712

OpenAI Model Spec: A Framework for Explicit Model Behavior

OpenAI has introduced the Model Spec, a formal, public framework designed to make intended AI model behavior explicit, legible, and revisable for users, developers, and researchers.

5713

OpenAI Safety Bug Bounty Program Launch

OpenAI has launched a public Safety Bug Bounty program to identify AI abuse and safety risks that fall outside conventional security vulnerabilities, specifically targeting agentic risks, proprietary information leaks, and platform integrity.

5714

OpenAI Teen Safety Policy Pack and gpt-oss-safeguard

OpenAI has released prompt-based safety policies and the open-weight gpt-oss-safeguard model to help developers implement age-appropriate protections for teenagers in AI applications.

5715

OpenAI expands product discovery in ChatGPT via Agentic Commerce Protocol

OpenAI has introduced enhanced visual shopping and product discovery capabilities in ChatGPT, powered by the expanded Agentic Commerce Protocol (ACP) to streamline how users find and compare products.

5716

OpenAI Foundation Update

OpenAI has announced that its Foundation will invest at least $1 billion over the next year across life sciences, economic impact, AI resilience, and community programs to ensure AGI benefits humanity.

5717

EVA End-to-End Evaluation Framework for Voice Agents

EVA is a new end‑to‑end framework that jointly evaluates voice agents on accuracy and conversational experience, revealing a consistent trade‑off between task success and user satisfaction.

5718

vLLM Model Runner V2 release notes / what's new

vLLM has introduced Model Runner V2 (MRV2), a ground-up re-implementation of the model runner that improves throughput and reduces latency through a GPU-native, async-first, and modular architecture.

5719

OpenAI Sora 2 Safety Framework

OpenAI has detailed the safety architecture for Sora 2, focusing on provenance signals, consent-based likeness management, and strict content filtering for audio and video.

5720

Domain-Specific Embedding Fine-Tuning with NVIDIA Nemotron – Under a Day

NVIDIA and Hugging Face released a single‑GPU, under‑a‑day pipeline that fine‑tunes the Llama‑Nemotron‑Embed‑1B‑v2 model on synthetic domain data, delivering >10% retrieval gains and up to 26% improvement on real enterprise datasets.

5721

OpenAI Internal Coding Agent Monitoring System

OpenAI has deployed a GPT-5.4 Thinking-powered monitoring system to detect misalignment and security violations in internal coding agents, identifying behaviors that often only emerge in complex, tool-rich workflows.

5722

OpenAI to acquire Astral

OpenAI is acquiring Astral to integrate its open-source Python tools, including uv, Ruff, and ty, into the Codex ecosystem to enable AI agents to participate in the entire software development lifecycle.

5723

Qwen3.5-Max-Preview Release on LMSys Arena

Qwen has deployed Qwen3.5-Max-Preview to the LMSys Arena for community evaluation ahead of its full release scheduled within two weeks.

5724

State of Open Source on Hugging Face: Spring 2026

Hugging Face reports a massive expansion of the open source AI ecosystem in 2025, characterized by China surpassing the U.S. in model downloads and the rapid emergence of robotics as the largest dataset category.

5725

Google DeepMind Measuring Progress Toward AGI: A Cognitive Framework

Google DeepMind has introduced a cognitive taxonomy and a three-stage evaluation protocol to empirically measure AI progress toward Artificial General Intelligence (AGI).

5726

Holotron-12B High Throughput Computer Use Agent

H Company released Holotron-12B, a multimodal computer-use model based on NVIDIA Nemotron-Nano-2 VL that uses a hybrid SSM-Attention architecture to achieve high inference throughput for agentic workloads.

5727

OpenAI GPT-5.4 mini and nano release notes

OpenAI has released GPT-5.4 mini and nano, high-efficiency small models that bring GPT-5.4 capabilities to high-volume workloads with significantly lower latency and cost.

5728

OpenAI Japan Teen Safety Blueprint

OpenAI Japan has introduced the Japan Teen Safety Blueprint, a framework prioritizing teen safety over convenience and privacy to protect younger users from AI-related risks.

5729

OpenAI Research: How Workers Use ChatGPT for Compensation Insights

OpenAI research reveals that US workers send nearly 3 million daily messages to ChatGPT seeking wage benchmarks and compensation guidance, particularly in high-skill, low-transparency roles.

5730

OpenAI Codex Security: Why the System Avoids SAST Report Seeding

OpenAI's Codex Security avoids starting with Static Application Security Testing (SAST) reports to prevent premature narrowing of analysis and to focus on validating whether security invariants actually hold through transformation chains.

5731

BAIR Introducing SPEX and ProxySPEX for Scalable LLM Interaction Discovery

BAIR has introduced SPEX and ProxySPEX, algorithms that use signal processing and coding theory to identify influential interactions between features, training data, and model components at scale.

5732

P-EAGLE: Parallel Speculative Decoding in vLLM

vLLM introduces P-EAGLE, a parallel speculative decoding method that generates all draft tokens in a single forward pass, delivering up to 1.69x speedup over vanilla EAGLE-3 on NVIDIA B200 GPUs.

5733

OpenAI Designing AI Agents to Resist Prompt Injection

OpenAI is shifting its defense strategy against prompt injection by treating it as a social engineering problem, focusing on constraining the impact of successful manipulations rather than relying solely on input filtering.

5734

OpenAI Responses API Computer Environment Update

OpenAI has equipped the Responses API with a shell tool and hosted container workspace, enabling models to execute real-world tasks via a command-line interface and persistent runtime context.

5735

NVIDIA Nemotron 3 Super Support in vLLM

vLLM now supports NVIDIA Nemotron 3 Super, a 120B parameter hybrid MoE model optimized for multi-agent AI with a 1 million token context window and high inference efficiency.

5736

Rakuten Integration of OpenAI Codex for Engineering Efficiency

Rakuten has integrated OpenAI Codex into its engineering stack, achieving a 50% reduction in mean time to recovery (MTTR) and compressing quarter-long development projects into weeks.

5737

Wayfair OpenAI Integration Case Study

Wayfair has integrated OpenAI models into its internal systems to automate product catalog tagging for 30 million items and streamline supplier support via the AI-powered tool Wilma.

5738

OpenAI Instruction Hierarchy Improvements and GPT-5 Mini-R

OpenAI has introduced a new reinforcement learning dataset, IH-Challenge, to train models to prioritize trusted instructions over untrusted ones, resulting in the GPT-5 Mini-R model with improved safety steerability and prompt injection robustness.

5739

ChatGPT Interactive Visuals for Math and Science

OpenAI has introduced dynamic visual explanations for over 70 core math and science concepts in ChatGPT to help users understand the relationships between variables and formulas in real time.

5740

vLLM Semantic Router v0.2 Athena release notes / what's new

vLLM Semantic Router v0.2 Athena introduces a rebuilt model stack, the experimental ClawOS orchestration layer, and advanced model selection primitives to transform semantic routing into a strategic system brain for multi-agent deployments.

5741

Hugging Face Storage Buckets Release

Hugging Face has introduced Storage Buckets, a mutable, S3-like object storage system backed by Xet for efficient handling of intermediate ML artifacts like checkpoints and processed data.

5742

Keep the Tokens Flowing: Lessons from 16 Open-Source RL Libraries

Hugging Face surveyed 16 open-source RL libraries and found that async RL training separates inference and training onto different GPU pools, uses a rollout buffer, and pushes weights asynchronously, with Ray dominating orchestration and NCCL broadcast the common weight‑sync method.

5743

Google DeepMind: 10 Years of AlphaGo's Impact

Google DeepMind reflects on how AlphaGo's 2016 victory over a world champion catalyzed a decade of AI breakthroughs in science, mathematics, and the pursuit of Artificial General Intelligence (AGI).

5744

OpenAI to Acquire Promptfoo

OpenAI is acquiring Promptfoo to integrate its AI security and evaluation tools into the OpenAI Frontier platform for building enterprise AI coworkers.

5745

LeRobot v0.5.0 release notes / what's new

Hugging Face has released LeRobot v0.5.0, introducing full Unitree G1 humanoid support, new VLA policies like Pi0-FAST and Wall-X, and significant dataset performance optimizations.

5746

Ulysses Sequence Parallelism for Million-Token Context Training

Hugging Face has integrated Ulysses Sequence Parallelism into Accelerate, Transformers, and TRL, enabling the training of LLMs with million-token contexts by distributing attention computation across multiple GPUs.

5747

OpenAI Codex Security Research Preview

OpenAI has introduced Codex Security, an application security agent that uses frontier models and automated validation to identify high-confidence vulnerabilities and provide actionable fixes.

5748

Balyasny Asset Management AI Research Engine Case Study

Balyasny Asset Management developed a centralized AI research engine using GPT-5.4 to reduce deep research tasks from days to hours and automate complex financial analysis.

5749

Descript GPT-5 dubbing pipeline boosts multilingual video localization

Descript unveiled a GPT‑5‑powered dubbing pipeline that jointly optimizes meaning and timing, raising dubbed video exports by 15% and improving duration adherence by up to 43 points, enabling scalable multilingual video localization.

5750

Bringing Robotics AI to Embedded Platforms: Dataset Recording, VLA Fine-Tuning, and On-Device Optimizations

Hugging Face and NXP provide a technical guide on deploying Vision-Language-Action (VLA) models on the i.MX 95 SoC, emphasizing dataset consistency, architectural decomposition, and asynchronous inference to achieve real-time robotic control.