6801

OpenAI Learning to Reason with LLMs

OpenAI demonstrates the reasoning capabilities of its latest models through a detailed walkthrough of a complex cipher decoding task, illustrating the step-by-step logical progression required to solve it.

6802

OpenAI o1-mini release notes

OpenAI has released o1-mini, a cost-efficient reasoning model optimized for STEM tasks that offers significantly lower latency and cost than o1-preview while maintaining high performance in math and coding.

6803

OpenAI o1 Contributions

OpenAI has published a comprehensive list of the internal and external contributors who developed the OpenAI o1 model, detailing the organizational roles and safety leadership involved in its creation.

6804

Coding with OpenAI o1

OpenAI o1 enhances software development by enabling users to build more complex and consistent code through advanced reasoning capabilities.

6805

OpenAI o1 and Quantum Physics

OpenAI o1 is a new series of AI models designed for complex reasoning in science, coding, and math by spending more time thinking before responding.

6806

OpenAI o1 and Economics

OpenAI o1 is a new series of AI models designed to reason through complex tasks in science, coding, math, and economics by spending more time thinking before responding.

6807

Decoding Genetics with OpenAI o1

OpenAI o1 is a new series of AI models designed for complex reasoning in science, coding, and math, providing geneticists like Catherine Brownstein with a tool to manage the vast complexity of genomic data.

6808

Ada Customer Service Automation with GPT-4

Ada has rebuilt its AI-native customer service platform using GPT-4, doubling its resolution rate from 30% to up to 60% (and over 80% for top performers).

6809

Hugging Face and TruffleHog Partnership for Secret Scanning

Hugging Face has partnered with Truffle Security to integrate TruffleHog's secret scanning capabilities into its automated pipeline and provide a native scanner for users to proactively scan their own account data.

6810

Qwen2-VL release: open-source 2B/7B vision-language models and 72B API with state-of-the-art image, video, and multilingual capabilities

Qwen released Qwen2-VL, a new vision-language model series (2B, 7B open-source, 72B API) that sets state-of-the-art performance on image, document, multilingual, and long-video understanding while supporting visual agent capabilities.

6811

LeRobotDataset video‑encoding format reduces robotics dataset size and speeds up training

Hugging Face released the LeRobotDataset video‑encoding format, shrinking robotics visual data to about 14 % of its original size while keeping loading speed and training performance intact.

6812

Arizona State University Integrates ChatGPT Edu for Personalized Education

Arizona State University has deployed ChatGPT Edu across more than 200 projects to personalize learning, advance research, and enhance operational efficiency for its 181,000 students.

6813

Hugging Face blog post highlights five under‑rated Hub tools and a free semantic‑search use case

Hugging Face announced five under‑rated Hub tools—ZeroGPU, multi‑process Docker, Gradio API, webhooks, and Nomic Atlas—and showed how to combine them into a free, auto‑updating semantic‑search app for Reddit data.

6814

Hugging Face Training Efficiency: Packing with Flash Attention 2

Hugging Face has introduced boundary-aware packing for instruction tuning examples, enabling up to 2x training throughput increase and 20% peak memory reduction when used with Flash Attention 2.

6815

OpenAI and Condé Nast Partnership for Content Integration

OpenAI has partnered with Condé Nast to integrate content from brands like Vogue and The New Yorker into ChatGPT and the SearchGPT prototype to improve information discovery and source attribution.

6816

Upwork OpenAI Integration and AI-First Strategy

Upwork has transitioned to an OpenAI-centric ecosystem, integrating GPT-4o and GPT-3.5 into its marketplace features, internal fraud detection, and corporate productivity via ChatGPT Enterprise.

6817

OpenAI launches fine‑tuning for GPT‑4o with free token quota and announces platform sunset

OpenAI launched fine‑tuning for GPT‑4o, enabling developers to customize the model with minimal data and offering free training tokens until September 23, while announcing the platform will close to new users after May 8 2026.

6818

Deploying Meta Llama 3.1 405B on Google Cloud Vertex AI

Hugging Face provides a guide for programmatically deploying the FP8 quantized version of Meta Llama 3.1 405B on Google Cloud Vertex AI using Text Generation Inference (TGI) and A3 machine series.

6819

OpenAI Disrupts Iranian Influence Operation Storm-2035

OpenAI has banned a cluster of ChatGPT accounts used by the Iranian influence operation Storm-2035 to generate political content for the 2024 U.S. election and other global events.

6820

Indeed Contextual Job Matching with OpenAI

Indeed integrated OpenAI's GPT models to provide personalized explanations for job recommendations, resulting in a 20% increase in started job applications and a 13% uplift in downstream success.

6821

OpenAI and The Met Museum Collaboration: Awakening Sleeping Beauties

OpenAI collaborated with the Metropolitan Museum of Art's Costume Institute to create an AI-powered chat experience allowing visitors to interact with a historical figure, Natalie Potter, based on curated historical datasets.

6822

Hugging Face Infini-Attention Reproduction Analysis

Hugging Face's attempt to reproduce Infini-Attention found that while gating convergence can be improved, the method's performance degrades with increased memory compression and remains less reliable than Ring Attention, YaRN, or RoPE scaling.

6823

OpenAI SWE-bench Verified release: human‑validated benchmark improves software‑engineering evaluation

OpenAI released SWE‑bench Verified, a human‑validated 500‑sample subset of the SWE‑bench software‑engineering benchmark that removes ambiguous issues and unfair tests, enabling more reliable evaluation of AI models’ coding abilities; GPT‑4o solves 33.2% of these samples, more than double its score on the original benchmark.

6824

Introduction to ggml

ggml is a lightweight, C/C++ machine learning library optimized for Transformer inference and on-device LLM execution across diverse hardware backends.

6825

Hugging Face Unified Tool Use API

Hugging Face has introduced a unified tool use API that allows developers to use the same code to implement tool calling across Mistral, Cohere, NousResearch, and Llama models.

6826

Falcon Mamba 7B Release Notes

The Technology Innovation Institute (TII) has released Falcon Mamba 7B, the first large-scale pure State Space Language Model (SSLM) that matches the performance of state-of-the-art transformer models while eliminating attention-based memory scaling issues.

6827

Qwen2-Audio Release Notes

Qwen2-Audio is an audio-language model capable of voice chat and audio analysis across more than eight languages and dialects, surpassing previous state-of-the-art performance on multiple benchmarks.

6828

Zico Kolter Joins OpenAI Board of Directors

OpenAI has appointed Zico Kolter, a professor and Director of the Machine Learning Department at Carnegie Mellon University, to its Board of Directors and Safety and Security Committee.

6829

Hugging Face acquires XetHub to upgrade Hub storage and collaboration

Hugging Face acquired XetHub to replace Git LFS with a more efficient storage backend, enabling incremental updates, trillion‑parameter model support, and better collaboration on massive AI datasets.

6830

GPT‑4o System Card – capabilities, safety mitigations, and risk assessment

OpenAI released the GPT‑4o System Card, detailing the omni‑model's capabilities, safety mitigations, and medium overall risk rating.

6831

Qwen2-Math Release Notes

Qwen has released Qwen2-Math, a series of specialized mathematical LLMs (1.5B, 7B, and 72B) that outperform several closed-source models, including GPT-4o, on math benchmarks.

6832

Rakuten and OpenAI Partnership: Leveraging Generative AI for Customer Insights

Rakuten is utilizing OpenAI's APIs, RAG, and Code Interpreter to transform unstructured data into automated customer service, review summaries, and B2B market insights.

6833

OpenAI Structured Outputs API feature announcement

OpenAI introduced Structured Outputs on August 6, 2024, a new API feature that guarantees model responses exactly match developer‑provided JSON Schemas, improving reliability for data‑centric applications.

6834

Hugging Face TextImage Augmentation pipeline release

Hugging Face and Albumentations AI released a TextImage Augmentation pipeline that jointly modifies document images and their text, enabling realistic synthetic data generation and robust fine‑tuning of vision‑language models on limited document datasets.

6835

Hugging Face 2024 Security Feature Highlights

Hugging Face has detailed its 2024 security landscape, introducing a suite of default protections for all users and advanced governance controls for Enterprise Hub users.

6836

Google releases Gemma 2 2B, ShieldGemma, and Gemma Scope

Google released Gemma 2 2B, ShieldGemma safety classifiers, and Gemma Scope sparse autoencoders, expanding open‑source LLM capabilities, moderation tools, and interpretability resources.

6837

Quanto quantization cuts memory for Transformer diffusion pipelines

Hugging Face quantization (Quanto) reduces GPU memory for Transformer diffusion models from ~12 GB to ~5 GB with minimal latency and quality impact.

6838

Hugging Face NVIDIA NIM API (serverless) launch and deprecation

Hugging Face launched the NVIDIA NIM API (serverless) for Enterprise Hub users, enabling pay‑as‑you‑go, serverless inference of open‑source LLMs on NVIDIA DGX Cloud H100 GPUs via an OpenAI‑compatible API.

6839

LAVE: Zero-shot VQA Evaluation on Docmatix with LLMs

Hugging Face introduces LAVE (LLM-Assisted VQA Evaluation) to address the rigidity of traditional VQA metrics, demonstrating a 50% accuracy gain in evaluating zero-shot performance on the Docmatix dataset.

6840

OpenAI SearchGPT Prototype Announcement

OpenAI has introduced SearchGPT, a temporary prototype that combines AI models with real-time web information to provide direct answers with clear source citations.

6841

OpenAI Rule-Based Rewards for Model Safety

OpenAI has introduced Rule-Based Rewards (RBRs), a method that uses explicit, step-by-step rules to align AI safety behavior without requiring extensive human data collection.

6842

Llama 3.1 Release Notes: Multilinguality, Long Context, and 405B Model

Meta has released Llama 3.1, featuring models in 8B, 70B, and 405B sizes with 128K context length, multilingual support for 8 languages, and a permissive license allowing synthetic data generation.

6843

Running Mistral 7B with Core ML

Hugging Face demonstrates how to run Mistral 7B on Mac using new Core ML features from WWDC 24, achieving a model size reduction to under 4GB using 4-bit block-wise quantization.

6844

GPT-4o mini release notes / what's new

OpenAI has released GPT-4o mini, a highly cost-efficient small model that outperforms GPT-3.5 Turbo and other small models on key reasoning and multimodal benchmarks.

6845

Docmatix Dataset Release

Hugging Face has released Docmatix, a Document Visual Question Answering (DocVQA) dataset featuring 2.4 million images and 9.5 million Q/A pairs, providing a 240x increase in scale over previous datasets.

6846

TGI Multi-LoRA: Deploy Once, Serve 30 Models

Hugging Face introduces Multi-LoRA serving in Text Generation Inference (TGI), allowing organizations to deploy a single base model and dynamically serve dozens of specialized fine-tuned adapters to reduce cost and operational complexity.

6847

ChatGPT Enterprise Compliance and Administrative Tools Update

OpenAI has introduced a Compliance API, SCIM for automated user management, and granular GPT controls to help enterprise customers meet regulatory requirements and scale AI deployments securely.

6848

OpenAI Prover‑Verifier Games improve legibility of language model outputs

OpenAI announced a prover‑verifier game training method that makes strong language models generate solutions that weaker models can verify, improving both correctness and human legibility.

6849

Argilla SDK Chatbot with distilabel – End‑to‑End Tutorial

Hugging Face released a tutorial showing how to build an Argilla 2.0 chatbot using distilabel‑generated synthetic data, fine‑tuned embeddings, lancedb vector storage, and a Gradio app deployed on Spaces.

6850

SmolLM Release: High-Performance Small Language Models

Hugging Face introduces SmolLM, a family of state-of-the-art small language models (135M, 360M, and 1.7B parameters) trained on a meticulously curated high-quality dataset.