OpenAI Learning to Reason with LLMs
OpenAI demonstrates the reasoning capabilities of its latest models through a detailed walkthrough of a complex cipher decoding task, illustrating the step-by-step logical progression required to solve it.
OpenAI o1-mini release notes
OpenAI has released o1-mini, a cost-efficient reasoning model optimized for STEM tasks that offers significantly lower latency and cost than o1-preview while maintaining high performance in math and coding.
OpenAI o1 Contributions
OpenAI has published a comprehensive list of the internal and external contributors who developed the OpenAI o1 model, detailing the organizational roles and safety leadership involved in its creation.
Coding with OpenAI o1
OpenAI o1 enhances software development by enabling users to build more complex and consistent code through advanced reasoning capabilities.
OpenAI o1 and Quantum Physics
OpenAI o1 is a new series of AI models designed for complex reasoning in science, coding, and math by spending more time thinking before responding.
OpenAI o1 and Economics
OpenAI o1 is a new series of AI models designed to reason through complex tasks in science, coding, math, and economics by spending more time thinking before responding.
Decoding Genetics with OpenAI o1
OpenAI o1 is a new series of AI models designed for complex reasoning in science, coding, and math, providing geneticists like Catherine Brownstein with a tool to manage the vast complexity of genomic data.
Ada Customer Service Automation with GPT-4
Ada has rebuilt its AI-native customer service platform using GPT-4, doubling its resolution rate from 30% to up to 60% (and over 80% for top performers).
Hugging Face and TruffleHog Partnership for Secret Scanning
Hugging Face has partnered with Truffle Security to integrate TruffleHog's secret scanning capabilities into its automated pipeline and provide a native scanner for users to proactively scan their own account data.
Qwen2-VL release: open-source 2B/7B vision-language models and 72B API with state-of-the-art image, video, and multilingual capabilities
Qwen released Qwen2-VL, a new vision-language model series (2B, 7B open-source, 72B API) that sets state-of-the-art performance on image, document, multilingual, and long-video understanding while supporting visual agent capabilities.
LeRobotDataset video‑encoding format reduces robotics dataset size and speeds up training
Hugging Face released the LeRobotDataset video‑encoding format, shrinking robotics visual data to about 14 % of its original size while keeping loading speed and training performance intact.
Arizona State University Integrates ChatGPT Edu for Personalized Education
Arizona State University has deployed ChatGPT Edu across more than 200 projects to personalize learning, advance research, and enhance operational efficiency for its 181,000 students.
Hugging Face blog post highlights five under‑rated Hub tools and a free semantic‑search use case
Hugging Face announced five under‑rated Hub tools—ZeroGPU, multi‑process Docker, Gradio API, webhooks, and Nomic Atlas—and showed how to combine them into a free, auto‑updating semantic‑search app for Reddit data.
Hugging Face Training Efficiency: Packing with Flash Attention 2
Hugging Face has introduced boundary-aware packing for instruction tuning examples, enabling up to 2x training throughput increase and 20% peak memory reduction when used with Flash Attention 2.
OpenAI and Condé Nast Partnership for Content Integration
OpenAI has partnered with Condé Nast to integrate content from brands like Vogue and The New Yorker into ChatGPT and the SearchGPT prototype to improve information discovery and source attribution.
Upwork OpenAI Integration and AI-First Strategy
Upwork has transitioned to an OpenAI-centric ecosystem, integrating GPT-4o and GPT-3.5 into its marketplace features, internal fraud detection, and corporate productivity via ChatGPT Enterprise.
OpenAI launches fine‑tuning for GPT‑4o with free token quota and announces platform sunset
OpenAI launched fine‑tuning for GPT‑4o, enabling developers to customize the model with minimal data and offering free training tokens until September 23, while announcing the platform will close to new users after May 8 2026.
Deploying Meta Llama 3.1 405B on Google Cloud Vertex AI
Hugging Face provides a guide for programmatically deploying the FP8 quantized version of Meta Llama 3.1 405B on Google Cloud Vertex AI using Text Generation Inference (TGI) and A3 machine series.
OpenAI Disrupts Iranian Influence Operation Storm-2035
OpenAI has banned a cluster of ChatGPT accounts used by the Iranian influence operation Storm-2035 to generate political content for the 2024 U.S. election and other global events.
Indeed Contextual Job Matching with OpenAI
Indeed integrated OpenAI's GPT models to provide personalized explanations for job recommendations, resulting in a 20% increase in started job applications and a 13% uplift in downstream success.
OpenAI and The Met Museum Collaboration: Awakening Sleeping Beauties
OpenAI collaborated with the Metropolitan Museum of Art's Costume Institute to create an AI-powered chat experience allowing visitors to interact with a historical figure, Natalie Potter, based on curated historical datasets.
Hugging Face Infini-Attention Reproduction Analysis
Hugging Face's attempt to reproduce Infini-Attention found that while gating convergence can be improved, the method's performance degrades with increased memory compression and remains less reliable than Ring Attention, YaRN, or RoPE scaling.
OpenAI SWE-bench Verified release: human‑validated benchmark improves software‑engineering evaluation
OpenAI released SWE‑bench Verified, a human‑validated 500‑sample subset of the SWE‑bench software‑engineering benchmark that removes ambiguous issues and unfair tests, enabling more reliable evaluation of AI models’ coding abilities; GPT‑4o solves 33.2% of these samples, more than double its score on the original benchmark.
Introduction to ggml
ggml is a lightweight, C/C++ machine learning library optimized for Transformer inference and on-device LLM execution across diverse hardware backends.
Hugging Face Unified Tool Use API
Hugging Face has introduced a unified tool use API that allows developers to use the same code to implement tool calling across Mistral, Cohere, NousResearch, and Llama models.
Falcon Mamba 7B Release Notes
The Technology Innovation Institute (TII) has released Falcon Mamba 7B, the first large-scale pure State Space Language Model (SSLM) that matches the performance of state-of-the-art transformer models while eliminating attention-based memory scaling issues.
Qwen2-Audio Release Notes
Qwen2-Audio is an audio-language model capable of voice chat and audio analysis across more than eight languages and dialects, surpassing previous state-of-the-art performance on multiple benchmarks.
Zico Kolter Joins OpenAI Board of Directors
OpenAI has appointed Zico Kolter, a professor and Director of the Machine Learning Department at Carnegie Mellon University, to its Board of Directors and Safety and Security Committee.
Hugging Face acquires XetHub to upgrade Hub storage and collaboration
Hugging Face acquired XetHub to replace Git LFS with a more efficient storage backend, enabling incremental updates, trillion‑parameter model support, and better collaboration on massive AI datasets.
GPT‑4o System Card – capabilities, safety mitigations, and risk assessment
OpenAI released the GPT‑4o System Card, detailing the omni‑model's capabilities, safety mitigations, and medium overall risk rating.
Qwen2-Math Release Notes
Qwen has released Qwen2-Math, a series of specialized mathematical LLMs (1.5B, 7B, and 72B) that outperform several closed-source models, including GPT-4o, on math benchmarks.
Rakuten and OpenAI Partnership: Leveraging Generative AI for Customer Insights
Rakuten is utilizing OpenAI's APIs, RAG, and Code Interpreter to transform unstructured data into automated customer service, review summaries, and B2B market insights.
OpenAI Structured Outputs API feature announcement
OpenAI introduced Structured Outputs on August 6, 2024, a new API feature that guarantees model responses exactly match developer‑provided JSON Schemas, improving reliability for data‑centric applications.
Hugging Face TextImage Augmentation pipeline release
Hugging Face and Albumentations AI released a TextImage Augmentation pipeline that jointly modifies document images and their text, enabling realistic synthetic data generation and robust fine‑tuning of vision‑language models on limited document datasets.
Hugging Face 2024 Security Feature Highlights
Hugging Face has detailed its 2024 security landscape, introducing a suite of default protections for all users and advanced governance controls for Enterprise Hub users.
Google releases Gemma 2 2B, ShieldGemma, and Gemma Scope
Google released Gemma 2 2B, ShieldGemma safety classifiers, and Gemma Scope sparse autoencoders, expanding open‑source LLM capabilities, moderation tools, and interpretability resources.
Quanto quantization cuts memory for Transformer diffusion pipelines
Hugging Face quantization (Quanto) reduces GPU memory for Transformer diffusion models from ~12 GB to ~5 GB with minimal latency and quality impact.
Hugging Face NVIDIA NIM API (serverless) launch and deprecation
Hugging Face launched the NVIDIA NIM API (serverless) for Enterprise Hub users, enabling pay‑as‑you‑go, serverless inference of open‑source LLMs on NVIDIA DGX Cloud H100 GPUs via an OpenAI‑compatible API.
LAVE: Zero-shot VQA Evaluation on Docmatix with LLMs
Hugging Face introduces LAVE (LLM-Assisted VQA Evaluation) to address the rigidity of traditional VQA metrics, demonstrating a 50% accuracy gain in evaluating zero-shot performance on the Docmatix dataset.
OpenAI SearchGPT Prototype Announcement
OpenAI has introduced SearchGPT, a temporary prototype that combines AI models with real-time web information to provide direct answers with clear source citations.
OpenAI Rule-Based Rewards for Model Safety
OpenAI has introduced Rule-Based Rewards (RBRs), a method that uses explicit, step-by-step rules to align AI safety behavior without requiring extensive human data collection.
Llama 3.1 Release Notes: Multilinguality, Long Context, and 405B Model
Meta has released Llama 3.1, featuring models in 8B, 70B, and 405B sizes with 128K context length, multilingual support for 8 languages, and a permissive license allowing synthetic data generation.
Running Mistral 7B with Core ML
Hugging Face demonstrates how to run Mistral 7B on Mac using new Core ML features from WWDC 24, achieving a model size reduction to under 4GB using 4-bit block-wise quantization.
GPT-4o mini release notes / what's new
OpenAI has released GPT-4o mini, a highly cost-efficient small model that outperforms GPT-3.5 Turbo and other small models on key reasoning and multimodal benchmarks.
Docmatix Dataset Release
Hugging Face has released Docmatix, a Document Visual Question Answering (DocVQA) dataset featuring 2.4 million images and 9.5 million Q/A pairs, providing a 240x increase in scale over previous datasets.
TGI Multi-LoRA: Deploy Once, Serve 30 Models
Hugging Face introduces Multi-LoRA serving in Text Generation Inference (TGI), allowing organizations to deploy a single base model and dynamically serve dozens of specialized fine-tuned adapters to reduce cost and operational complexity.
ChatGPT Enterprise Compliance and Administrative Tools Update
OpenAI has introduced a Compliance API, SCIM for automated user management, and granular GPT controls to help enterprise customers meet regulatory requirements and scale AI deployments securely.
OpenAI Prover‑Verifier Games improve legibility of language model outputs
OpenAI announced a prover‑verifier game training method that makes strong language models generate solutions that weaker models can verify, improving both correctness and human legibility.
Argilla SDK Chatbot with distilabel – End‑to‑End Tutorial
Hugging Face released a tutorial showing how to build an Argilla 2.0 chatbot using distilabel‑generated synthetic data, fine‑tuned embeddings, lancedb vector storage, and a Gradio app deployed on Spaces.
SmolLM Release: High-Performance Small Language Models
Hugging Face introduces SmolLM, a family of state-of-the-art small language models (135M, 360M, and 1.7B parameters) trained on a meticulously curated high-quality dataset.