AMD EPYC Turin CPU delivers 2× LLM inference throughput over Genoa
AMD’s 5th‑gen EPYC Turin CPU delivers roughly double the LLM inference throughput of Genoa, enabling lower latency and higher throughput for Hugging Face workloads.
Scaling AI Data Processing with Hugging Face and Dask
Hugging Face and Dask enable the scaling of AI-based data processing from small local samples to hundreds of millions of rows using distributed computing and multi-GPU parallel inference.
Gradio 5 Release Notes
Hugging Face has released Gradio 5, a production-ready framework for building performant, scalable, and secure machine learning web applications using Python.
OpenAI and Hearst Content Partnership
OpenAI has partnered with Hearst to integrate content from over 20 magazine brands and 40+ newspapers into its AI products, including ChatGPT, to provide users with more reliable and cited journalism.
Hugging Face Transformers 4.45.0 Dynamic Speculative Decoding
Hugging Face and Intel Labs introduced dynamic speculative decoding in Transformers 4.45.0, accelerating text generation by up to 2.7x by dynamically adjusting the number of draft tokens based on model confidence.
Improving Parquet Deduplication on Hugging Face Hub
Hugging Face is optimizing its storage architecture to improve Parquet file deduplication, proposing content-defined row groups to reduce storage overhead during dataset updates.
Open FinLLM Leaderboard launch – comprehensive zero‑shot benchmark for financial language models
Hugging Face launched the Open FinLLM Leaderboard, a zero‑shot benchmark covering 40 finance‑specific tasks across seven categories to evaluate LLM readiness for real‑world financial applications.
OpenAI Canvas beta launch: collaborative writing and coding interface for ChatGPT
OpenAI introduced Canvas, a beta interface that lets ChatGPT collaborate on writing and coding projects with inline editing, version control, and specialized shortcuts, initially rolling out to Plus and Team users.
OpenAI Establishes $4 Billion Credit Facility for Financial Flexibility
OpenAI has established a $4 billion revolving credit facility with a consortium of global banks to increase liquidity and support the scaling of AI research and infrastructure.
Chinese AI Global Expansion Analysis
Chinese AI companies are accelerating international expansion due to domestic market saturation, intense price wars, and regulatory pressures, targeting Southeast Asia, the Middle East, and Western consumer markets.
OpenAI Funding Announcement October 2024
OpenAI has raised $6.6 billion in new funding at a $157 billion post-money valuation to accelerate frontier AI research and increase compute capacity.
OpenAI Realtime API Release
OpenAI has launched the Realtime API in public beta, enabling developers to build low-latency, multimodal speech-to-speech experiences using GPT-4o.
GPT-4o Vision Fine-Tuning API Release
OpenAI has introduced vision fine-tuning for GPT-4o, allowing developers to customize the model with image-text datasets to improve specialized visual understanding and object detection.
OpenAI Prompt Caching API Release
OpenAI has introduced Prompt Caching for GPT-4o, GPT-4o mini, o1-preview, and o1-mini, offering a 50% discount and reduced latency for reused input tokens.
OpenAI Model Distillation API Integration
OpenAI has introduced an integrated Model Distillation suite to allow developers to use outputs from frontier models like o1-preview and GPT-4o to fine-tune and improve the performance of smaller models like GPT-4o mini.
OpenAI and Altera: Creating Collaborative Digital Humans with GPT-4o
Altera has developed autonomous AI agents, termed digital humans, that can collaborate with people in environments like Minecraft using a brain-inspired architecture powered by GPT-4o.
BenCzechMark: A Comprehensive Evaluation Suite for Czech LLMs
Hugging Face and academic partners have released BenCzechMark, the first comprehensive evaluation suite for Czech language models, featuring 50 tasks across 9 categories and a novel duel-based scoring mechanism.
OpenAI Disrupts STORM-0817 Iran-Linked Malware Activity
OpenAI disabled accounts used by the Iran-based threat actor STORM-0817 to develop Android malware, scrape Instagram profiles, and perform reconnaissance on Pakistani cybersecurity professionals.
OpenAI Disrupts SweetSpecter China-Linked Cyber Activity
OpenAI identified and banned accounts linked to the China-based adversary SweetSpecter, who attempted to use ChatGPT for offensive cyber operations and targeted OpenAI employees with spear phishing attacks.
OpenAI Investigation: Fake Russian Troll Error Message Hoax
OpenAI has debunked a viral post claiming to expose a Russian troll account's GPT-4o error message, revealing the incident was a manually created hoax likely originating in the United States.
OpenAI Disrupts CyberAv3ngers Iran-linked Cyber Research Activity
OpenAI has banned accounts linked to the Iran-affiliated threat actor CyberAv3ngers, who used LLMs for reconnaissance, code debugging, and vulnerability research targeting industrial control systems.
OpenAI Disrupts Operation Stop News Russian Influence Activity
OpenAI banned a cluster of ChatGPT accounts used by a Russia-origin influence operation, dubbed Stop News, which used AI-generated text and images to mimic news outlets and establish deceptive partnerships.
OpenAI Disrupts Operation A2Z Multilingual Influence Activity
OpenAI banned a cluster of accounts using its API to run a multilingual influence operation, dubbed Operation A2Z, which leveraged AI to manage fake personas and generate political content across X and Facebook.
OpenAI Disrupts Bet Bot Gambling Spam Network
OpenAI banned a set of accounts using its API via an Israel-based startup to run a gambling spam network on X, utilizing AI-generated personas to lure users to gambling sites.
OpenAI Disrupts Operation STORM-2035 Iranian Influence Activity
OpenAI banned ChatGPT accounts used by an Iranian-origin actor, known as Storm-2035, to generate deceptive long-form articles and social media comments targeting the U.S. election and other global political issues.
OpenAI Disrupts Rwandan Election Political Commenting Network
OpenAI banned a network of ChatGPT accounts in Rwanda used to generate partisan election content and attempt to manipulate X trends through high-volume comment spamming.
OpenAI Tort Report: Disrupting Abusive Reporting Activity
OpenAI banned accounts involved in 'Tort Report,' a Category 1 influence operation that used AI to generate generic reports against independent Vietnamese media outlets on Facebook and YouTube.
OpenAI Disrupts 'Corrupt Comment' Influence Operation
OpenAI banned a cluster of API activity used to generate English-language comments on X targeting the Anti-Corruption Foundation and Alexei Navalny's associates.
Converting Vertex-Colored Meshes to Textured Meshes
Hugging Face introduces a method and the InstantTexture library to convert vertex-colored 3D meshes into UV-mapped, textured meshes for better application compatibility.
OpenAI Moderation API Update: omni-moderation-latest Model Release
OpenAI has released omni-moderation-latest, a GPT-4o-based multimodal moderation model that improves harm detection accuracy across text and images, particularly for non-English languages.
Minnesota Enterprise Translation Office ChatGPT Integration
The State of Minnesota's Enterprise Translation Office has integrated ChatGPT to accelerate government translation services, reducing turnaround times from weeks to under 48 hours and saving over $100,000 per month.
OpenAI and GEDI Strategic Partnership for Italian News Content
OpenAI and GEDI have partnered to integrate high-quality Italian-language news content from publications like La Repubblica and La Stampa into ChatGPT and SearchGPT.
Llama 3.2 Release Notes: Multimodal Vision and On-Device Small Language Models
Meta has released Llama 3.2, introducing multimodal vision capabilities in 11B and 90B sizes and lightweight 1B and 3B text-only models optimized for on-device deployment.
Mercado Libre Verdi AI Development Platform Launch
Mercado Libre unveiled Verdi, an AI development platform built on GPT‑4o that lets its developers create secure, autonomous LLM applications, starting with AI‑driven customer‑service mediation handling 10% of disputes.
Hugging Face Daily Papers Features Guide
Hugging Face's Daily Papers page provides a community-curated hub for AI research, featuring tools for author claiming, paper submission, and direct interaction between researchers and developers.
FineVideo Dataset Release
Hugging Face has released FineVideo, a high-quality open video dataset containing 43k videos (3.4k hours) with rich, structured annotations for video understanding and generative AI training.
Optimizing and Deploying Hugging Face Models with Optimum-Intel and OpenVINO GenAI
Hugging Face and Intel provide a streamlined workflow using Optimum-Intel and OpenVINO GenAI to optimize and deploy Transformers models on Intel hardware, specifically targeting edge and client-side C++ and Python environments.
Genmab "AI Everywhere" rollout: enterprise ChatGPT expansion and 100+ custom GPTs for biotech
Genmab expanded ChatGPT Enterprise to over 2,000 staff and deployed 100+ custom GPTs, saving each user 3.5 hours weekly and accelerating biotech research and documentation.
Qwen2.5 Release Notes: New Foundation, Coder, and Math Models
Qwen has released Qwen2.5, a comprehensive suite of open-weight dense decoder-only models including general-purpose LLMs, specialized Coder and Math variants, and an updated Qwen2-VL-72B.
Qwen2.5 LLM series release
Qwen announced the Qwen2.5 series – open-source decoder-only LLMs from 0.5B to 72B parameters that double the capability of Qwen2 while adding a larger 18‑trillion‑token dataset, major gains in knowledge, coding, math, and alignment, and a 128K token context window.
Qwen2.5-Coder Release Notes
Qwen has released Qwen2.5-Coder, a series of open-source coding models trained on 5.5 trillion tokens that outperform larger models in code generation, reasoning, and mathematics.
Qwen2.5-Math Release Notes
Qwen has released Qwen2.5-Math, a series of open-source mathematical LLMs supporting bilingual reasoning and Tool-Integrated Reasoning (TIR) to outperform leading closed-source models on complex math benchmarks.
Fine-tuning LLMs to 1.58-bit: Extreme Quantization with BitNet
Hugging Face demonstrates that existing LLMs, such as Llama 3 8B, can be fine-tuned to 1.58-bit ternary precision using a dynamic warmup quantization strategy, significantly reducing memory and energy requirements while maintaining strong performance.
Arco Educação and OpenAI Partner to Enhance Teaching in Brazil
Arco Educação is partnering with OpenAI to implement GPT-4 powered tools, specifically a Teacher Assistant, to reduce administrative burdens and create personalized lesson plans for students with diverse learning needs in Brazil.
Hugging Face SQL Console for Datasets
Hugging Face has introduced a browser-based SQL Console powered by DuckDB WASM that allows users to query, filter, and transform datasets directly on the Hub without backend dependencies.
OpenAI Safety and Security Practices Update
OpenAI has established an independent Board oversight committee to govern critical safety and security measures for model development and deployment.
HuggingChat Community Tools Release
Hugging Face has introduced Community Tools on HuggingChat, allowing users to integrate any public Hugging Face Space as a tool for LLMs to use directly within the chat interface.
Accelerate 1.0.0 Release Candidate Announcement
Accelerate 1.0.0 release candidates add FP8, DeepSpeed multi‑model, torch.compile, and new data‑loader/pipeline features while stabilizing the API for large‑scale training and inference.
OpenAI o1-preview release notes / what's new
OpenAI has released o1-preview and o1-mini, a new series of reasoning models designed to solve complex problems in science, coding, and mathematics through extended thinking time.
OpenAI Learning to Reason with LLMs
OpenAI demonstrates the reasoning capabilities of its latest models through a detailed walkthrough of a complex cipher decoding task, illustrating the step-by-step logical progression required to solve it.