7001

Cosmopedia: Large-Scale Synthetic Data for LLM Pre-training

Hugging Face introduces Cosmopedia, the largest open synthetic dataset for LLM pre-training, containing 30 million files and 25 billion tokens generated by Mixtral-8x7B-Instruct-v0.1.

7002

Phi-2 on Intel Meteor Lake: Local LLM Inference

Hugging Face demonstrates how to run the Microsoft Phi-2 model locally on Intel Meteor Lake (Core Ultra) processors using 4-bit quantization via OpenVINO and Optimum Intel.

7003

Holiday Extras ChatGPT Enterprise Implementation

Holiday Extras deployed ChatGPT Enterprise across its organization, saving over 500 hours per week and achieving an estimated $500k in annual savings.

7004

OpenAI and Salesforce Integration for Enterprise Trust and Safety

OpenAI provides the foundational generative AI models for the Salesforce Einstein AI Platform, enabling secure, enterprise-ready AI capabilities across Sales, Service, and Commerce Clouds.

7005

Superhuman AI Integration with OpenAI API

Superhuman has integrated OpenAI's API to launch a suite of AI-powered email features that double the speed of inbox processing for its users.

7006

Quanto: a PyTorch quantization backend for Optimum

Hugging Face introduces Quanto, a versatile and device-agnostic PyTorch quantization backend for Optimum designed to simplify low-precision model deployment across any modality.

7007

Hugging Face Train on DGX Cloud

Hugging Face launched Train on DGX Cloud, a no-code service for Enterprise Hub organizations to fine-tune open models using NVIDIA H100 and L40S GPUs.

7008

WebSight Dataset and Sightseer Model

Hugging Face introduced WebSight, a synthetic dataset of 2 million screenshot-to-HTML pairs, and Sightseer, a vision-language model capable of converting web screenshots into functional HTML code.

7009

CPU Optimized Embeddings with Optimum Intel and fastRAG

Hugging Face and Intel have introduced a method to accelerate embedding models on Xeon CPUs using Optimum Intel and fastRAG, achieving up to 4.5x latency reduction and 4x throughput improvement via int8 quantization.

7010

OpenAI Global News Partnerships with Le Monde and Prisa Media

OpenAI has partnered with Le Monde and Prisa Media to integrate their news content into ChatGPT, providing users with attributed summaries and direct links to original articles.

7011

Healthify and OpenAI Collaboration for AI Health Coaching

Healthify has integrated OpenAI's GPT-4 Vision, GPT-4 Turbo, and Embeddings models to scale personalized nutrition and fitness coaching, resulting in a 50% increase in food tracking engagement.

7012

OpenAI Governance Review and Leadership Confirmation

OpenAI has completed an independent review by WilmerHale, confirming full confidence in Sam Altman and Greg Brockman's leadership and implementing significant governance reforms.

7013

OpenAI Board of Directors Expansion

OpenAI has appointed Dr. Sue Desmond-Hellmann, Nicole Seligman, and Fidji Simo to its Board of Directors, while CEO Sam Altman has rejoined the board.

7014

Lifespan Healthcare System Uses GPT-4 to Improve Patient Health Literacy

Lifespan, Rhode Island's largest healthcare system, has deployed GPT-4 to simplify complex surgical consent forms from a college reading level to a middle school level to improve patient understanding and well-being.

7015

Match Group Adopts ChatGPT Enterprise to Enhance Productivity and Product Development

Match Group has deployed ChatGPT Enterprise to thousands of employees to improve cross-functional collaboration, engineering efficiency, and product usability testing, while leveraging OpenAI's API for customer-facing features.

7016

Paradigm Integrates GPT-4 to Accelerate Clinical Trial Patient Matching

Paradigm uses OpenAI's GPT-4 API to automate the evaluation of medical records for clinical trial eligibility, increasing accuracy by 10% and enabling the screening of hundreds of patients per minute.

7017

OpenAI and Elon Musk: Clarification of Partnership and Mission

OpenAI released a statement clarifying its historical relationship with Elon Musk, the necessity of its for-profit transition to fund AGI development, and its commitment to broad benefit.

7018

ConTextual: Benchmarking Multimodal Reasoning in Text-Rich Scenes

Hugging Face and UCLA researchers have introduced ConTextual, a dataset and leaderboard designed to evaluate how Large Multimodal Models (LMMs) jointly reason over text and visual cues in complex, text-rich images.

7019

Hugging Face and Argilla Enable Collective Community Dataset Building

Hugging Face and Argilla have introduced a streamlined workflow using Hugging Face Spaces and Argilla to allow communities to collectively build high-quality, open-source datasets.

7020

Text-Generation Pipeline on Intel Gaudi 2 AI Accelerator

Hugging Face introduces a custom text-generation pipeline for Intel Gaudi 2 AI accelerators, enabling streamlined deployment of Llama 2 models (7b, 13b, and 70b) via Optimum Habana.

7021

StarCoder2 and The Stack v2 Release

BigCode has released StarCoder2, a family of transparently trained open code LLMs in 3B, 7B, and 15B parameter sizes, powered by the massive new Stack v2 dataset.

7022

Hugging Face TTS Arena: Benchmarking Text-to-Speech Models

Hugging Face has launched the TTS Arena, a crowdsourced, side-by-side comparison tool and leaderboard using an Elo rating system to objectively measure text-to-speech model quality.

7023

AI Watermarking 101: Tools and Techniques

Hugging Face provides a comprehensive overview of AI watermarking techniques across images, text, and audio to combat deepfakes and ensure content provenance.

7024

Introduction to Matryoshka Embedding Models

Matryoshka Embedding models allow for variable-size embeddings that can be truncated without significant performance loss, enabling a flexible trade-off between storage, speed, and accuracy.

7025

Fine-Tuning Gemma Models in Hugging Face

Hugging Face provides a guide on using Parameter-Efficient Fine-Tuning (PEFT) and Low-Rank Adaptation (LoRA) to customize Google's Gemma models on GPUs and Cloud TPUs.

7026

Hugging Face and Haize Labs Introduce Red-Teaming Resistance Leaderboard

Hugging Face and Haize Labs have launched the Red-Teaming Resistance (RTR) Benchmark to evaluate LLM robustness against high-quality, human-like adversarial prompts across specific safety violation categories.

7027

Google Gemma Open LLM Release

Google has released Gemma, a family of open-access large language models based on Gemini, available in 2B and 7B parameter sizes with base and instruction-tuned variants.

7028

Open Ko-LLM Leaderboard

Hugging Face and Upstage have launched the Open Ko-LLM Leaderboard to provide a fair, transparent evaluation ecosystem for Korean Large Language Models using private test sets to prevent contamination.

7029

Hugging Face PEFT New LoRA Merging Methods

Hugging Face has introduced several new merging methods to the PEFT library, enabling users to combine multiple LoRA adapters from the same base model on the fly to synthesize new capabilities.

7030

Synthetic Data with Open-Source LLMs Cuts Cost, Latency, and Carbon for Custom Models

Hugging Face shows how using open‑source LLMs to generate synthetic data and fine‑tune a small RoBERTa model reduces inference cost from $3061 to $2.7, latency from seconds to 0.13 s, and CO₂ emissions from ~1 t to 0.12 kg while matching GPT‑4 accuracy on financial sentiment classification.

7031

OpenAI Sora: Video Generation Models as World Simulators

OpenAI has introduced Sora, a diffusion transformer model capable of generating high-fidelity video up to one minute long, suggesting that scaling video generation is a path toward general-purpose physical world simulators.

7032

OpenAI Disrupts State-Affiliated Threat Actors Using AI

OpenAI, in partnership with Microsoft Threat Intelligence, terminated accounts of five state-affiliated threat actors from China, Iran, North Korea, and Russia who used AI for limited cybersecurity tasks.

7033

AMD Pervasive AI Developer Contest

AMD and Hugging Face have partnered to launch the Pervasive AI Developer Contest, offering developers free access to AMD hardware and cash prizes to build AI applications in Generative AI, Robotics AI, and PC AI.

7034

ChatGPT Memory and New User Controls

OpenAI has introduced a memory feature for ChatGPT that allows the model to remember details across conversations to provide more personalized responses without requiring repeated information.

7035

Hugging Face TGI Messages API Release

Hugging Face has introduced a Messages API for Text Generation Inference (TGI) starting with version 1.4.0, enabling OpenAI Chat Completion API compatibility for open LLMs.

7036

Qwen1.5 Release Notes / What's New

Qwen has released Qwen1.5, a series of open-source base and chat models ranging from 0.5B to 110B parameters, featuring improved human alignment, multilingual capabilities, and native Hugging Face transformers integration.

7037

SegMoE: Segmind Mixture of Diffusion Experts

SegMoE is a framework for creating Mixture-of-Experts (MoE) Diffusion models by replacing Feed-Forward layers in Stable Diffusion architectures with sparse MoE layers to improve prompt understanding.

7038

NPHardEval Leaderboard: Evaluating LLM Reasoning via Computational Complexity

Hugging Face introduces the NPHardEval leaderboard, a dynamic benchmark that uses computational complexity classes to quantitatively measure the logical reasoning abilities of Large Language Models.

7039

PatchTST Integration in Hugging Face

Hugging Face has integrated PatchTST, a Transformer-based model that uses time series patching and channel-independence to improve long-term forecasting and enable transfer learning.

7040

Hugging Face Text Generation Inference now supports AWS Inferentia2

Hugging Face announced the general availability of Text Generation Inference on AWS Inferentia2 via Amazon SageMaker, enabling cost‑effective, high‑throughput LLM serving as an alternative to GPU deployments.

7041

Constitutional AI with Open LLMs

Hugging Face introduces an end-to-end recipe and the llm-swarm tool to implement Constitutional AI (CAI) on open models, enabling scalable self-alignment based on user-defined principles without expensive human feedback.

7042

OpenAI Biological Threat Evaluation: Building an Early Warning System

OpenAI conducted a human-centric study to determine if GPT-4 increases the ability of users to create biological threats compared to internet-only access, finding mild but not statistically significant uplifts in accuracy and completeness.

7043

Enterprise Scenarios Leaderboard: Evaluating LLMs for Real-World Use Cases

Hugging Face and Patronus AI have launched the Enterprise Scenarios Leaderboard to evaluate language models on six real-world business tasks, moving beyond academic benchmarks to measure practical enterprise utility.

7044

Accelerating StarCoder on Intel Xeon with Optimum Intel

Hugging Face and Intel demonstrate over 7x inference acceleration for the StarCoder-15B model on 4th Gen Intel Xeon processors by combining 8-bit quantization and assisted generation.

7045

Hugging Face Hallucinations Leaderboard launch and initial findings

Hugging Face launched the Hallucinations Leaderboard to benchmark LLMs on factuality and faithfulness errors across multiple open-source datasets, offering transparent rankings that guide model selection and research.

7046

AI Secure LLM Safety Leaderboard

Hugging Face and the Secure Learning Lab have released the LLM Safety Leaderboard, powered by the DecodingTrust framework to evaluate LLM trustworthiness across eight critical safety dimensions.

7047

OpenAI New Embedding Models and API Updates January 2024

OpenAI has released new text-embedding-3-small and text-embedding-3-large models with improved performance and lower pricing, alongside updates to GPT-4 Turbo and GPT-3.5 Turbo.

7048

Qwen-VL-Plus and Qwen-VL-Max Release

Qwen has released Qwen-VL-Plus and Qwen-VL-Max, large visual language models that match GPT-4V and Gemini Ultra in multimodal tasks and outperform them in Chinese text comprehension.

7049

Hugging Face and Google Cloud Strategic Partnership

Hugging Face and Google Cloud have entered a strategic partnership to democratize machine learning by integrating open models with Google Cloud's AI infrastructure and hardware.

7050

Open-source LLMs as LangChain Agents

Hugging Face demonstrates that open-source LLMs, specifically Mixtral-8x7B, are now capable of powering agent workflows and can outperform GPT-3.5 in general-purpose reasoning tasks.