1801

OFASys: A Framework for Multimodal Multitask Learning

Qwen introduces OFASys, an AI framework that simplifies multimodal multitask learning by allowing users to define complex tasks and modalities via a single-line Instruction interface.

1802

Chinese CLIP: Contrastive Vision-Language Pretraining in Chinese

Qwen has released Chinese CLIP, a vision-language model designed to overcome the cultural and linguistic limitations of English-centric CLIP models in cross-modal retrieval and image classification.

1803

OpenAI Applied AI Research: Perspectives on AGI, Safety, and Continuous Learning

Lilian Weng of OpenAI's Applied AI Research team discusses the path toward AGI, the critical importance of model alignment and safety, and the role of continuous learning in AI development.

1804

Zero-shot image segmentation with CLIPSeg

Hugging Face introduces CLIPSeg, a zero-shot image segmentation model that uses CLIP embeddings to create segmentation masks from either text or image prompts without requiring category-specific training.

1805

Hugging Face Model Cards Documentation Framework

Hugging Face has released a suite of tools and resources, including a GUI-based creator tool and a standardized template, to improve the accessibility and standardization of machine learning model documentation.

1806

OpenAI Point-E: Generating 3D Point Clouds from Text Prompts

OpenAI has introduced Point-E, a system that generates 3D point clouds from complex text prompts in 1-2 minutes on a single GPU, prioritizing sampling speed over absolute sample quality.

1807

OpenAI text-embedding-ada-002 release

OpenAI announced the text-embedding-ada-002 model, a unified, faster, cheaper embedding model that outperforms previous models on most tasks.

1808

Hugging Face Audio Datasets Guide – How to Load, Process, and Stream Audio Data

Hugging Face announced a comprehensive guide showing that the 🤗 Datasets library can load, preprocess, and stream any audio dataset from the Hub with just a few lines of Python code, enabling efficient research on speech and audio tasks.

1809

Hugging Face Ethics and Society Newsletter #2: Addressing Bias in Machine Learning

Hugging Face outlines a sociotechnical framework for mitigating machine learning bias by treating biases as risk factors that must be addressed across task definition, dataset curation, and model training.

1810

Habana Gaudi2 vs Nvidia A100 80GB Performance Benchmarks

Hugging Face benchmarks show that Habana Gaudi2 provides approximately twice the throughput of Nvidia A100 80GB for both training and inference across BERT, Stable Diffusion, and T5-3B models.

1811

Hugging Face Announces Bumblebee: Transformers and Stable Diffusion in Pure Elixir

Hugging Face released Bumblebee, a pure‑Elixir implementation of Transformers that brings models from GPT‑2 to Stable Diffusion to the Elixir ecosystem, enabling native CPU/GPU inference without external dependencies.

1812

Illustrating Reinforcement Learning from Human Feedback (RLHF)

Hugging Face explains Reinforcement Learning from Human Feedback (RLHF), a three-step process used to align large language models with complex human values by optimizing them using a reward model based on human preferences.

1813

OpenAI Supercomputing Infrastructure and Backend Systems Engineering

OpenAI engineer Jake Stangel describes the technical challenges of managing billion-dollar supercomputing clusters where extreme scale reveals hardware and kernel issues unseen by other users.

1814

Deep Learning with Proteins – Hugging Face guide to protein language models and folding

Hugging Face announced a tutorial series showing how to fine‑tune protein language models and use ESMFold for protein folding, demonstrating that transfer learning techniques from NLP can be applied directly to protein sequence tasks.

1815

Time Series Transformer probabilistic forecasting with 🤗 Transformers

Hugging Face released a vanilla Transformer model for global probabilistic time‑series forecasting, demonstrating state‑of‑the‑art performance on the Tourism Monthly benchmark.

1816

Stable Diffusion Core ML on Apple Silicon – How to Run and Optimize

Hugging Face released Core ML‑converted Stable Diffusion checkpoints for Apple Silicon, enabling on‑device image generation in Python or Swift with up to 18 seconds per image on an M1 Max.

1817

OpenAI Introducing ChatGPT

OpenAI has released ChatGPT, a conversational AI model based on the GPT-3.5 series, trained using Reinforcement Learning from Human Feedback (RLHF) to interact in a dialogue format.

1818

VQ-Diffusion: Conditional Latent Diffusion in Discrete Space

VQ-Diffusion is a conditional latent diffusion model that operates on a quantized discrete latent space, offering faster inference and higher image quality than traditional autoregressive models.

1819

Hugging Face 2023 Internship Program

Hugging Face has announced its 2023 internship program, offering roles across Open Source, Science, and Social Impact teams to democratize responsible machine learning.

1820

Hugging Face Diffusion Models Class and Community Event

Hugging Face announced a free Diffusion Models Class launching November 28, 2022, accompanied by a live community event on November 30 featuring researchers from Stability AI, Meta, and Runway.

1821

Hugging Face Director of Machine Learning Insights Part 4

Four Machine Learning Directors share industry-specific insights on the impact, challenges, and integration pitfalls of ML in e-commerce, engineering, education, and SaaS.

1822

Hugging Face Inference Solutions Overview November 2022

Hugging Face announced a suite of free and paid inference options—including a widget, API, Inference Endpoints, and Spaces—to simplify model testing, deployment, and production scaling.

1823

Hugging Face Accelerating Document AI

Hugging Face provides a comprehensive guide to using open-source multimodal models to automate document classification, parsing, and visual question answering for enterprise workflows.

1824

Sentiment Analysis on Encrypted Data with Homomorphic Encryption

Hugging Face demonstrates how to use the Concrete-ML library to perform sentiment analysis on encrypted data using a combination of BERT transformers and XGBoost with Fully Homomorphic Encryption (FHE).

1825

Hugging Face and arXiv Integration for Machine Learning Demos

Hugging Face has integrated Hugging Face Spaces with arXivLabs to provide interactive machine learning demos directly on arXiv paper abstract pages.

1826

OFA: Towards Building a One-For-All Model

OFA is a unified multimodal pretrained model that unifies understanding and generation tasks across modalities into a single framework using instruction-based multitask pretraining.

1827

Hugging Face Pricing Update November 2022

Hugging Face has transitioned to a compute-based monetization model, sunsetting the Paid tier of the Inference API in favor of Inference Endpoints and hardware upgrades for Spaces.

1828

Contrastive Search for Human-Level Text Generation in Transformers

Hugging Face has integrated Contrastive Search into the transformers library, a decoding method that prevents model degeneration and maintains semantic coherence across 16 languages using off-the-shelf models.

1829

Dreambooth Stable Diffusion fine‑tuning guide with Diffusers

Hugging Face released detailed recommendations for training Stable Diffusion with Dreambooth using the Diffusers library, showing that low learning rates, enough steps, prior preservation for faces, and text‑encoder fine‑tuning yield the highest quality results.

1830

OpenAI DALL·E API public beta announcement

OpenAI announced that the DALL·E image generation API is now available in public beta, letting developers integrate state‑of‑the‑art image creation and built‑in moderation into their apps.

1831

Fine-Tuning Whisper for Multilingual ASR with Hugging Face Transformers

Hugging Face provides a comprehensive guide on fine-tuning OpenAI's Whisper model for multilingual automatic speech recognition (ASR), demonstrating a 31.5% absolute WER improvement on Hindi using only 8 hours of data.

1832

Hugging Face Optimum Intel and OpenVINO Integration

Hugging Face has integrated Intel OpenVINO into Optimum Intel, enabling accelerated inference and quantization for Transformer models on Intel hardware.

1833

Evaluating Language Model Bias with 🤗 Evaluate

Hugging Face added bias metrics—toxicity, language polarity, and HONEST—to the 🤗 Evaluate library, enabling systematic measurement of harmful language in causal language models.

1834

Distributed Training with PyTorch DDP, Accelerate, and Transformers Trainer

Hugging Face explains how to implement distributed training using three levels of abstraction: native PyTorch DDP, the Accelerate library, and the high-level Transformers Trainer API.

1835

OpenAI Scaling Laws for Reward Model Overoptimization

OpenAI identifies that optimizing against an imperfect proxy reward model eventually degrades ground truth performance, following a predictable scaling law based on model parameters and dataset size.

1836

MTEB: Massive Text Embedding Benchmark

Hugging Face introduced MTEB, a massive and multilingual benchmark consisting of 56 datasets across 8 tasks to evaluate the performance of text embedding models.

1837

Hugging Face Inference Endpoints

Hugging Face Inference Endpoints is a managed service that allows users to deploy machine learning models from the Hugging Face Hub to scalable, secure cloud infrastructure with a few clicks.

1838

Stable Diffusion JAX and Flax Integration

Hugging Face Diffusers version 0.5.1 introduces support for Flax, enabling high-speed Stable Diffusion inference on Google TPUs via JAX.

1839

Hugging Face BLOOM Inference Optimization

Hugging Face achieved a 5x reduction in latency and a 50x increase in throughput for the BLOOM model by transitioning from Pipeline Parallelism to Tensor Parallelism and implementing custom CUDA kernels.

1840

Hugging Face Introduces DOI Support for Models and Datasets

Hugging Face now lets users generate Digital Object Identifiers (DOIs) for Hub models and datasets, providing permanent, citable links that persist across versions.

1841

Japanese Stable Diffusion release by rinna

rinna released Japanese Stable Diffusion, a Japanese‑language fine‑tuned version of Stable Diffusion that generates culturally appropriate images from Japanese prompts.

1842

Hugging Face Zero-Shot Evaluation on the Hub

Hugging Face has introduced zero-shot evaluation for causal language models on the Hub, enabling users to benchmark models up to 66 billion parameters without writing code.

1843

OpenAI DALL·E Beta Now Available Without Waitlist

OpenAI has removed the waitlist for the DALL·E beta, allowing immediate sign-up for a system that currently supports over 1.5 million users creating more than 2 million images daily.

1844

Hugging Face AutoTrain Image Classification

Hugging Face has added Image Classification to AutoTrain, enabling users to train custom image categorization models without writing code or configuring hyperparameters.

1845

Hugging Face Accelerate: Running Large Models with PyTorch

Hugging Face Accelerate enables the execution of massive AI models on consumer hardware by leveraging PyTorch's meta device and sharded checkpoints to manage memory across GPUs, CPU RAM, and disk.

1846

SetFit: Efficient Few-Shot Learning Without Prompts

Hugging Face introduces SetFit, a prompt-free framework for few-shot fine-tuning of Sentence Transformers that achieves high accuracy with minimal labeled data.

1847

Hugging Face Ethics and Society Newsletter #1

Hugging Face introduces its Ethics and Society newsletter and outlines a decentralized, value-driven approach to operationalizing AI ethics through collaboration, transparency, and responsibility.

1848

OpenAI Whisper Release

OpenAI has released Whisper, an automatic speech recognition (ASR) system trained on 680,000 hours of multilingual and multitask supervised data to improve robustness across accents, noise, and languages.

1849

Incredibly Fast BLOOM Inference with DeepSpeed and Accelerate

Hugging Face demonstrates sub‑millisecond per‑token generation for the 176B‑parameter BLOOM model using DeepSpeed‑Inference tensor parallelism and Accelerate pipeline parallelism on 8×80 GB A100 GPUs.

1850

Diffusers 0.3 release adds image‑to‑image, textual inversion, inpainting, GPU optimizations, Mac MPS, ONNX support and new docs

Hugging Face announced Diffusers version 0.3, introducing image‑to‑image, textual inversion, experimental inpainting, smaller‑GPU optimizations, Mac MPS support, an ONNX exporter, expanded documentation, and a wave of community projects.