OpenAI Spring Update: Introducing GPT-4o
OpenAI has released GPT-4o, a new flagship model capable of real-time reasoning across audio, vision, and text, while expanding advanced capabilities to ChatGPT free users.
OpenAI GPT-4o Release and ChatGPT Free Tier Enhancements
OpenAI has launched GPT-4o, a flagship model providing GPT-4 level intelligence with improved speed and multimodal capabilities across text, voice, and vision, now available to both free and paid users.
Transformers Agents 2.0 release notes / what's new
Hugging Face introduces Transformers Agents 2.0, a modular agent framework featuring iterative problem-solving agents and high performance on the GAIA benchmark using Llama-3-70B-Instruct.
Qwen-Max-0428 Release Notes
Qwen has released Qwen-Max-0428, an instruction-tuned chat model that outperforms Qwen1.5-110B-Chat on MT-Bench and ranks in the top 10 of the Chatbot Arena leaderboard.
Building Cost-Efficient Enterprise RAG Applications with Intel Gaudi 2 and Intel Xeon
Hugging Face demonstrates how combining Intel Gaudi 2 accelerators and Intel Xeon CPUs via the OPEA framework allows enterprises to build RAG applications with superior performance-per-dollar compared to H100-based systems.
Hugging Face Enterprise Hub now available on AWS Marketplace
Hugging Face has integrated Enterprise Hub with the AWS Marketplace, allowing organizations to upgrade their accounts and manage billing directly through their AWS accounts.
OpenAI Model Spec
OpenAI has introduced the Model Spec, a framework of objectives, rules, and default behaviors designed to standardize and transparently shape how AI models behave in the OpenAI API and ChatGPT.
OpenAI Approach to Data and AI
OpenAI is introducing Media Manager to allow creators to control their content's use in AI training and detailing its multi-pronged data sourcing strategy to balance model diversity with creator rights.
OpenAI Content Provenance and Authenticity Initiatives
OpenAI is adopting the C2PA open standard for digital content certification and developing internal tools like image detection classifiers and audio watermarking to help users identify AI-generated content.
OpenAI and Stack Overflow API Partnership
OpenAI and Stack Overflow have partnered to integrate vetted technical data via OverflowAPI into OpenAI models and ChatGPT, while Stack Overflow will use OpenAI models to develop OverflowAI.
Hugging Face Open Leaderboard for Hebrew LLMs
Hugging Face has launched an open leaderboard specifically designed to evaluate and improve Large Language Models (LLMs) in Hebrew, addressing the challenges of low-resource and morphologically complex languages.
Artificial Analysis LLM Performance Leaderboard on Hugging Face
Hugging Face has integrated the Artificial Analysis LLM Performance Leaderboard, providing AI engineers with a unified metric system for comparing the quality, price, and speed of over 100 serverless LLM API endpoints.
Hugging Face Inference Endpoints ASR and Diarization Pipeline
Hugging Face introduces a custom inference handler for deploying a modular pipeline combining Whisper ASR, Pyannote diarization, and speculative decoding on Inference Endpoints.
OpenAI Disrupts "Bad Grammar" Russian-Linked Influence Operation
OpenAI banned a Russian-linked network dubbed "Bad Grammar" that used AI models to automate the generation of political comment spam on Telegram across multiple countries.
OpenAI Disrupts China-linked Influence Operation "Spamouflage"
OpenAI banned accounts associated with the operation known as Spamouflage, a China-linked influence campaign that used AI to generate content and conduct research for targeting critics of the Chinese government.
OpenAI Disrupts Russian 'Doppelganger' Influence Operation
OpenAI banned four clusters of accounts linked to the Russian 'Doppelganger' influence operation that used AI to generate multilingual content targeting Europe and North America.
OpenAI Operation Zero Zeno: Israel-linked Influence Activity
OpenAI banned a cluster of accounts operated by the Israeli political campaign firm STOIC to disrupt an AI-generated influence operation targeting multiple countries.
OpenAI Disrupts Iran-linked Influence Operation IUVM
OpenAI banned accounts associated with IUVM, an Iran-linked influence network that used AI to generate and translate pro-Iran, anti-US, and anti-Israel content.
Improving Prompt Consistency with Structured Generations
Hugging Face and Dottxt research demonstrates that using structured generation to constrain LLM outputs reduces performance variance and improves ranking consistency across different prompt formats and shot orders.
StarCoder2-15B-Instruct-v0.1 release notes / what's new
Hugging Face introduces StarCoder2-15B-Instruct-v0.1, the first entirely self-aligned code LLM trained with a fully transparent and permissive pipeline that outperforms CodeLlama-70B-Instruct on HumanEval.
OpenAI and Financial Times Strategic Partnership
OpenAI and the Financial Times have entered a strategic partnership and licensing agreement to integrate attributed FT journalism into ChatGPT and provide FT employees with ChatGPT Enterprise access.
Qwen1.5-110B release notes / what's new
Qwen has released Qwen1.5-110B, the first model in the Qwen1.5 series to exceed 100 billion parameters, delivering competitive performance against Llama-3-70B and significant improvements over Qwen1.5-72B.
Moderna and OpenAI Partnership for AI-Driven Drug Development
Moderna has deployed ChatGPT Enterprise across its organization to accelerate mRNA medicine development and optimize business operations through custom GPTs.
OpenAI Introducing ChatGPT and Whisper APIs
OpenAI has released the GPT-3.5 Turbo and Whisper APIs, offering developers significant cost reductions and high-performance speech-to-text capabilities.
GPT-4 API General Availability and Completions API Deprecation
OpenAI has made the GPT-4 API generally available to all paying customers and announced a deprecation plan for older models in the Completions, Embeddings, and Edits APIs.
OpenAI Child Safety Commitment: Adopting Safety by Design Principles
OpenAI and other industry leaders have committed to implementing Safety by Design principles to mitigate generative AI risks to children, focusing on the prevention of child sexual abuse material (CSAM) and exploitation material (CSEM).
OpenAI Enterprise API Feature Updates April 2024
OpenAI has introduced new enterprise-grade security, administrative controls, Assistants API enhancements, and cost-management options to support scaling AI solutions for large organizations.
Hugging Face Open Chain of Thought Leaderboard
Hugging Face has introduced the Open Chain of Thought Leaderboard to measure the specific accuracy gain provided by chain-of-thought prompting across various LLMs on challenging reasoning tasks.
Jack of All Trades (JAT) Multi-Purpose Transformer Agent
Hugging Face introduces Jack of All Trades (JAT), a single transformer-based agent capable of performing diverse sequential decision-making tasks across Atari, BabyAI, Meta-World, and MuJoCo environments.
OpenAI The Instruction Hierarchy: Prioritizing Privileged Instructions to Prevent Prompt Injection
OpenAI has introduced an instruction hierarchy that trains LLMs to prioritize system prompts over untrusted user input, significantly reducing vulnerability to prompt injections and jailbreaks.
The Open Medical-LLM Leaderboard: Benchmarking Large Language Models in Healthcare
Hugging Face has introduced the Open Medical-LLM Leaderboard, a standardized platform to evaluate and compare the performance of LLMs across diverse medical datasets to improve reliability and patient safety.
Meta Llama 3 Release Notes
Meta has released Llama 3, an open-access LLM family featuring 8B and 70B parameter models with improved tokenization and training on 15 trillion tokens.
CodeQwen1.5 Release Notes
Qwen has released CodeQwen1.5-7B, an open-source code LLM supporting 92 programming languages and 64K token context windows to enhance developer productivity.
Ryght Case Study: Building a Life Sciences Generative AI Platform with Hugging Face
Ryght has launched Ryght Preview, an enterprise-grade generative AI platform for healthcare and life sciences that leverages Hugging Face's Expert Support, TGI, and TEI to provide secure, flexible, and high-performance AI copilots.
Running Privacy-Preserving Inferences on Hugging Face Endpoints
Hugging Face and Zama have enabled the deployment of Fully Homomorphic Encryption (FHE) models via Hugging Face Endpoints, allowing users to perform machine learning inferences on encrypted data without decrypting it.
LiveCodeBench Leaderboard: Contamination-Free Evaluation for Code LLMs
Hugging Face has introduced the LiveCodeBench leaderboard, a new benchmark developed by researchers from UC Berkeley, MIT, and Cornell to evaluate LLM code generation and reasoning capabilities while preventing benchmark contamination using time-windowed problem sets.
Gradio Reload Mode for Faster AI App Development
Gradio's reload mode enables developers to apply source code changes to AI applications instantly without restarting the server, significantly reducing development latency.
Idefics2 8B Vision-Language Model Release – Architecture, Data, and Performance
Hugging Face released Idefics2, an 8B open‑source vision‑language model that outperforms other 8‑B models on VQA and OCR benchmarks and is ready for fine‑tuning via 🤗 Transformers.
OpenAI Japan Launch and GPT-4 Japanese Custom Model
OpenAI has opened its first Asian office in Tokyo and released a GPT-4 custom model optimized for the Japanese language that operates up to 3x faster than GPT-4 Turbo.
Vision Language Models Explained
Hugging Face provides a comprehensive guide to Vision Language Models (VLMs), detailing their architecture, open-source options, evaluation benchmarks, and new experimental support for fine-tuning via the TRL library.
Hugging Face and Google Cloud Vertex AI Model Garden Integration
Hugging Face has launched 'Deploy on Google Cloud,' enabling users to deploy thousands of open foundation models to Vertex AI or Google Kubernetes Engine (GKE) via the Hugging Face Hub or Vertex Model Garden.
CodeGemma Release Notes / What's New
Google has released CodeGemma, a family of open-access code-specialist LLMs based on Gemma, trained on 500 billion additional tokens of code, mathematics, and English language data.
Hugging Face Public Policy Program Overview and Submitted Materials
Hugging Face announced a comprehensive public‑policy program that provides U.S., EU, and U.K. policymakers with detailed position papers, testimony, and comment letters, reflecting its cross‑functional commitment to responsible openness and shaping AI regulation.
Klarna AI Assistant Integration and Performance Results
Klarna has deployed an OpenAI-powered AI assistant that handles two-thirds of its customer service chats, performing the work of 700 full-time agents while improving resolution times and profit.
Hugging Face and Wiz Research Partnership for AI Security
Hugging Face has partnered with Wiz to integrate advanced vulnerability management and cloud security posture management to protect its platform and the broader AI/ML ecosystem.
Text2SQL with Hugging Face Dataset Viewer API and DuckDB-NSQL-7B
Hugging Face demonstrates how to use the DuckDB-NSQL-7B model and the Dataset Viewer API to convert natural language questions into SQL queries for analyzing over 120,000 open datasets.
OpenAI Fine-Tuning API Improvements and Custom Models Program Expansion
OpenAI has introduced new control features for its fine-tuning API and expanded its Custom Models program to include assisted fine-tuning and fully custom-trained models for domain-specific needs.
SetFit Inference Acceleration with 🤗 Optimum Intel on Xeon
Hugging Face demonstrates how to achieve up to 7.8x faster inference throughput for SetFit models on Intel Xeon CPUs using post-training static quantization via the 🤗 Optimum Intel library.
Qwen1.5-32B release notes / what's new
Qwen has released Qwen1.5-32B and Qwen1.5-32B-Chat, models designed to balance high performance with lower memory and inference costs compared to the 72B version.
Hugging Face and Cloudflare Workers AI Integration
Hugging Face integrated Cloudflare Workers AI to provide serverless GPU inference for popular open models, allowing developers to deploy AI applications with a pay-per-request pricing model.