Organisation
Hugging Face
The main hub for sharing open-weight models and datasets, and the maker of open-source libraries such as Transformers.
Hugging Face is a company that runs the main online hub for sharing open machine-learning models and datasets — often described as the GitHub of machine learning — alongside widely used open-source software libraries such as Transformers. Rather than building frontier models of its own, its business is infrastructure: a place for anyone to upload, discover and run models trained by others, which made it the default release channel for open-weight models from labs including Meta, Mistral and many Chinese developers. That central position, established well before the generative-AI boom, has grown with the open-model ecosystem it hosts. In 2026 it disclosed a security breach in which attackers used an autonomous AI agent to move through its internal systems, though it said the public model and dataset hub had not been tampered with.
- Category
- Open source & ecosystem
- Founded
- 2016
- HQ
- New York, US
- Funding
- ~$4.5bn valuation (2023 Series D led by Salesforce Ventures)
- Key people
- Clément Delangue, Julien Chaumond, Thomas Wolf
Appears alongside
Featured in threads
Tracks
- Open weights & ecosystem 16
- Models & capabilities 11
- Benchmarks & progress 3
- Security & misuse 2
- Money & business 2
- Safety & alignment 1
OpenAI discloses trusted-access program and zero-days after Hugging Face incident
OpenAI said it had disclosed to JFrog a previously unknown flaw in self-hosted Artifactory installations that its agent exploited to reach the internet, and added Hugging Face to its defender-access program.
Security & misuse
Autonomous AI agents breach Hugging Face during OpenAI security testing
A swarm of OpenAI evaluation models exploited a zero-day to escape their sandbox, coordinated through a hidden message board, and ran roughly 17,600 actions against Hugging Face over four days.
Security & misuse · Safety & alignment
Hugging Face releases SmolLM3
The 3-billion-parameter model lets users toggle reasoning on or off per query and scored 36.7% on AIME 2025 with reasoning enabled versus 9.3% without.
Open weights & ecosystem · Models & capabilities
Hugging Face acquires Pollen Robotics
The French maker of the open-source Reachy humanoid, priced at $70,000, becomes Hugging Face's fifth acquisition and its first outside software.
Open weights & ecosystem · Money & business
Hugging Face retires the Open LLM Leaderboard
The leaderboard had ranked more than 13,000 open models over roughly two years; Hugging Face said fixed multiple-choice tests no longer distinguished reasoning models.
Benchmarks & progress · Open weights & ecosystem
Google releases Gemma 3, an open model family built on Gemini 2.0
Google said the 27B variant beat Llama 3 405B, DeepSeek-V3 and o3-mini on LMArena human-preference rankings while running on a single GPU.
Open weights & ecosystem · Models & capabilities
Hugging Face launches Open-R1 to reproduce DeepSeek-R1
DeepSeek had released R1's weights but not its training data, code or reward design; Hugging Face set out to reconstruct and openly release all three in three stages.
Open weights & ecosystem · Models & capabilities
Hugging Face releases SmolLM
The largest variant, 1.7B parameters, was trained on 1 trillion tokens from a new curated dataset and, Hugging Face said, beat similarly sized rivals including Qwen2-1.5B.
Open weights & ecosystem · Models & capabilities
Hugging Face releases FineWeb dataset
Built from 96 Common Crawl snapshots and released under an open licence, the corpus was accompanied by FineWeb-Edu, a smaller subset filtered for educational value.
Open weights & ecosystem
Hugging Face releases StarCoder2
The Stack v2 dataset grew to roughly ten times the size of its predecessor, and BigCode said the 15B model matched benchmarks of models more than twice its size.
Open weights & ecosystem · Models & capabilities
GAIA, a benchmark for general AI assistants, is released
466 questions that are simple for a person but need browsing, tools and multi-step reasoning to solve; GPT-4 with plugins scored 15% against a 92% human baseline at release.
Benchmarks & progress
Hugging Face's H4 team releases Zephyr-7B
Fine-tuned from Mistral 7B using AI-generated preference data and no human annotation, it scored 7.34 on MT-Bench against Llama 2 70B Chat's 6.86.
Open weights & ecosystem · Models & capabilities
TII releases Falcon 180B
At 180 billion parameters, trained on 3.5 trillion tokens, TII said it rivalled PaLM 2 — but its licence barred hosting the model as a paid service without permission.
Open weights & ecosystem · Models & capabilities
Hugging Face releases IDEFICS, an open Flamingo reproduction
Built entirely from public data and models, the 80B-parameter version reportedly matched the closed Flamingo it reproduced on several benchmarks.
Open weights & ecosystem · Models & capabilities
Hugging Face documents biases in using GPT-4 as a judge
Testing GPT-4 as a stand-in for human preference judges, Hugging Face found it favoured longer answers and its own family's outputs, correlating with humans only moderately.
Open weights & ecosystem · Benchmarks & progress
TII releases Falcon under Apache 2.0
Falcon-40B outperformed Meta's larger Llama 65B on the Open LLM Leaderboard despite using under half the training compute, largely on the strength of its filtered web dataset.
Open weights & ecosystem · Models & capabilities
Hugging Face releases StarCoder
The BigCode project released StarCoder, a 15B open code model trained on permissively-licensed repositories, with an OpenRAIL licence.
Open weights & ecosystem · Models & capabilities
BigScience releases BLOOM open multilingual model
Over 1,000 researchers from more than 70 countries trained the 176-billion-parameter model in the open on a French public supercomputer, releasing checkpoints and optimiser states alongside weights.
Open weights & ecosystem · Models & capabilities
Hugging Face raises $100m Series C
Hugging Face raised $100m in Series C funding led by Lux Capital, growing from 30 to 120 staff in a year while serving over 10,000 companies.
Money & business · Open weights & ecosystem
Also mentioned in 35 entries
Referenced in passing — Hugging Face isn't the main subject of these.
- July 2026Anthropic discloses Claude gained unauthorized access to real systems during security evaluations
- July 2026Thinking Machines Lab releases Inkling-Small, a distilled open-weight model
- July 2026OpenAI reports alignment failures in an internal long-horizon research model
- July 2026Moonshot AI launches Kimi K3
- June 2026Moonshot AI ships Kimi K2.7-Code
- June 2026Google releases Gemma 4 12B, an encoder-free multimodal open model
- June 2026MiniMax releases MiniMax-M3, combining frontier coding, 1M context and native multimodality
- April 2026Moonshot AI releases Kimi K2.6 open-weight flagship
- April 2026OpenAI expands Trusted Access for Cyber with a fine-tuned GPT-5.4-Cyber model
- March 2026Claude Opus 4.6 shown gaming a benchmark after detecting it was being evaluated
- February 2026Shanghai AI Laboratory open-sources Intern-S1-Pro, a 1-trillion-parameter scientific model
- December 2025MiniMax releases M2.1 update
- December 2025Mistral launches Mistral 3 model family
- December 2025DeepSeek releases DeepSeek-V3.2 and V3.2-Speciale
- October 2025Moonshot AI releases Kimi Linear architecture model
- September 2025DeepSeek releases DeepSeek-V3.1-Terminus
- September 2025Alibaba releases Qwen3-Next, Qwen3-VL and Qwen3-Omni
- August 2025xAI open-sources Grok 2 weights
- August 2025DeepSeek releases DeepSeek-V3.1 with hybrid reasoning mode
- June 2025Baidu open-sources the ERNIE 4.5 model family
- June 2025MiniMax releases MiniMax-M1, world's first open-weight large-scale hybrid-attention reasoning model
- March 2025Alibaba releases Qwen2.5-Omni multimodal model
- March 2025DeepSeek releases DeepSeek-V3-0324 update
- March 2025Mistral releases Mistral Small 3.1
- March 2025Manus markets a fully autonomous agent from China
- February 2025Perplexity open-sources decensored DeepSeek R1 variant
- February 2025Physical Intelligence open-sources π0
- November 2024Alibaba releases QwQ-32B-Preview reasoning model
- November 2024Tencent open-sources Hunyuan-Large MoE model
- September 2024Meta releases Llama 3.2 with vision and edge models
- July 2024Meta releases Llama 3.1 405B
- December 2023Mistral releases Mixtral 8x7B
- July 2023Meta releases Llama 2 for commercial use
- June 2023vLLM releases PagedAttention inference engine
- May 2023Direct Preference Optimization paper reframes RLHF as a classification loss