Hugging Face's biannual landscape report, published August 14, shows how the open-weights ecosystem evolved from January through August 2026. The distribution is extreme: 85.6% of models have fewer than 200 lifetime downloads, while 1.5% of repositories account for 99.2% of all downloads.

The Hub grew to 2.96 million model repositories (up from 2.43 million at the start of the year), 1 million datasets, and 1.44 million Spaces.

ResourceCount (Aug 2026)Start-of-Year CountNet Growth
Model repositories2.96 million2.43 million+530 thousand
Datasets1 million
Spaces1.44 million
FIG. 02 Hugging Face Hub scale as of August 2026 (Jan–Aug growth period) — Hugging Face State of Open Models, Summer 2026

China's research labs — Moonshot, MiniMax, Xiaomi, Z.ai — skipped the small-to-large progression that defined earlier open-model strategies and deployed frontier-scale models directly. China's monthly parameter ceiling ran between 754B and 2.78 trillion in 2026; the U.S. ceiling stayed below 130B in five of seven months. NVIDIA's Nemotron 3 Ultra (561B) and Thinking Machines Lab's Inkling (952B) were the primary U.S. exceptions above 100B. Most other U.S. releases above 100B are ports or conversions built on Chinese base models.

GeographyMonthly Parameter Ceiling (2026)Months Below 130BNotable Exceptions
China754B – 2.78 trillion0 of 7Moonshot, MiniMax, Xiaomi, Z.ai — frontier-scale from launch
U.S.Mostly below 130B5 of 7NVIDIA Nemotron 3 Ultra (561B); Thinking Machines Lab Inkling (952B)
FIG. 03 China vs. U.S. monthly parameter ceiling comparison, 2026 — Hugging Face State of Open Models, Summer 2026

Two forces enabled the shift. First, the community quantization layer: a trillion-parameter release becomes runnable within days, removing the practical barrier that once forced labs to ship smaller models first. Second, large scale stopped being a cost differentiator. Xiaomi and Meituan both cleared a trillion parameters in 2026 without significant open-weights presence a year prior. Portfolio strategy signals intent: a frontier-only portfolio bets on benchmark position and API demand; a full-spectrum portfolio — Tencent and Alibaba Qwen cover everything from under 1B to above 70B — bets on becoming the standardized family.

Within U.S. institutions, the top publishers by new repository count are hardware vendors. AMD and NVIDIA each released more than 200 new model repositories in 2026, far ahead of any model lab; LiquidAI ranked third at roughly 100. Google and Meta now rank below NVIDIA in new model releases. Meta has shifted toward closed flagship models. Open source has migrated from model labs to hardware and infrastructure companies, with freely available models serving as proof that the underlying silicon works.

PublisherNew Model Repositories (2026)CategoryNotable Shift
AMD> 200Hardware vendorOpen models as silicon proof-of-work
NVIDIA> 200Hardware vendorRanks above Google and Meta in new releases
LiquidAI~ 100AI labThird-highest U.S. publisher
GoogleBelow NVIDIATech companyDown relative to hardware vendors
MetaBelow NVIDIATech companyShifted toward closed flagship models
FIG. 04 Top U.S. institutions by new model repositories released, Jan–Aug 2026 — Hugging Face State of Open Models, Summer 2026

The attention-versus-adoption split defines what matters. The top 25 repositories by 2026 downloads and the top 25 by likes share exactly one entry. Not a single model published in 2026 appears in the downloads top 25; thirteen of those 25 date from 2022. all-MiniLM-L6-v2 was pulled 1.55 billion times in seven months against 5,156 likes. Kimi-K3 was pulled about 60 times per like received. Likes record what the field finds significant in the weeks after release; downloads record what is actually running in production.

ModelDownloads (7 months)LikesDownloads per LikeSignal Type
all-MiniLM-L6-v21.55 billion5,156~300,000Production adoption — dates to 2022
Kimi-K3~60Community attention — new 2026 release
FIG. 05 Attention (likes) vs. adoption (downloads): selected examples from the 2026 download top-25 — Hugging Face State of Open Models, Summer 2026

For architects evaluating migration paths: the quantization layer makes most of the frontier technically accessible, but accessibility is not production readiness. The 85.6% of repos with sub-200 downloads is noise. Treat likes as an early-warning signal for what to evaluate, downloads as the signal for what has been validated at scale.