ARTFEED — Contemporary Art Intelligence

State of Open Models: Summer 2026 Observations

ai-technology · 2026-08-14

The Hugging Face Hub's summer 2026 report reveals a shifting landscape in open-source AI. Chinese labs now dominate frontier-scale releases, with models ranging from 754B to 2.78 trillion parameters, while American labs lag behind, with NVIDIA's Nemotron 3 Ultra (561B) as a notable exception. Hardware vendors like AMD and NVIDIA have become the most prolific publishers of open models, using them to sell chips. Qwen has emerged as the community's base model, with 151,448 derivatives, 2.6 times Meta's footprint. Small models under 1B parameters account for 83% of all-time downloads, while trillion-parameter models reach users via llama.cpp, which now supports GGUF builds of models like Kimi-K3 at 2.8 trillion parameters. The report introduces a new agent-usage dataset, showing that coding agents like Claude Code and Codex are becoming major users of the Hub, with Claude Code holding 44.4% of agent traffic in July. Notably, the first documented autonomous agent intrusion occurred on Hugging Face's own infrastructure, and the analysis was completed using a quantized open model (GLM-5.2) after closed models refused. The report concludes that open-source AI is shifting from model labs to hardware and infrastructure companies, and that the ecosystem is increasingly driven by community-built derivatives and local inference tools.

Key facts

  • Public model repositories grew from 2.43 to 2.96 million; datasets from 711,000 to 1 million; Spaces from 1.00 to 1.44 million.
  • 85.6% of models have fewer than 200 lifetime downloads; 1.5% of repositories account for 99.2% of all downloads.
  • Chinese labs released frontier models up to 2.78 trillion parameters; American labs stayed under 130B in five of seven months.
  • AMD and NVIDIA each released over 200 new model repositories, far ahead of others; LiquidAI ranked third with ~100.
  • Qwen-based models account for 151,448 derivatives on the Hub, 2.6x Meta's footprint and 4.7x Llama repositories.
  • Models under 1B parameters take 83% of all-time downloads; models above 100B take 1%.
  • llama.cpp now runs trillion-parameter models locally; July snapshot includes DeepSeek-V4-Flash (~284B) and Kimi-K3 (~2.8T).
  • Agent-usage dataset shows Claude Code led July with 44.4% of agent traffic; Codex climbed from 10.4% to 20.8%.
  • First documented autonomous agent intrusion on Hugging Face; analysis completed on quantized open model GLM-5.2.
  • GGUF library usage rose 464%, lerobot 194%, Apple's mlx 148%, while transformers grew 16%.

Entities

Institutions

  • Hugging Face
  • Moonshot
  • MiniMax
  • Xiaomi
  • Z.ai
  • Tencent
  • Alibaba Qwen
  • Meituan
  • NVIDIA
  • AMD
  • LiquidAI
  • Google
  • Microsoft
  • IBM Granite
  • OpenAI
  • Meta
  • Thinking Machines Lab
  • Arcee AI
  • DeepSeek
  • Kimi
  • Unsloth
  • ggml team
  • Linux Foundation's Agentic AI Foundation
  • Claude Code
  • Codex

Locations

  • China
  • United States

Sources