Launch tempo
Used on: Home (the record and the tape); The idea; Ledger · Fig. 2
- Source
- Hugging Face public models API, pulled 2026-09-01 (22:08 UTC). ↗
- Method
- Lab list of 42 organisation ids in labs.json; 6,977 repositories read. Pipeline tags kept: text-generation, image-text-to-text, any-to-any. Quantizations and derivative repositories excluded (GGUF, AWQ, GPTQ, MLX, FP8, INT4 and similar). Families deduplicated by stripping size and variant suffixes; each family dated by its earliest repository creation and scored by its peak likes. Tiers: flagship 1,000+, major 500 to 999, broad 250 to 499. Window 2024-09-01 to the pull date.
- Assumptions
- Likes are a popularity proxy, not a capability measure. The lab list is a judgement call. ModelScope is not scanned, so labs releasing only there are invisible. Family deduplication is heuristic.
- Uncertainty
- Likes accumulate over time, so recent months are undercounted at any fixed threshold. Counts are floors, not estimates. Rows count releases, not evaluation campaigns; tier and modality variants of one model stay separate.
- Falsification
- Name a qualifying release the pull missed, or a counted family that is a derivative. The pull re-runs without credentials.
Coverage
Used on: Home (the record); The idea · Fig. 1; Ledger · Fig. 1
- Source
- Model cards, technical reports, release posts and independent evaluation reports, walked in the rubric’s source hierarchy.
- Method
- Rows are open-weight releases judged frontier-capable: listed as frontier or notable by Epoch AI, ranked near the frontier on a public index, or a flagship release from a frontier lab (the rule is recorded per row in the data file). API-only comparison rows are excluded. Every yes or partial cell carries a URL. The four labels on the site map to the rubric's public tiers: nothing found = T0, model card only = T1, developer evaluation = T2, independent evaluation = T3; the rubric also keeps an internal 0 to 5 scale. 19 of 39 rows coded as of 2026-09-01; the rest render as unreviewed, never as T0.
- Assumptions
- Independence ranks above a lab’s own disclosure: a lab-run CBRN evaluation is T2, an independent methodology-public evaluation is T3. "No source found" is an observation, not proof of absence.
- Uncertainty
- Two passes: one coder with a stated source walk per row, then an independent second reviewer who opened every T3 source URL and spot-checked T1 rows. Coding notes record every ambiguous call.
- Falsification
- Write to info@pacificcompute.org with the evaluation we missed and its source.
Cite this dataset
Pacific Compute Ledger, version 17df23f, 2026-09-01. pacificcompute.org/ledger.