Methodology · 8 of 8 files published · 599 catalogued sources

How this compendium is built

Every figure on this site comes from a primary source that you can open in one click. Figures that could not be verified are written as undisclosed rather than estimated.

What this is

Global Distillation is a compendium of one technique: training a small student model on the behaviour of a large teacher. The same technique is a research method, a product line, a pricing strategy, a contract term and, since 2025, an accusation. Each of those readings gets its own section, built from the same catalogue of sources.

The site is static. Perspective research is written into JSON files that follow a published schema; the browser fetches them and renders every table and chart from the same data a reader can download. There is no server-side model, no summarisation step between the source and the figure, and no number that exists only in a chart.

Where the numbers come from

Sources are ranked in this order: the paper or technical report; the vendor’s own model card, pricing page or terms of service; a filing, statute or official memorandum; then reporting by a named outlet. A claim that exists only in reporting is attributed to whoever made it, in the text, not presented as fact.

Prices are list prices for the standard tier on the date shown, in US dollars per million tokens, excluding batch, cache and volume discounts, because those vary by contract. Benchmark scores are the figures the model’s own authors published, which is a real limitation: labs choose their comparisons. Where an independent reproduction exists, both are shown.

Download counts, star counts and paper counts are activity measures collected from public APIs. They measure attention, not quality, and they are labelled that way wherever they appear.

What this method cannot tell you

Three of the most interesting questions are unanswerable from public evidence, and this site does not pretend otherwise: whether a given closed small model was distilled, what a training run really cost, and whether a particular model was trained on a competitor’s outputs. Evidence offered in public disputes is circumstantial and is reported here as a claim by a named party, with its rebuttal.

Benchmark retention percentages compare a student against its own teacher on one benchmark. They do not transfer across benchmarks, and they say nothing about robustness, long-context behaviour, or how a model degrades outside the distribution it was distilled on.

How a model is classified as distilled

3 rows
How a model is classified as distilled — Three levels of evidence. A model is never moved up a level by inference.
LevelWhat it takesShown as
StatedThe vendor says so in a technical report, model card or launch post.distilled
Documented methodA paper or repository describes the exact procedure and the teacher, even if the vendor avoids the word.distilled (method)
UndisclosedA small tier is widely assumed to be distilled but the vendor has never said and no paper describes it.undisclosed

The third level is deliberately not counted in any figure on this site.

Update cadence

6 rows
Update cadence — What changes daily, and what changes when the research is redone.
Data streamCadenceHow it is collected
arXiv paper counts and new preprintsDaily, 06:17 UTCarXiv API query for knowledge distillation, counted per year and for the last 30 days.
Hugging Face downloadsDaily, 06:17 UTCHub API, 30-day download totals for a fixed watchlist.
Repository starsDaily, 06:17 UTCGitHub API for the training, serving and evaluation repositories used in distillation work.
News and Hacker News itemsDaily, 06:17 UTCSearch of Hacker News and named outlets; items are listed, never summarised into a claim.
Perspective research filesOn revisionRewritten end to end when a perspective is revisited; each file carries its own compiled date.
Prices and terms of serviceOn revision, checked against the vendor pageList prices for the standard tier on the date given in the row; batch and cache discounts excluded.

A figure is never carried forward silently. If a stream fails, the panel says so rather than showing yesterday’s number.

Sources: info.arxiv.org · huggingface.co · docs.github.com · hn.algolia.com

What each data file contains

8 rows
What each data file contains — One JSON file per perspective, validated against data/SCHEMA.md before it is served. — Units: Sources in sources.
FileContainsStatusSources sourcesCompiled
data/academic.jsonPapers, benchmark retention, reproductionspublished484 September 2026
data/financial.jsonToken prices, training-run costs, market eventspublished633 September 2026
data/political.jsonStatutes, memoranda, export controls, disputespublished1004 September 2026
data/company.jsonCorporate stance, products, terms of servicepublished753 September 2026
data/developer.jsonLibraries, platforms, GPU budgets, recipespublished583 September 2026
data/customer.jsonBuying guidance, model specificationspublished784 September 2026
data/library.jsonDistillation methods, formulas, trade-offspublished664 September 2026
data/timeline.jsonEvery dated event, 2006 to todaypublished1113 September 2026

Counts are of catalogued primary sources, not of citations: one source is often cited by several figures.

Sources: global-distillation.com

Sources for this page · 12 endpoints