A public compendium · updated daily
How frontier intelligence is compressed, priced, and contested.
Distillation moves capability from a large teacher model into a small, cheap student. It is the quiet engine behind most models people actually pay for, the subject of an open dispute between the largest labs, and a live regulatory question in three jurisdictions. This compendium tracks all of it.
Methodology · 8 of 8 files published · 599 catalogued sources
How this compendium is built
Every figure on this site comes from a primary source that you can open in one click. Figures that could not be verified are written as undisclosed rather than estimated.
What this is
Global Distillation is a compendium of one technique: training a small student model on the behaviour of a large teacher. The same technique is a research method, a product line, a pricing strategy, a contract term and, since 2025, an accusation. Each of those readings gets its own section, built from the same catalogue of sources.
The site is static. Perspective research is written into JSON files that follow a published schema; the browser fetches them and renders every table and chart from the same data a reader can download. There is no server-side model, no summarisation step between the source and the figure, and no number that exists only in a chart.
Where the numbers come from
Sources are ranked in this order: the paper or technical report; the vendor’s own model card, pricing page or terms of service; a filing, statute or official memorandum; then reporting by a named outlet. A claim that exists only in reporting is attributed to whoever made it, in the text, not presented as fact.
Prices are list prices for the standard tier on the date shown, in US dollars per million tokens, excluding batch, cache and volume discounts, because those vary by contract. Benchmark scores are the figures the model’s own authors published, which is a real limitation: labs choose their comparisons. Where an independent reproduction exists, both are shown.
Download counts, star counts and paper counts are activity measures collected from public APIs. They measure attention, not quality, and they are labelled that way wherever they appear.
What this method cannot tell you
Three of the most interesting questions are unanswerable from public evidence, and this site does not pretend otherwise: whether a given closed small model was distilled, what a training run really cost, and whether a particular model was trained on a competitor’s outputs. Evidence offered in public disputes is circumstantial and is reported here as a claim by a named party, with its rebuttal.
Benchmark retention percentages compare a student against its own teacher on one benchmark. They do not transfer across benchmarks, and they say nothing about robustness, long-context behaviour, or how a model degrades outside the distribution it was distilled on.
How a model is classified as distilled
3 rows| Level | What it takes | Shown as |
|---|---|---|
| Stated | The vendor says so in a technical report, model card or launch post. | distilled |
| Documented method | A paper or repository describes the exact procedure and the teacher, even if the vendor avoids the word. | distilled (method) |
| Undisclosed | A small tier is widely assumed to be distilled but the vendor has never said and no paper describes it. | undisclosed |
The third level is deliberately not counted in any figure on this site.
Update cadence
6 rows| Data stream | Cadence | How it is collected |
|---|---|---|
| arXiv paper counts and new preprints | Daily, 06:17 UTC | arXiv API query for knowledge distillation, counted per year and for the last 30 days. |
| Hugging Face downloads | Daily, 06:17 UTC | Hub API, 30-day download totals for a fixed watchlist. |
| Repository stars | Daily, 06:17 UTC | GitHub API for the training, serving and evaluation repositories used in distillation work. |
| News and Hacker News items | Daily, 06:17 UTC | Search of Hacker News and named outlets; items are listed, never summarised into a claim. |
| Perspective research files | On revision | Rewritten end to end when a perspective is revisited; each file carries its own compiled date. |
| Prices and terms of service | On revision, checked against the vendor page | List prices for the standard tier on the date given in the row; batch and cache discounts excluded. |
A figure is never carried forward silently. If a stream fails, the panel says so rather than showing yesterday’s number.
Sources: info.arxiv.org · huggingface.co · docs.github.com · hn.algolia.com
What each data file contains
8 rows| File | Contains | Status | Sources sources | Compiled |
|---|---|---|---|---|
| data/academic.json | Papers, benchmark retention, reproductions | published | 48 | 4 September 2026 |
| data/financial.json | Token prices, training-run costs, market events | published | 63 | 3 September 2026 |
| data/political.json | Statutes, memoranda, export controls, disputes | published | 100 | 4 September 2026 |
| data/company.json | Corporate stance, products, terms of service | published | 75 | 3 September 2026 |
| data/developer.json | Libraries, platforms, GPU budgets, recipes | published | 58 | 3 September 2026 |
| data/customer.json | Buying guidance, model specifications | published | 78 | 4 September 2026 |
| data/library.json | Distillation methods, formulas, trade-offs | published | 66 | 4 September 2026 |
| data/timeline.json | Every dated event, 2006 to today | published | 111 | 3 September 2026 |
Counts are of catalogued primary sources, not of citations: one source is often cited by several figures.
Sources: global-distillation.com