Methodology
Principles
- Nothing is invented. When a value is not publicly available the site says Not publicly disclosed or Unknown.
- Every meaningful value is stored as a fact with value, source, source type, URL, published / retrieved / last-verified timestamps, confidence and verification status.
- Estimates, community data, calculations and editorial seeds are labelled as such and never presented as verified.
- History is appended, never overwritten: metrics are time series, prices are dated observations, events are an append-only log.
- There is no single "best AI" score. Separate measurements are shown with their inputs.
Data integrity & bias controls
We treat reliability as something to check, not assert. Every ingest applies these controls, and anything that fails goes to the public review queue rather than onto the site as fact:
- Prices are checked against the official page. A plan is marked “verified” only when its exact price appears next to the plan's name on the vendor's own pricing page. Pages that redirect to another site, block automated access in robots.txt, or render prices with JavaScript are labelled “not checkable” - never assumed correct. Mismatches are reviewed against the page text and corrected from the official figure.
- No single-language bias. Public-interest signals sum Wikipedia views across English and 9 other major languages, so products popular outside the English-speaking world are not systematically under-counted.
- No keyword inflation. Hacker News discussion is measured with specific phrases for ambiguous names (“Cursor IDE”, not “cursor”), each story counted once.
- No double counting. When a product's Wikipedia title redirects to another tracked entity's article, the views are not attributed twice; suspicious step changes (e.g. article moves) are flagged on the product page.
- No single-signal winners. The coverage penalty stops an entity with one strong signal from outranking broadly measured ones.
- No paid placement. Rankings and inclusion cannot be bought. Companies can send corrections, which are published only when a primary source confirms them.
- Conflicts stay visible. When sources disagree, both claims are kept and labelled.
Known residual biases we disclose rather than hide: Hacker News skews toward developer tools; Wikipedia attention favours consumer brands; benchmark coverage favours models that Epoch AI has evaluated; and seed data for products without machine-readable pricing pages remains unverified until checked.
Where data comes from
Source adapters run in the ingestion pipeline (npm run ingest). Each adapter is isolated: failures are retried with exponential backoff, rate-limited per host, protected by a circuit breaker and reported without stopping the run. Official sources rank above independent press, which ranks above community data. Full list with health on the sources page.
| Source | Tier | Used for | Cadence |
|---|---|---|---|
| OpenRouter models API | 2 | Structured listing of 400+ hosted models: prices per token, context length, modalities, supported parameters and listing date. Prices are OpenRouter's listed prices, which usually mirror first-party pricing but can differ. | every run (recommended 6h) |
| Epoch AI Benchmarking Hub | 1 | Independent benchmark runs (GPQA Diamond, SWE-bench Verified, FrontierMath, ...) plus externally reported leaderboards and the Epoch Capabilities Index (ECI). | daily |
| Artificial Analysis (via OpenRouter) | 2 | Artificial Analysis Intelligence Index values as embedded in the OpenRouter listing. Shown separately and never merged with other benchmarks. | every run |
| Wikimedia Pageviews API | 3 | Daily human (agent=user) page views of the entity's Wikipedia article in English plus 9 other major languages (de, fr, es, ja, zh, ko, pt, ru, it), English redirects included. A proxy for public interest - not users. | daily |
| Hacker News (Algolia API) | 3 | Count of HN stories whose title matches the entity's search phrase. Community signal used for discussion/trend detection only; keyword matching may include false positives. | daily |
| GitHub REST API | 1 | Stars, forks, open issues, last push and license for official repositories. | daily |
| Hugging Face Hub API | 1 | Downloads (last 30 days) and likes for official model repositories. | daily |
| npm downloads API | 1 | Daily downloads of official SDK packages. Includes CI/automated installs. | daily |
| PyPI Stats | 1 | Daily downloads (without mirrors) of official Python packages, last ~180 days. | daily |
| Wikidata | 3 | Structured company facts (inception, headquarters, founders, CEO, employees, revenue, parent). Community-maintained; shown with the statement's point-in-time where available. | weekly |
| Official company feeds | 1 | Official blogs/newsrooms: OpenAI, Google, Google DeepMind, Hugging Face, Mistral, NVIDIA, GitHub, AWS ML. | every run (recommended 1-3h) |
| Technology press feeds | 2 | TechCrunch AI, The Verge AI, Ars Technica, MIT Technology Review AI. | every run |
| Internet Archive Wayback Machine | 3 | Historical snapshots of the OpenRouter listing used to backfill price history. | weekly |
| Official pricing pages | 1 | Each product's official pricing page, fetched (robots.txt respected) to confirm every seed price appears verbatim. Plans are marked verified only on an exact match; pages rendered by JavaScript are reported as not checkable. | every run (pages cached 3h) |
| Editorial seed | 3 | Hand-curated structural facts (which company owns a product, category, official URLs) and seed pricing. Seed prices are marked 'unverified' until re-checked against the linked official pricing page. | manual |
| AI Stats Live methodology | 1 | Rankings, momentum, discovery scores and normalised prices computed from the sources above. | every run |
| Similarweb (traffic) | 2 | Website traffic estimates. Requires a paid API key; disabled unless SIMILARWEB_API_KEY is set. | not configured |
| App store metrics | 2 | App downloads/ratings. No free, terms-compliant source is configured; reserved adapter. | not configured |
Rankings
Rankings measure attention and adoption signals, not quality. For each signal we take the trailing window sum (daily sources) or the latest snapshot (GitHub, Hugging Face), log-scale it relative to the day's maximum (0–100), and take the weighted mean of the signals that are available. The result is multiplied by 0.4 + 0.6 × coverage, where coverage is the share of total weight with data - so an entity with only one strong signal cannot outrank broadly measured ones.
Because inputs are stored as daily series, every day's ranking is recomputed from history, making movement arrows (▲ 4, ▼ 2) reproducible. Snapshot-based rankings (developer, open source) only have history from the first day a snapshot was recorded - older values are never back-filled.
Public interest (Wikipedia views), developer discussion (Hacker News) and SDK downloads (npm + PyPI) over a trailing 28-day window.
- Wikipedia pageviews (28d)40%
- Hacker News stories (28d)30%
- SDK downloads, npm + PyPI (28d)30%
- Signals are proxies for attention and adoption, not user counts.
- No website-traffic or app-download data is included (no terms-compliant free source is configured).
- Products without a dedicated Wikipedia article or package lose that signal (see coverage).
- HN matching is by title phrase and may include unrelated stories for generic names.
Wikipedia article views only - a proxy for mainstream public interest.
- Wikipedia pageviews (28d)100%
- Only entities with their own English Wikipedia article are ranked.
- Signals are proxies for attention and adoption, not user counts.
- No website-traffic or app-download data is included (no terms-compliant free source is configured).
SDK/package downloads, developer discussion and GitHub stars.
- SDK downloads, npm + PyPI (28d)45%
- Hacker News stories (28d)25%
- GitHub stars30%
- Package downloads include CI and automated installs.
- Rank history starts on the first day GitHub star snapshots were recorded.
GitHub stars and forks, plus Hugging Face downloads for open-weight model families.
- GitHub stars45%
- GitHub forks20%
- Hugging Face downloads (30d)35%
- Snapshot signals: rank history accumulates from the first ingest run.
- Hugging Face downloads are summed over the organisation's models (optionally filtered by family name).
Company article views plus the combined attention of the company's tracked AI products.
- Wikipedia pageviews (28d)30%
- Product Wikipedia views (28d)35%
- Hacker News stories (28d)35%
- Big-tech company articles attract attention unrelated to AI.
- Footprint depends on how many of the company's products are tracked.
Enterprise presence is not ranked: no measurable, public, terms-compliant source is configured. Website traffic and app downloads are not used for the same reason; the Similarweb and app-store adapters are present but disabled.
Trend & momentum
Trend (7d) compares the last 7 days with the average week of the previous 28 days. Momentum (30d) compares the last 30 days with the 30 before. Both are weighted across Wikipedia views (40%), HN stories (30%) and SDK downloads (30%), each capped at −90%/+500% and requiring a minimum baseline (700 weekly views, 2 weekly stories, 5,000 weekly downloads) so tiny numbers do not produce huge percentages. "Why is this trending?" lists the per-signal changes as likely contributing signals; no cause is inferred.
Discovery score
A calculated metric for emerging entities, separate from popularity. Products launched within 18 months: 35% recency, 25% trend, 20% momentum, 20% event activity (60 days). Models released within 90 days: 50% recency, 30% capability (ECI), 20% event activity.
Verification & confidence
- Verified
Read directly from a primary or structured source at the retrieval time shown.
- Reported
Stated by a named source (company, publication, benchmark body) but not independently confirmed.
- Estimated
A derived estimate - always labelled; never shown as fact.
- Community-sourced
From a community-maintained reference (Wikidata) - generally reliable, open to error.
- Calculated
Computed by our documented methodology from the underlying sources.
- Unverified seed
Editorial seed awaiting automated verification (e.g. consumer prices).
- Unknown
No value could be established.
Confidence (high / medium / low) reflects the source tier and extraction method. Low-confidence items (e.g. unclassified official posts, unmatched benchmark organisations) are routed to the review queue rather than published silently.
Freshness
- Verified <24h - Read from its source in the last 24 hours.
- Verified 1–7d - Verified between 1 and 7 days ago.
- Verified 7–30d - Verified between 7 and 30 days ago.
- Stale >30d - Not verified for more than 30 days - treat with caution.
- Not verified - No verification date recorded (editorial seed or unknown).
If a source fails, the last known value stays visible with its original verification date - it ages into "stale" instead of being silently presented as current.
Conflicting sources
When sources disagree (e.g. an editorial license vs the GitHub repository license, or country per seed vs Wikidata), both claims are kept, the page states "Sources disagree", and a review item is created. Where fields conflict, the source priority is Official > Filing > Benchmark body > Trusted news > Reference > Community > Editorial - applied only where that source type actually supports the field.
Price normalisation
API prices are normalised to USD per 1M tokens. "Blended" prices assume a 3:1 input:output token ratio (stated wherever used). Subscription plans are converted to a monthly equivalent (annual ÷ 12); one-time, usage-based and custom prices are never converted. Image, video and audio generation prices are not normalised until a structured source exists. The cost calculator is informational and excludes reasoning tokens, tool fees and discounts.
Event engine
Events come from official feeds, press feeds, model listings (new models, retirement dates), Epoch benchmark runs and change detection between ingest snapshots (prices, context windows). Articles are classified by rule into normalised types (MODEL_LAUNCH, PRICE_INCREASE, ACQUISITION, …), linked to entities by name and alias (short or ambiguous names require exact case), and clustered: articles within 72 hours with similar titles or the same primary entity and type become one event. The strongest source - official first, then earliest - is shown as primary. Article text is not copied.
Limitations
- Wikipedia views and HN stories measure attention, not usage; products without their own article lose that signal. Views of redirect titles are summed with the article so page moves do not create false jumps.
- HN phrase matching can include unrelated stories for generic names (e.g. "Cursor", "Gemini").
- Package downloads include CI and mirrors (PyPI figures exclude known mirrors).
- Model release dates fall back to the OpenRouter listing date when Epoch has no record; this is labelled.
- Consumer prices and capability lists are editorial seeds until automated verification of official pricing pages is added.
- Regional availability and privacy/compliance details are only shown where captured from vendor pages, and are labelled as vendor claims.