LYRENTH
DocsPricingBenchmarksIndex statsAboutBlogFor site ownersContact
Index stats

The AI-readable
web index, live.

Production truth, ticking: every counter below reads from the live index, where each page of the public web becomes one canonical, clean AIDocument. First the index; then what it delivers; then the audit of the raw web it replaces.

Lyrenthweb indexGET /v1/stats · every 0.5sLIVE
AIDocuments in the canonical index
1,974,004,159
one canonical document per URL, written once, read by everyone
Indexed domains
149,193,150
distinct domains represented
Audited pages
1,971,891,982
99.9% of the corpus, scored on seven signals
the corpus, to scale: one cell = 10,000,000 AIDocumentscell 198 is filling now: 40.0%
What the index delivers
Measured token reduction
94%
raw page100%
aidocument6%

median across the measured benchmark set, from the API's own economics

JS rendering absorbed
50%
served staticrendered by the index

share of pages that only exist after JavaScript runs; the index runs the browser so your agent never does

Corpus freshness
24h/ 90d

recrawl cadence by plan; re-index any page on read with force_refresh

Before Lyrenth

The raw web, audited

This is the part the index replaces. Every page is scored on seven content signals measured on the source page as our crawler found it: the state of the raw web, not of what we serve. Agents reading through the index never see these problems: they get the cleaned AIDocument, and the token difference is measured on the benchmarks page.

05105.6/10
AI readability, corpus mean
how the score is composed: segment width = weight in the score · fill = corpus mean
01Structured dataweight 20%

Whether the page carries JSON-LD blocks (Article, Product, FAQPage…). Structured data lets agents extract facts without parsing prose.

19%
02Heading hygieneweight 15%

Single H1, monotonic descent through H2 / H3. Predictable structure makes a page easier to skim and section.

62%
03Static renderabilityweight 20%

Share of pages served without a headless-Chromium escalation. Static-renderable pages cost less and never arrive empty to first-pass scrapers.

50%
04Content densityweight 20%

Ratio of meaningful markdown to raw HTML. High density means most of the page is content, not nav / chrome / ads.

64%
05Title & descriptionweight 10%

Non-empty, sensible-length title and description that are not generic placeholders.

65%
06Open accessweight 10%

Share of indexed pages that reach readers (and agents) without paywall or login-wall markers. Paywalled pages are indexed and scored too, so this is a measured rate across the corpus, not a definition.

99.9%
07Content depthweight 5%

Whether the page has enough words to be substantive on its own. Sub-stub pages get partial credit; pages with no body fail outright.

81%
How to read this

What these numbers mean

For AI agent builders

A high site-wide structured-data percentage means agents can extract facts cheaply. Low static-renderability or density means more of your input tokens pay for headless rendering and chrome, not content. Lyrenth normalizes these into a single AIDocument shape regardless of how the source page is built.

Read the quickstart

For site owners

The signals above are exactly what Lyrenth measures per page on your verified domains. Verify a domain to see your own AI Readiness score on the dashboard, broken down page by page.

Verify a domain

Methodology: every audited page is scored on seven content signals, each 0.0 to 1.0. Per-page scores roll up to a domain average and a corpus-wide mean. Raw page content stays in our private index; only aggregate counts and means are public. Counters are live; signal aggregates recompute about every 10 minutes.