Social intelligence · YouTube · August 29, 2026
4,577,544 public YouTube channel profiles — subscribers, videos, views, region, bio and links — queryable over one REST API. Filter by subscriber band, region, channel-creation date or field coverage; sort by reach. Read from each channel's public About page, no login and no YouTube Data API quota, pay on success.
4,577,544
YouTube channels — one record per channel, refreshed continuously.
8.2%
1K+ subscribers
27.8%
have a bio
20.7%
have a linked URL
Snapshot August 29, 2026 — public fields only. A seeded index discovered via Common Crawl and Wikidata, not a sample of YouTube.
4,577,544
public YouTube channels, one record per channel, read from the public About page — no login and no API quota. Common Crawl accounts for 100% of attributed discoveries; 56,099 records predate discovery-source tracking entirely.
8.2%
of channels have 1,000 or more subscribers — the opposite shape from this catalog's X Users and Instagram Users datasets. Homepage/backlink discovery casts a far wider net than a notable-person seed list, so this index is long-tail, not top-heavy.
91.9%
of indexed channels show a creation date between 2006 and 2011, even though YouTube's real channel growth has exploded since. That is a seed-method artifact, not a fact about YouTube — Common Crawl's web-archive snapshots disproportionately capture channels old enough to have accumulated backlinks and mentions.
27.8%
of channels publish a bio, and just under a quarter carry an outbound link — lower signal density than this catalog's other social datasets, since the discovery net here reaches deep into small, sparsely-filled channels.
Every number on this page describes the channels in the index, not YouTube's creator base. Channels are discovered two ways: Common Crawl's web-archive sweep (by far the larger source) and Wikidata (notable people and organizations with a linked channel). Common Crawl's method — following links and mentions across archived web pages — pulls in everything from mega-creators to single-video hobby channels, which is why this index looks nothing like a curated "top creators" list. It is the right shape for finding and enriching channels at any scale; it is the wrong shape for claims like "the average YouTube channel".
Unlike this catalog's notable-person-seeded social datasets, this one is not top-heavy: the modal band is under 1,000 subscribers, and only 8.2% clear the 1,000-subscriber line. 10,153 channels sit above a million — up to 511M on MrBeast. Some of that floor is a hidden count, not a genuinely tiny channel: 937,614 channels have their subscriber count hidden on their public About page and are stored as zero — filter on the availability flags to tell the two apart.
| Subscribers | Channels | Share |
|---|---|---|
| Under 1K | 4,202,089 | 91.8% |
| 1K–9.9K | 251,779 | 5.5% |
| 10K–99K | 83,772 | 1.8% |
| 100K–999K | 29,751 | 0.6% |
| 1M+ | 10,153 | 0.2% |
Creation dates cluster overwhelmingly in YouTube's first six years: 91.9% of indexed channels joined between 2006 and 2011, peaking at 826,790 channels in 2007 alone. That is not a fact about when people actually started YouTube channels — it is the crawl-discovery method showing through. Old channels have had more time to accumulate the backlinks and web mentions that Common Crawl's archive captures; a channel created last year simply has not had that time yet.
A profile is only useful if the fields you need are populated. Subscriber counts are available for most channels, but bios and links are sparser here than on this catalog's other social datasets — a direct consequence of the discovery net reaching deep into small, minimally-filled-out channels alongside the mega-creators.
Channel region is a self-declared About page field, and the facet below returns only the top buckets, not every value in the index — 10.6% of channels are covered by these 12 regions. Treat it as a top-region view, not a full geographic breakdown.
| Region | Channels | Share of tagged |
|---|---|---|
| United States | 170,208 | 35% |
| United Kingdom | 31,878 | 6.6% |
| Germany | 26,770 | 5.5% |
| Brazil | 23,532 | 4.8% |
| Canada | 21,770 | 4.5% |
| Japan | 18,052 | 3.7% |
| Spain | 16,029 | 3.3% |
| France | 15,472 | 3.2% |
| Italy | 14,396 | 3% |
| India | 11,977 | 2.5% |
| Australia | 10,724 | 2.2% |
| Russia | 9,267 | 1.9% |
Shares here are of the 485,904 channels these top regions cover, not the full 4,577,544-record index — the region facet is a top-N aggregation by construction, not a data-quality gap.
Straight from a followers_desc query — nothing hand-picked.
| Channel | Subscribers | Joined |
|---|---|---|
| MrBeast | 511,000,000 | 2012 |
| T-Series | 314,000,000 | 2006 |
| Cocomelon - Nursery Rhymes | 202,000,000 | 2006 |
| SET India | 190,000,000 | 2006 |
| Stokes Twins | 143,000,000 | 2008 |
| ✿ Kids Diana Show | 138,000,000 | 2015 |
| 김프로KIMPRO | 134,000,000 | 2017 |
| Like Nastya | 133,000,000 | 2016 |
| Zee Music Company | 122,000,000 | 2014 |
| Alejo Igoa | 120,000,000 | 2014 |
| WWE | 114,000,000 | 2007 |
| Goldmines | 110,000,000 | 2012 |
One record per channel. Grouped for readability; the API returns a flat object. No video metadata or captions — for those, use the live YouTube endpoints.
Identity
channel_idchannel_namechannel_urlprofile_picProfile
biolinksregionjoined_dateAudience
followers_countvideos_countviews_countfollowers_count_availablevideos_count_availableviews_count_availableStatus & provenance
discovery_sourcestatusdiscovered_athydrated_atThe search endpoint takes the filters below; combine any of them. Page with page and page_size (≤100 per page, and page × page_size ≤ 10,000).
Full-text & identity
Audience
Field coverage
Dates
Sort
relevancefollowers_descfollowers_ascviews_descvideos_deschydrated_at_deschydrated_at_ascEvery query authenticates with an x-api-key header and reads the stored search index — there is no live crawl, proxy or YouTube Data API quota to manage, and you are billed pay on success: charged for results, not failed requests. Three endpoints cover it:
GET /datasets/youtube-creators/search — filter, sort, page the population.GET /datasets/youtube-creators/items/{channel_id} — one channel.GET /datasets/youtube-creators/facets — counts for region or discovery source.The same three are exposed as MCP tools — datasets_youtube_creators_search, datasets_youtube_creators_item and datasets_youtube_creators_facets — so an agent can call them directly. Need video metadata, captions or a live read of a channel that is not in the index? Those are the live YouTube endpoints.
Cite this
Crawlora (2026). YouTube Creators Dataset. 4,577,544 public YouTube channel profiles, seeded index; public fields only. https://crawlora.net/datasets/youtube-creators.
Search with filters
# Channels with 1M+ subscribers, biggest first
curl "https://api.crawlora.net/api/v1/datasets/youtube-creators/search?min_followers=1000000&sort=followers_desc" \
-H "x-api-key: $CRAWLORA_API_KEY"One channel by channel_id
# One channel by channel_id
curl "https://api.crawlora.net/api/v1/datasets/youtube-creators/items/UCX6OQ3DkcsbYNE6H8uQQuVA" -H "x-api-key: $CRAWLORA_API_KEY"Facet the population
# Facet the population by region (or discovery_source)
curl "https://api.crawlora.net/api/v1/datasets/youtube-creators/facets?facet=region" -H "x-api-key: $CRAWLORA_API_KEY"Channels with 1M+ subscribers, biggest first
# Channels created since 2020 with a public bio
curl "https://api.crawlora.net/api/v1/datasets/youtube-creators/search?joined_after=2020-01-01&has_bio=true&sort=followers_desc" \
-H "x-api-key: $CRAWLORA_API_KEY"The same figures behind the bars and the line — plain and machine-readable for search engines and AI answer engines that cannot parse a chart.
| Subscribers | Channels | Share |
|---|---|---|
| Under 1K | 4,202,089 | 91.8% |
| 1K–9.9K | 251,779 | 5.5% |
| 10K–99K | 83,772 | 1.8% |
| 100K–999K | 29,751 | 0.6% |
| 1M+ | 10,153 | 0.2% |
| Joined | Channels | Share of index |
|---|---|---|
| 2005 | 15,850 | 0.3% |
| 2006 | 709,045 | 15.5% |
| 2007 | 826,790 | 18.1% |
| 2008 | 767,179 | 16.8% |
| 2009 | 735,947 | 16.1% |
| 2010 | 583,047 | 12.7% |
| 2011 | 582,505 | 12.7% |
| 2012 | 86,505 | 1.9% |
| 2013 | 42,509 | 0.9% |
| 2014 | 38,901 | 0.8% |
| 2015 | 32,787 | 0.7% |
| 2016 | 27,621 | 0.6% |
| 2017 | 23,400 | 0.5% |
| 2018 | 18,514 | 0.4% |
| 2019 | 15,508 | 0.3% |
| 2020 | 25,584 | 0.6% |
| 2021 | 15,249 | 0.3% |
| 2022 | 10,978 | 0.2% |
| 2023 | 8,795 | 0.2% |
| 2024 | 5,836 | 0.1% |
| 2025 | 3,884 | 0.1% |
| 2026 (partial) | 1,109 | 0% |
Query 4.6M public YouTube channel profiles — subscribers, videos, views, region, bio and links — through one REST API. Creator discovery, audience research, lead enrichment or channel vetting: clean JSON, public fields only, pay on success.
The dataset holds 4,577,544 public YouTube channel profiles, deduplicated to one record per channel_id. 56,099 records predate discovery-source tracking and carry no field at all, so the discovery-source breakdown sums slightly short of the total rather than landing in an "unknown" bucket.
No, and it matters — though not in the direction you might expect. Channels are discovered two ways: Common Crawl's web-archive sweep (100% of attributed discoveries) and Wikidata (notable people and organizations with a linked channel). Because Common Crawl follows links and mentions across the open web rather than a curated notable-person list, this index reaches deep into small channels — only 8.2% have 1,000+ subscribers, and 91.9% show a creation date between 2006 and 2011 purely because older channels have had more time to accumulate the web mentions this method finds. Use it as a broad discovery surface, not as a census or a basis for population-level claims about YouTube creators.
Every field is a public, credential-free channel field read from the channel's public About page: channel_id, channel name, channel URL, profile picture, linked external URLs, subscriber/video/view counts (with an availability flag for each, since channels can hide them), region, bio, channel-creation date, the discovery source, and a hydrated_at timestamp. Nothing behind authentication is collected — no login, no API key, no YouTube Data API quota. There is no video metadata, no captions and no comments; for those, use the live YouTube endpoints instead.
It is a discovery-method artifact, not a fact about YouTube. 91.9% of indexed channels (4,204,513) show a creation date in that six-year window, peaking at 826,790 channels in 2007 — even though YouTube's actual channel creation has grown enormously since. Common Crawl's web-archive snapshots capture pages that have accumulated backlinks and mentions over time, so long-established channels are systematically over-represented relative to newer ones that simply have not had the same time to be discovered and linked.
region is a self-declared field from the channel's public About page. It is sparse by construction — the facet endpoint returns only the top buckets, and the 12 regions shown on this page cover 10.6% of the 4,577,544-record index. Filter on region to narrow a search; do not treat its partial coverage as a data-quality gap.
Every record in this snapshot was hydrated between July 28, 2026 and August 29, 2026. Each record carries its own hydrated_at timestamp, and you can filter or sort on it (hydrated_after, hydrated_before, sort=hydrated_at_asc) to find the stalest records yourself.
Yes. Three endpoints cover it: search (full-text and faceted, with subscriber, video, view, date and field-coverage filters), one channel by channel_id, and facet counts for region or discovery source. The same three are exposed as MCP tools — datasets_youtube_creators_search, datasets_youtube_creators_item and datasets_youtube_creators_facets — so an agent can call them directly. Reads hit the stored index and never trigger a live crawl or consume YouTube Data API quota; billing is pay-on-success.