GDELT API endpoint
Use Crawlora's GDELT Television Search API to extract supported public GDELT data as structured JSON. This page includes request parameters, cURL examples, response schema, error behavior, credit cost, and a Playground link for testing before integration.
/gdelt/tv-searchSearch GDELT's Television 2.0 AI index of US television news Developers commonly use this endpoint for data enrichment, monitoring, research dashboards, internal automation, and agent-native workflows that need repeatable structured public web data. Authentication uses the documented Crawlora headers, and usage is metered with the credit cost shown on this page.
Request parameters are generated from the active endpoint catalog. Required values must be sent before Crawlora can call the upstream public web data source.
| Parameter | Type | Required | Default | Description | Example |
|---|---|---|---|---|---|
| transcript | array | No | Search machine-generated speech-to-text transcripts (GDELT's asr: operator). Repeatable; multiple values are OR'd together. Short phrases only (GDELT caps each at 5 words). | ||
| caption | array | No | Search human-provided closed captioning (GDELT's cap: operator). Repeatable; multiple values are OR'd together. | ||
| concept | array | No | Search Google Knowledge Graph concepts extracted from captioning, by MID code (GDELT's capnlp: operator). Repeatable; multiple values are OR'd together. | ||
| onscreen_text | array | No | Search OCR'd onscreen text/chyrons (GDELT's ocr: operator). Repeatable; multiple values are OR'd together. Short phrases only (GDELT caps each at 5 words). | ||
| visual | array | No | Search visual object/activity labels from computer vision (GDELT's visual: operator). Repeatable; multiple values are OR'd together. | ||
| exclude_transcript | array | No | Exclude clips whose speech-to-text transcript matches this value (GDELT's -asr: operator). Repeatable; every value must be absent. | ||
| exclude_caption | array | No | Exclude clips whose closed captioning matches this value (GDELT's -cap: operator). Repeatable; every value must be absent. | ||
| exclude_concept | array | No | Exclude clips whose extracted concepts match this MID code (GDELT's -capnlp: operator). Repeatable; every value must be absent. | ||
| exclude_onscreen_text | array | No | Exclude clips whose OCR'd onscreen text matches this value (GDELT's -ocr: operator). Repeatable; every value must be absent. | ||
| exclude_visual | array | No | Exclude clips whose visual labels match this value (GDELT's -visual: operator). Repeatable; every value must be absent. | ||
| station | string | Yes | Station to search. Allowed values: CNN, MSNBC, FOXNEWS, BBCNEWS, KGO, KPIX, KNTV | ||
| show | string | No | Limit to an exact show name. | ||
| day_of_week | integer | No | Limit to a day of week, 0 (Sunday) through 7 (Saturday), PST. Minimum: 0. Maximum: 7. | ||
| timespan | string | No | Relative time window ending now, e.g. 1h, 7d, 3m, 1y. Cannot be combined with from/to. GDELT's TV archive starts July 6, 2010. | ||
| from | string | No | Start of an absolute time window. Accepts YYYY-MM-DD, RFC3339, or GDELT's raw YYYYMMDDHHMMSS. Cannot be combined with timespan. | ||
| to | string | No | End of an absolute time window. Same formats as from. | ||
| sort | string | No | Result order. Allowed values: relevance, datedesc, dateasc | ||
| maxrecords | integer | No | Rows to return. GDELT's own hard cap is 3000 for this endpoint; there is no pagination cursor beyond it. Minimum: 1. Maximum: 3000. | ||
| x-api-key (header) | string | Yes | API key required |
curl -X GET "https://api.crawlora.net/api/v1/gdelt/tv-search?station=CNN&sort=relevance" \ -H "x-api-key: $CRAWLORA_API_KEY"
Send your scraping API key in the x-api-key header. Use the console API Keys page to rotate or select the active key.
Endpoint usage is metered in credits. The plan prices, included credits, limits, and overage rates below match the active backend billing configuration.
| Plan | Price | Included credits | Daily cap | Rate limit | Overage |
|---|---|---|---|---|---|
| Free | $0/mo | 2,000 | 500 daily credits | 5/min | No overage |
| Starter | $9/mo | 20,000 | 5,000 daily credits | 15/min | $0.75/1,000 overage credits when enabled |
| Growth | $29/mo | 100,000 | 25,000 daily credits | 45/min | $0.45/1,000 overage credits when enabled |
| Pro | $79/mo | 400,000 | No daily cap | 120/min | $0.30/1,000 overage credits |
| Business | $199/mo | 1,200,000 | No daily cap | 300/min | $0.20/1,000 overage credits |
| Enterprise | $499/mo | 5,000,000 | No daily cap | 1,000/min | $0.12/1,000 overage credits |
This endpoint is executed through Crawlora's managed scraping infrastructure.
- `matched_at` is the exact timestamp of the matched moment within the broadcast; `show_start_at` is when the containing show began. - `visual_entities` is a normalized list (GDELT's own upstream field is a comma-separated string). - `count` is the number of clips actually returned (0 is a valid, legitimate "no matches" result, distinct from an upstream error). - `source_url` is the exact upstream GDELT Television 2.0 AI API URL used for the call. - A GDELT-side rate limit, an all-too-common ASR/caption/OCR phrase, a date outside the archive's range, or an unrecognized/malformed upstream response surfaces as a `503` upstream error, not a `200` with empty/wrong data. Example response: ```json { "code": 200, "msg": "OK", "data": { "query": "cap:\"climate change\" station:CNN", "station": "CNN", "sort": "relevance", "maxrecords": 50, "count": 1, "source_url": "https://api.gdeltproject.org/api/v2/tvai/tvai?format=json&maxrecords=50&mode=clipgallery&query=cap%3A%22climate+change%22+station%3ACNN", "fetched_at": "2026-08-20T10:00:00Z", "clips": [ { "station": "cnn", "show": "Inside Politics", "show_start_at": "2019-05-03T16:00:00Z", "matched_at": "2019-05-03T17:00:56Z", "transcript": "climate change the", "caption": "climate", "onscreen_text": "SWEEPING PLAN LIVE", "visual_entities": ["person", "news", "facial expression"], "clip_url": "https://archive.org/details/CNNW_20190503_160000_Inside_Politics/start/3656", "thumbnail_url": "http://data.gdeltproject.org/televisionexplorer/thumbnails/CNNW_20190503_160000_Inside_Politics-003656.jpg" } ] } } ```
Crawlora does not silently return bad data when the upstream page cannot be used.
| Status | Common failure case |
|---|---|
| 400 | Invalid input or missing required parameter |
| 429 | Plan or endpoint rate limit exceeded |
| 500 | Internal execution error |
| 502 | Upstream platform failed, returned unusable HTML, or served a challenge page that could not be resolved |
When possible, Crawlora returns structured error context so your integration can retry, back off, or inspect the request.
| Status | Description | Schema |
|---|---|---|
| 400 | Bad Request | #/definitions/app.Response |
| 500 | Internal Server Error | #/definitions/app.Response |
| 503 | Service Unavailable | #/definitions/app.Response |
{
"code": 200,
"msg": "OK",
"data": {
"query": "cap:\"climate change\" station:CNN",
"station": "CNN",
"sort": "relevance",
"maxrecords": 50,
"count": 1,
"source_url": "https://api.gdeltproject.org/api/v2/tvai/tvai?format=json&maxrecords=50&mode=clipgallery&query=cap%3A%22climate+change%22+station%3ACNN",
"fetched_at": "2026-08-20T10:00:00Z",
"clips": [
{
"station": "cnn",
"show": "Inside Politics",
"show_start_at": "2019-05-03T16:00:00Z",
"matched_at": "2019-05-03T17:00:56Z",
"transcript": "climate change the",
"caption": "climate",
"onscreen_text": "SWEEPING PLAN LIVE",
"visual_entities": [
"person",
"news",
"facial expression"
],
"clip_url": "https://archive.org/details/CNNW_20190503_160000_Inside_Politics/start/3656",
"thumbnail_url": "http://data.gdeltproject.org/televisionexplorer/thumbnails/CNNW_20190503_160000_Inside_Politics-003656.jpg"
}
]
}
}Request schema
No body schema
Response schema
#/definitions/gdelt.tvSearchResponseDoc
| Field | Type | Required | Enum | Bounds | Example | Description |
|---|---|---|---|---|---|---|
| code | integer | No | 200 | |||
| data | gdelt.TVClipSearchResponse | No | ||||
| data.clips | array | No | ||||
| data.clips[].caption | string | No | climate | |||
| data.clips[].caption_concepts | string | No | /M/0d063v | |||
| data.clips[].clip_url | string | No | https://archive.org/details/CNNW_20190503_160000_Inside_Politics/start/3656 | |||
| data.clips[].matched_at | string | No | 2019-05-03T17:00:56Z | |||
| data.clips[].onscreen_text | string | No | BREAKING NEWS: SEA LEVELS RISING | |||
| data.clips[].show | string | No | Inside Politics | |||
| data.clips[].show_start_at | string | No | 2019-05-03T16:00:00Z | |||
| data.clips[].station | string | No | cnn | |||
| data.clips[].thumbnail_url | string | No | http://data.gdeltproject.org/televisionexplorer/thumbnails/CNNW_20190503_160000_Inside_Politics-003656.jpg | |||
| data.clips[].transcript | string | No | climate change the | |||
| data.clips[].visual_entities | array | No | person,news,weather | |||
| data.count | integer | No | 3 | |||
| data.fetched_at | string | No | 2026-08-20T10:00:00Z | |||
| data.maxrecords | integer | No | 50 | |||
| data.query | string | No | cap:"climate change" station:CNN | |||
| data.sort | string | No | relevance | |||
| data.source_url | string | No | https://api.gdeltproject.org/api/v2/tvai/tvai?format=json&maxrecords=50&mode=clipgallery&query=cap%3A%22climate+change%22+station%3ACNN | |||
| data.station | string | No | CNN | |||
| msg | string | No | OK |
Use environment variables for secrets and keep Crawlora API keys server-side.
curl -X GET "https://api.crawlora.net/api/v1/gdelt/tv-search?station=CNN&sort=relevance" \
-H "x-api-key: $CRAWLORA_API_KEY"Crawlora is designed for responsible structured public web data workflows. Customers are responsible for using Crawlora in compliance with applicable laws, third-party rights, target-platform rules, and Crawlora terms.
Read Crawlora terms