Baidu API endpoint
Use Crawlora's Search Baidu web results API to extract supported public Baidu data as structured JSON. This page includes request parameters, cURL examples, response schema, error behavior, credit cost, and a Playground link for testing before integration.
/baidu/searchReturns normalized Baidu result cards for a query, read from Baidu's mobile search page: title, destination URL, snippet, publisher, and date, with page-based pagination. Each result's `type` is `web` (standard pages), `video` (video cards), or `baike` (Baidu Baike encyclopedia entries); answer widgets such as weather, entity panels, and recommendation lists are not returned. Baidu challenges a portion of requests with a security verification page; the service retries across browser profiles and returns 503 if every attempt is challenged, so callers should retry on 503. Developers commonly use this endpoint for data enrichment, monitoring, research dashboards, internal automation, and agent-native workflows that need repeatable structured public web data. Authentication uses the documented Crawlora headers, and usage is metered with the credit cost shown on this page.
Request parameters are generated from the active endpoint catalog. Required values must be sent before Crawlora can call the upstream public web data source.
| Parameter | Type | Required | Default | Description | Example |
|---|---|---|---|---|---|
| q | string | Yes | Search query | openai | |
| page | integer | No | 1 | 1-based result page; defaults to 1, maximum 10 Minimum: 1. Maximum: 10. | 1 |
| x-api-key (header) | string | Yes | API key required |
curl -X GET "https://api.crawlora.net/api/v1/baidu/search?q=openai&page=1" \ -H "x-api-key: $CRAWLORA_API_KEY"
Send your scraping API key in the x-api-key header. Use the console API Keys page to rotate or select the active key.
Endpoint usage is metered in credits. The plan prices, included credits, limits, and overage rates below match the active backend billing configuration.
| Plan | Price | Included credits | Daily cap | Rate limit | Overage |
|---|---|---|---|---|---|
| Free | $0/mo | 2,000 | 500 daily credits | 5/min | No overage |
| Starter | $9/mo | 20,000 | 5,000 daily credits | 15/min | $0.75/1,000 overage credits when enabled |
| Growth | $29/mo | 100,000 | 25,000 daily credits | 45/min | $0.45/1,000 overage credits when enabled |
| Pro | $79/mo | 400,000 | No daily cap | 120/min | $0.30/1,000 overage credits |
| Business | $199/mo | 1,200,000 | No daily cap | 300/min | $0.20/1,000 overage credits |
| Enterprise | $499/mo | 5,000,000 | No daily cap | 1,000/min | $0.12/1,000 overage credits |
This endpoint is executed through Crawlora's managed scraping infrastructure.
Some targets require real browser execution because the data is loaded through JavaScript, dynamic rendering, or interaction-like browser behavior.
For supported endpoints, Crawlora can route requests through a managed browser cluster. This allows Crawlora to execute JavaScript, load dynamic content, apply browser-level request behavior, and normalize the rendered result into JSON.
You do not need to operate your own Playwright, Puppeteer, Chrome, proxy, queue, or retry infrastructure.
- Each result's `type` is one of `web` (standard pages), `video` (video cards), or `baike` (Baidu Baike encyclopedia entries). Answer widgets such as weather panels, entity panels, Tieba cards without a link, and recommendation lists are not results and are not returned. - `source` and `date` are the publisher name and date Baidu shows on the card; `date` is Baidu's own text, such as `2025年10月30日` or `1月19日`, and is omitted when Baidu shows none. - A single page holds about 10 cards, of which roughly 8 to 10 are returned as results. - Baidu challenges a share of requests, and a repeating client address, with a security verification page. The service retries up to three times across mobile browser profiles on fresh connections and never tries to solve the verification; if every attempt is challenged it returns `503`, and the call can simply be retried. - A response page with no result cards and no verification marker is reported as `503` (parser drift), never as an empty success. - `pagination.next_page` is present only when Baidu offers a next page and `page` is below `10`.
Crawlora does not silently return bad data when the upstream page cannot be used.
| Status | Common failure case |
|---|---|
| 400 | Invalid input or missing required parameter |
| 429 | Plan or endpoint rate limit exceeded |
| 500 | Internal execution error |
| 502 | Upstream platform failed, returned unusable HTML, or served a challenge page that could not be resolved |
When possible, Crawlora returns structured error context so your integration can retry, back off, or inspect the request.
| Status | Description | Schema |
|---|---|---|
| 400 | Missing or invalid parameters | #/definitions/app.Response |
| 429 | Rate limit exceeded | #/definitions/app.Response |
| 500 | Internal server error | #/definitions/app.Response |
| 503 | Baidu upstream request failed or was challenged with a security verification | #/definitions/app.Response |
{
"code": 200,
"msg": "OK",
"data": {
"query": "openai",
"results": [
{
"position": 1,
"type": "web",
"title": "OpenAI万亿估值之谜:营收百亿、支出千亿与市值神话的...",
"url": "https://baijiahao.baidu.com/s?id=1847370812297480431",
"description": "据路透社援引知情人士爆料,OpenAI正推进最早于2027年的上市计划",
"source": "新浪财经",
"date": "2025年10月30日"
}
],
"pagination": {
"page": 1,
"next_page": 2
}
}
}Request schema
No body schema
Response schema
#/definitions/baidu.searchResponseDoc
| Field | Type | Required | Enum | Bounds | Example | Description |
|---|---|---|---|---|---|---|
| code | integer | No | 200 | |||
| data | baidu.SearchResponse | No | ||||
| data.pagination | baidu.SearchPagination | No | ||||
| data.pagination.next_page | integer | No | 2 | |||
| data.pagination.page | integer | No | 1 | |||
| data.query | string | No | openai | |||
| data.results | array | No | ||||
| data.results[].date | string | No | 2025年10月30日 | |||
| data.results[].description | string | No | 据路透社援引知情人士爆料,OpenAI 正推进最早于2027年的上市计划 | |||
| data.results[].position | integer | No | 1 | |||
| data.results[].source | string | No | 新浪财经 | |||
| data.results[].title | string | No | OpenAI 万亿估值之谜 | |||
| data.results[].type | string | No | web, video, baike | web | ||
| data.results[].url | string | No | https://baijiahao.baidu.com/s?id=1847370812297480431 | |||
| msg | string | No | OK |
Use environment variables for secrets and keep Crawlora API keys server-side.
curl -X GET "https://api.crawlora.net/api/v1/baidu/search?q=openai&page=1" \
-H "x-api-key: $CRAWLORA_API_KEY"Crawlora is designed for responsible structured public web data workflows. Customers are responsible for using Crawlora in compliance with applicable laws, third-party rights, target-platform rules, and Crawlora terms.
Read Crawlora terms