Data Hub Content Search
Search Data Hub content with a phrase and filters, then return matching page records with source metadata and content text.
概要
Data Hub Content Search turns a complete phrase into a structured page of matching content records. It brings together the result title, text, summary, source details, publication timing, and URLs so research teams can move from a topic to evidence without manually opening every search result.
The search is useful for monitoring a phrase across public content sources, collecting a reproducible result set, and comparing what different sources publish about the same topic. Filters for time, domain, account, and organization help narrow the result set to the slice that matters.
Data notes
Each record comes from the Data Hub content search index and represents one matched content result. Results are returned in the requested page order, while the search response also reports the estimated total number of matches. Some source metadata, cached links, classification flags, and full content fields can be missing, so those values may be null. The original URL, display URL, snippet, summary, and full content are separate fields and should not be treated as interchangeable.
The search matches the complete phrase against content titles and summaries. It can restrict results to included domains, remove excluded domains, and apply exact account or organization filters. The returned language and timestamps follow the source values; the API documentation does not define a common timezone for those timestamps.
What the results look like
Each row represents one content result:
| Result ID | Content title | Source site | Publication time | Content text | Original URL |
|---|---|---|---|---|---|
| 0 | 小米汽车交付节奏持续提升 | Example Site | 2026-08-16 14:23:10 | 小米汽车近期交付节奏持续提升。数据显示,相关门店试驾热度上升,交付安排正在稳步推进。 | https://www.example.com/article/123 |
| 1 | 新能源汽车销量观察 | Industry News | 2026-08-15 10:05:00 | 新能源汽车市场的月度销量变化受到多项因素影响。 | https://www.example.com/article/456 |
Additional fields include account and organization names, crawl time, language, snippets, summaries, display URLs, cached page URLs, site icons, and source classification flags.
Use cases
- For topic monitoring, compare “Content title”, “Publication time”, and “Original URL” across repeated searches to identify newly published coverage.
- For source research, group “Source site”, “Account name”, and “Organization name” to see which publishers contribute to a topic.
- For evidence review, read “Content text”, “Short snippet”, and “Content summary” together, then use “Display URL” or “Original URL” to open the source page.
適用範囲と境界
Each call returns one page of up to 200 matching content records. Pages 1 through 50 can be requested, covering at most 10,000 records per search.
- 適している用途
- When you need matching Data Hub content for a complete phrase, topic, brand, or event.
- When you need content titles, summaries, source details, publication times, and URLs for monitoring or research.
- 適さない用途
- When you need more than 10,000 matches from one search.
- When you need unrestricted full-text crawling beyond the records returned by Data Hub search.
失敗時の処理
作者が宣言した失敗時と再試行の動作です。連携時にはシステムプロンプトに含めることをおすすめします。
- 1An empty result list means that no content matched the phrase and filters; broaden the phrase or remove a restrictive filter.
- 2For authentication or authorization errors, verify the API Key configured for the execution environment before retrying.
- 3For rate-limit or temporary service errors, retry later without changing the search unless the error persists.
入力
この App の呼び出しに必要なパラメータで、manifest.json の input.schema から生成されています。
| フィールド | 業務名称 | 型 | 必須 | デフォルト | 列挙値/制約 | 例 | 説明 |
|---|---|---|---|---|---|---|---|
| keyword | Search phrase | string | はい | — | — | iphone | Required complete phrase matched against content titles and summaries. Do not leave it empty. |
| freshness | Publication time range | string | いいえ | noLimit | — | — | Limit results to a relative period such as oneMonth, a YYYY-MM-DD date, or a YYYY-MM-DD..YYYY-MM-DD date range. The default noLimit searches without a time limit. |
| include | Included domains | string | いいえ | — | — | — | Only return results from these domains. Separate multiple domains with | or ,; leave empty to include all domains. |
| exclude | Excluded domains | string | いいえ | — | — | — | Exclude results from these domains. Separate multiple domains with | or ,; leave empty to exclude none. |
| account_name | Account name | string | いいえ | — | — | — | Filter for an exact content account name; leave empty to use no account filter. |
| office_name | Organization name | string | いいえ | — | — | — | Filter for an exact organization name; leave empty to use no organization filter. |
| page | Result page | integer | いいえ | 1 | 1–50 | — | Page number to retrieve. Pages are 200 records each; page 1 is the first page and page 50 is the last page within the 10,000-record search limit. |
| sort_field | Sort field | string(enum) | いいえ | publishTime | publishTime / crawlerTime / id / cleanUpdateTime | — | Field used to order results. The default publishTime shows the newest published content first when paired with the default descending order. |
| sort_order | Sort direction | string(enum) | いいえ | desc | asc / desc | — | Use ascending or descending order for the selected sort field. The default desc returns the newest or largest values first. |
出力
1 件のレコードのフィールド構造で、manifest.json の output.schema から生成されています。
| フィールド | 業務名称 | 型 | 例 | 説明 |
|---|---|---|---|---|
| account_name | Account name | string | China News Network | Name of the account associated with the content; may be null when the source does not provide one. |
| cached_page_url | Cached page URL | string | — | URL of a cached version of the page when available; the source may return null. |
| content | Content text | string | 小米汽车近期交付节奏持续提升。数据显示,相关门店试驾热度上升,交付安排正在稳步推进。 | Full content text when available. It is independent from snippet and summary and may be null. |
| date_last_crawled | Last crawled time | string | 2026-08-17 09:12:30 | Most recent crawl timestamp in the source format; may be null. |
| date_published | Publication time | string | 2026-08-16 14:23:10 | Publication timestamp in the source format; may be null. |
| display_url | Display URL | string | https://www.example.com/article/123 | URL presented for the result; may be null. |
| id | Result ID | string | 0 | Stable identifier assigned to the search result; may be null when the source omits it. |
| is_family_friendly | Family-friendly flag | boolean | false | Whether the source marks the page as suitable for family audiences; may be null. |
| is_navigational | Navigational flag | boolean | false | Whether the source classifies the page as navigational; may be null. |
| language | Content language | string | zh | Language code reported for the content; may be null. |
| name | Content title | string | 小米汽车交付节奏持续提升 | Title of the matched content; may be null. |
| office_name | Organization name | string | — | Organization associated with the content; may be null. |
| site_icon | Site icon URL | string | — | URL of the source site's icon when available; may be null. |
| site_name | Source site | string | Example Site | Name of the site or source that published the content; may be null. |
| snippet | Short snippet | string | 小米汽车近期交付节奏持续提升。 | Short extract returned for the result; it is independent from content and summary and may be null. |
| summary | Content summary | string | 小米汽车近期交付节奏持续提升,相关门店试驾热度上升。 | Summary returned for the result; it is independent from content and snippet and may be null. |
| url | Original URL | string | https://www.example.com/article/123 | Original URL of the matched content; may be null. |
レコード Schema
出力はレコード単位で 1 件ずつ返されます。 detail.output.idFieldHint
{
"type": "object",
"properties": {
"account_name": {
"type": "string",
"title": "Account name",
"description": "Name of the account associated with the content; may be null when the source does not provide one.",
"prefill": "China News Network"
},
"cached_page_url": {
"type": "string",
"title": "Cached page URL",
"description": "URL of a cached version of the page when available; the source may return null."
},
"content": {
"type": "string",
"title": "Content text",
"description": "Full content text when available. It is independent from snippet and summary and may be null.",
"prefill": "小米汽车近期交付节奏持续提升。数据显示,相关门店试驾热度上升,交付安排正在稳步推进。"
},
"date_last_crawled": {
"type": "string",
"title": "Last crawled time",
"description": "Most recent crawl timestamp in the source format; may be null.",
"prefill": "2026-08-17 09:12:30"
},
"date_published": {
"type": "string",
"title": "Publication time",
"description": "Publication timestamp in the source format; may be null.",
"prefill": "2026-08-16 14:23:10"
},
"display_url": {
"type": "string",
"title": "Display URL",
"description": "URL presented for the result; may be null.",
"prefill": "https://www.example.com/article/123"
},
"id": {
"type": "string",
"title": "Result ID",
"description": "Stable identifier assigned to the search result; may be null when the source omits it.",
"prefill": "0"
},
"is_family_friendly": {
"type": "boolean",
"title": "Family-friendly flag",
"description": "Whether the source marks the page as suitable for family audiences; may be null.",
"prefill": false
},
"is_navigational": {
"type": "boolean",
"title": "Navigational flag",
"description": "Whether the source classifies the page as navigational; may be null.",
"prefill": false
},
"language": {
"type": "string",
"title": "Content language",
"description": "Language code reported for the content; may be null.",
"prefill": "zh"
},
"name": {
"type": "string",
"title": "Content title",
"description": "Title of the matched content; may be null.",
"prefill": "小米汽车交付节奏持续提升"
},
"office_name": {
"type": "string",
"title": "Organization name",
"description": "Organization associated with the content; may be null."
},
"site_icon": {
"type": "string",
"title": "Site icon URL",
"description": "URL of the source site's icon when available; may be null."
},
"site_name": {
"type": "string",
"title": "Source site",
"description": "Name of the site or source that published the content; may be null.",
"prefill": "Example Site"
},
"snippet": {
"type": "string",
"title": "Short snippet",
"description": "Short extract returned for the result; it is independent from content and summary and may be null.",
"prefill": "小米汽车近期交付节奏持续提升。"
},
"summary": {
"type": "string",
"title": "Content summary",
"description": "Summary returned for the result; it is independent from content and snippet and may be null.",
"prefill": "小米汽车近期交付节奏持续提升,相关门店试驾热度上升。"
},
"url": {
"type": "string",
"title": "Original URL",
"description": "Original URL of the matched content; may be null.",
"prefill": "https://www.example.com/article/123"
}
},
"required": [],
"additionalProperties": false
}連携方法
この App は MCP、API、SDK、ファイルエクスポートのいずれでも連携でき、どのチャネルも同じ能力と料金を共有します。すべてのリクエストは Authorization: Bearer ヘッダーで認証し、認証情報には API Key(長期有効。Data Hub コンソールで作成)を使用します。MCP クライアントは OAuth によるキー不要のログインにも対応しています。CLI や Skill など、さらに多くの連携方法も準備中です。
MCP(Model Context Protocol)を使うと、Claude や Cursor などの AI クライアントからこの App を直接呼び出せます。クライアントと認証方式を選び、下の設定をコピーしてください。
クライアント設定
Bearer の後の値を長期有効な API Key に置き換えてください。あらゆるクライアント、CI、ヘッドレス環境で利用できます。
{
"mcpServers": {
"meme_today__data-hub-content-search": {
"type": "http",
"url": "https://mcp-v2.octoparse.com?pin=meme_today/data-hub-content-search",
"headers": { "Authorization": "Bearer <YOUR_API_KEY>" }
}
}
}AI に設定を任せる
設定を手作業で編集したくない場合は、インストールプロンプトをコピーして任意の AI クライアントに貼り付けてください。AI が自分のやり方でセットアップを完了します(プロンプトは AI があなたに API Key を尋ねる形になっているため、認証情報がチャット履歴や共有設定に残りません)。
detail.access.mcp.composeHint
料金
正常に返されたレコードの件数に応じて課金されます。失敗した実行には課金されません。 20 件を 1 つの課金単位とし、端数は 20 件に切り上げて課金されます。
複数の課金イベントはそれぞれ独立して累計されます。詳細は各項目をご覧ください。失敗した実行には課金されません。
Credits を確認今すぐ試す
パラメータを入力して実行してください。結果は実際の呼び出しによるものです。