采集卡片

Bing Keyword Ideas 采集器

同样的展开,跑在 Bing 的自动补全上——DuckDuckGo 用的也是它。给种子词一份便宜的第二意见:两个引擎都提示的短语,是人们真的会输入的;只有 Bing 有的那条,往往是一个值得拿下的空档。

本页内容

它采什么

搜索与 SEO

同样的展开,跑在 Bing 的自动补全上——DuckDuckGo 用的也是它。给种子词一份便宜的第二意见:两个引擎都提示的短语,是人们真的会输入的;只有 Bing 有的那条,往往是一个值得拿下的空档。 它是一张采集器卡片,也就是说全部约定只有一条:填好表单,回来一张表。不会打开浏览器窗口,不占用你套餐里的任何自动化名额,请求走的是普通 HTTP,带着真实浏览器的 TLS 指纹。

用你自己的地址跑就很好;运行时可以选代理,但这个目标并不需要。无论走的是哪条路,每一行都会记下来——见下面的溯源列。

你要提供什么——9 项

应用画出的那张表单,直接读自这张卡片本身。下面每一条提示,都是运行对话框里显示在该字段旁边的那一条,所以这里没有任何脱离产品另写一遍的说明。

它唯一需要的那个答案

seeds
Seeds · list · 必填

The words you already know. One per line, and each is expanded into a few dozen queries of its own — at most 25 per run. Two or three real ones beat a long list of guesses.

其余的都可填可不填

useAlphabet
A–Z · boolean

The classic sweep: 26 requests, and between them they surface most of what the engine will complete this seed into.

useDigits
0–9 · boolean

Versions, years, model numbers and "top 10" phrasings.

useQuestions
Questions · boolean

What people ask ABOUT it — the phrasing that wins featured snippets and answers a video title.

useCommercial
Buying intent · boolean

The words somebody uses once they have decided to spend money. The smallest set and usually the most valuable.

usePrepositions
Modifiers · boolean

What it is for, what it works with, what it is unlike — where the long tail lives.

market
Market · text

Bing’s market code — language and country together, like en-us, en-gb or de-de. One field rather than two, because this is the shape Bing asks for.

maxPerSeed
Keywords per seed · number

Where to stop for each seed. Duplicates are removed before this is counted, so it is a cap on distinct keywords rather than on rows collected.

mustContain
Must contain · list

Keep only keywords containing one of these words. Left empty, everything the engine returned is kept — which is usually what you want the first time you run a seed.

会拿回什么——6 列

每次运行生成一个数据集——一张带类型的表,归你的工作区所有:可以排序、筛选、在网格里直接编辑、整张导出,也可以通过本地 API 读回来。下面就是它创建时带的列。

来自 Bing Keyword Ideas 的 2 列

键名类型
keywordKeywordtext
searchTermSeedtext

每个采集器都会写的 4 列

每张卡片上都是同样这四列,好让一张表在几个月后仍能回答它的行是怎么来的:来自哪个服务、什么时候、请求带的是哪个环境的身份,以及那个身份当时是否已登录。

键名类型
platformPlatformtext
collected_atCollected atdatetime
profileCollected byprofile
logged_inSigned inselect

跑它的四种方式

「采集器」标签页。 挑中这张卡片,填好表单,按「开始」。从这里发起的每一次运行都会生成一张新表,以你搜索的内容和时间命名。想先试试也可以——它只采一页,什么都不写,并告诉你哪些列拿回了内容。

问助手。 它手里有整个目录,所以这张卡片是一句话,而不是一张表单。它会按你说的话把上面那些参数填好,并在开跑之前拿给你看。

通过 MCP 或本地 API。 同一张卡片,从编程 Agent 里调用——MCP 服务器上的 argus_run_scraper,或者本地 APIPOST http://127.0.0.1:39219/v1/scrapers/run。旁边有一个什么都不写入的示例调用。

argus_run_scraper — Bing Keyword Ideas
{
  "kind": "bing_keywords",
  "inputs": {
    "seeds": [
      "…"
    ]
  }
}

作为工作流里的一个步骤。 Run scraper 步骤把这个采集器放进一条自动化流程的中间——先采、再筛、再发信——整棵流程照样一个窗口都不开。它也是唯一能推翻「一次运行一张表」规则的调用方:指定一张你命名的表,它会在表不存在时按上面的列建好,之后每次运行都写进这张表,追加或者按你指定的匹配列就地更新。这个步骤会把表名、表 id 和行数交给下一步。

它到哪里为止

它不会去驱动页面。凡是需要真实浏览器的事——背后没有列表接口的网站,或者任何操作你自己账号的事——都该交给自动化流程。那是另一件工具,而不是这件工具的劣化版。

它不登录任何地方。未登录读取公开页面是已成定论的做法——Bright Data 未登录采集 Meta 并且胜诉,hiQ 登录后采集 LinkedIn 则败诉——所以这张卡片不要账号,也不持有账号。