抓取与访问

AI Bot 策略生成器

根据内容授权选择,生成明确的 robots.txt 爬虫策略。

免费 · 即时

按用途选择允许或禁止

禁止训练爬虫不等于禁止搜索可见性爬虫——它们是不同 UA。想被 AI 搜索推荐,至少保留第二组。

模型训练GPTBot, ClaudeBot, Google-Extended, Applebot-Extended, Bytespider, Meta-ExternalAgent, CCBot
AI 搜索可见性OAI-SearchBot, PerplexityBot
用户代抓 / 助手ChatGPT-User, Amazonbot
# AI 爬虫策略(由 InsightWonder Tools 生成)

# 模型训练: 允许
User-agent: GPTBot
User-agent: ClaudeBot
User-agent: Google-Extended
User-agent: Applebot-Extended
User-agent: Bytespider
User-agent: Meta-ExternalAgent
User-agent: CCBot
Disallow:

# AI 搜索可见性: 允许
User-agent: OAI-SearchBot
User-agent: PerplexityBot
Disallow:

# 用户代抓 / 助手: 允许
User-agent: ChatGPT-User
User-agent: Amazonbot
Disallow:
01

这个工具检查什么?

按 AI 运营方和用途(训练、搜索、用户代抓)逐项选择允许或禁止,生成可直接粘贴的 robots.txt 片段,使用正确的真实 UA 令牌。

02

结果该怎么看?

禁止训练爬虫(GPTBot、ClaudeBot)不等于禁止搜索可见性爬虫(OAI-SearchBot、PerplexityBot)——它们是不同用途的不同 UA。生成器刻意把两类分开。

03

下一步该做什么?

把片段粘贴进 robots.txt,再用 AI 爬虫访问检查确认服务器层的真实表现。

参考来源: RFC 9309 Robots Exclusion Protocol · OpenAI crawler docs

InsightWonder

访问规则会在你不知情时改变

robots.txt 的改动和 CDN 的 bot 规则通常没人通知你。InsightWonder 会定期复检,某个爬虫开始被拦时告警。

开始深度分析
  • 定期复检 + 变化告警
  • 跨 5 个 AI 引擎的可见度与引用分析
  • 按影响排序的修复建议,并生成内容
AI Bot 策略生成器 —— 生成 robots.txt 规则