Cloudflare 推出 BotBase for Operators
BotBase for Operators: A clearer path to joining Cloudflare's directory of bots and agents
Last month, on our second Content Independence Day, we announced a couple of features designed to give website owners more visibility and control over automated traffic: BotBase added a searchable directory of known bots to the Cloudflare dashboard, while Business Insights helped owners understand how crawlers interact with their content. We know that the ecosystem of bots is vast, making it all the more important for site owners to be able to manage bot traffic sustainably.
上个月,在我们的第二个“内容独立日”,我们宣布了一系列旨在让网站所有者对自动化流量有更多可见性和控制权的功能:BotBase 在 Cloudflare 仪表盘中新增了一个可搜索的已知机器人目录,而 Business Insights 则帮助所有者理解爬虫如何与其内容互动。我们知道机器人生态系统庞大,这使得网站所有者能够可持续地管理机器人流量变得尤为重要。
But this ecosystem goes both ways. While website owners need to decide which automated traffic they allow, bot operators need a clear way to identify themselves, explain what their bots do, and keep that information current. BotBase works best when both sides can participate.
但这个生态系统是双向的。网站所有者需要决定允许哪些自动化流量,而机器人运营商则需要一种清晰的方式来标识自己、解释其机器人的功能,并保持这些信息的更新。当双方都能参与时,BotBase 才能发挥最佳效果。
When we launched BotBase, we said we would build tools to bring bot operators into this ecosystem. Until now, their experience largely ended at submission. After pressing submit, an operator had no easy way to check the submission's status, understand why it was rejected, or update an existing entry. Today, we start to change that with the launch of BotBase for Operators, tackling what bot operators need first: transparency.
当我们推出 BotBase 时,我们说过会构建工具,将机器人运营商引入这个生态系统。直到现在,他们的体验基本止步于提交。点击提交后,运营商没有简便的方法来检查提交状态、了解被拒原因或更新现有条目。今天,我们通过推出 BotBase for Operators 开始改变这一现状,首先解决机器人运营商最迫切的需求:透明度。
A new home for bot submissions
机器人提交的新家园
Imagine you’re a bot operator looking to submit your bot to BotBase. Where on the dashboard would you look for such a submission form? Previously, the form lived under Manage Account → Configurations, which tied the bot clearly to your account, but didn’t acknowledge its connection to the bots ecosystem.
想象一下,你是一位机器人运营商,想要向 BotBase 提交你的机器人。在仪表盘中,你会去哪里找这样的提交表单?以前,表单位于“管理账户 → 配置”下,这虽然将机器人与你的账户明确关联,但并未体现其与机器人生态系统的联系。
Starting today, the bot submission experience has a home next to the rest of your bot and trust tools: Protect & Connect → Application Security → BotBase (new!). All customers can access this today directly from the Cloudflare dashboard.
从今天起,机器人提交体验有了一个与你的其他机器人和信任工具相邻的家:Protect & Connect → Application Security → BotBase(新增!)。所有客户今天都可以直接从 Cloudflare 仪表盘访问。
Here, we’ve split BotBase for Operators by use case:
在这里,我们按使用场景将 BotBase for Operators 分成了几个部分:
- Bots directory — browse, search, and filter the bots Cloudflare already tracks (the same catalogue you can explore on Cloudflare Radar).
- Submission form — submit a new bot.
- Submission history — track everything you have submitted.
- 机器人目录 — 浏览、搜索和筛选 Cloudflare 已跟踪的机器人(与你在 Cloudflare Radar 上可探索的目录相同)。
- 提交表单 — 提交新的机器人。
- 提交历史 — 跟踪你提交的所有内容。
Finding BotBase solves the "where" problem. The "what happens next" problem is the one that we’ve heard is deeply important to bot operators, so we’ll cover that in the rest of this post.
找到 BotBase 解决了“在哪里”的问题。而“接下来会发生什么”的问题,我们听说对机器人运营商来说至关重要,所以我们将在本文其余部分进行阐述。
See where your submission stands
查看你的提交状态
We spoke to many bot operators, and the resounding feedback was this: submitting a bot feels like a black box. You fill in the form, press submit, and wait, with no way to tell whether anything happened next.
我们与许多机器人运营商交流过,得到的强烈反馈是:提交机器人就像面对一个黑箱。你填写表单,点击提交,然后等待,却无法知道接下来是否发生了什么。
Now, the Submission history tab shows every bot submitted from your account, each with a clear status:
现在,“提交历史”标签页会显示您账户提交的每个机器人,每个都带有清晰的状态:
- Waiting for review — we have received your submission and it is in our queue.
- Accepted — we have reviewed it and your bot is now tracked in the directory.
- Rejected — something in the submission needs to change. We tell you why, with steps you can act on, so you can fix it and resubmit.
- 等待审核 — 我们已收到您的提交,正在排队处理中。
- 已接受 — 我们已审核通过,您的机器人现在已被目录收录。
- 已拒绝 — 提交内容中有需要修改的地方。我们会告知您原因,并提供可操作步骤,以便您修复后重新提交。
Open any submission to see its full details. If it was rejected, you will see the reason why. If it was accepted but we adjusted how your bot is classified, you will see what we changed.
打开任意提交即可查看完整详情。如果被拒绝,您会看到原因。如果被接受但我们调整了机器人的分类方式,您会看到我们做了哪些更改。
Previously, operators would need to email support just to ask whether their bot got reviewed or to check on their submission's progress. That's exactly the gap we’re closing with this new tab.
以前,运营者需要发送邮件给支持团队,仅仅为了询问机器人是否已审核或查看提交进度。这正是我们通过这个新标签页要弥补的缺口。
Today, the submission form is no longer a black box. Every operator can now view the record of every bot they've submitted starting from today’s launch, with a status you can check anytime. We also provide a way to filter “My bots,” from the Bots directory screen, so you can see all bots that have been submitted under the account with which you’re currently logged in.
如今,提交表单不再是黑盒。从今天上线开始,每位运营者都可以查看自己提交的每个机器人的记录,并随时查看状态。我们还提供了从“机器人目录”页面筛选“我的机器人”的功能,这样您可以看到当前登录账户下提交的所有机器人。
Keep your bot's information up to date
保持机器人信息最新
A bot's identification details can change over time. You might redesign your website and end up hosting your IP list at a new endpoint. Or you might move from an IP allowlist to signing your traffic with Web Bot Auth, and need your entry to match. Before today, the only way to reflect either change was to fill out the whole form again and submit a brand-new entry. Now, you can edit a submission you have already made.
机器人的识别信息可能会随时间变化。您可能重新设计网站,将IP列表托管到新的端点;或者从IP白名单切换到使用Web Bot Auth对流量进行签名,需要条目与之匹配。在今天之前,要反映这些更改,唯一的方法是重新填写整个表单并提交全新条目。现在,您可以编辑已提交的条目。
You can also cancel a submission that is still waiting for review.
您也可以取消仍在等待审核的提交。
We encourage every operator to keep their bot's information current. Accurate details are a key component of how a bot earns and keeps Verified status, which increasingly determines whether sites across Cloudflare's network can easily allow it based on its behavior. Of course, it is ultimately up to the individual site owner to decide what traffic is allowed and what is not.
我们鼓励每位运营者保持机器人信息的最新状态。准确的信息是机器人获得并保持“已验证”状态的关键组成部分,这越来越决定Cloudflare网络上的网站能否根据其行为轻松允许它。当然,最终由各个网站所有者决定允许或不允许哪些流量。
A submission form built on an updated, pragmatic taxonomy
基于更新、实用分类法的提交表单
Picture a bot. Maybe it only crawls pages to build a search index. Maybe it also acts on a user's behalf, or pulls in data for something else entirely. How it uses what it reads matters just as much as what it does.
想象一个机器人。它可能只爬取页面以构建搜索索引。它可能代表用户执行操作,或者为其他完全不同的目的拉取数据。它如何使用读取的内容与其行为本身同样重要。
The new intake form asks you to describe your bot the way it actually behaves. It follows the same behavior and content use model we introduced on July 1, so instead of squeezing your bot into a single label, you now tell us three things.
新的提交表单要求你按照你的机器人实际行为来描述它。它遵循我们于7月1日推出的相同行为与内容使用模型,因此,你不再需要将你的机器人压缩进单一标签,而是需要告诉我们三件事。
First, what your bot does. Maybe it only does one thing, like indexing pages for search. Maybe it's an agent acting on a user's behalf, or it collects data, trains models, or supports SEO tools. You can select every behavior that applies, not just the closest match.
首先,你的机器人做什么。也许它只做一件事,比如为搜索索引页面。也许它是一个代表用户行动的代理,或者它收集数据、训练模型,或支持SEO工具。你可以选择所有适用的行为,而不仅仅是最接近的匹配项。
Second, how it uses what it reads. A crawler that skims a page for a search snippet is not the same as one that stores that page to train a model. You tell us the level of content use your bot needs, using the same Content Signals model website owners already use to set their own rules. For example, a site's robots.txt might read Content-Signal: search=yes, ai-train=no, use=reference, telling every crawler it's fine to index the page for search and keep a reference, but not to train a model on it. Your bot's content-use declaration is what gets checked against exactly that kind of preference.
其次,它如何使用所读取的内容。一个仅为搜索摘要而浏览页面的爬虫,与一个存储该页面以训练模型的爬虫是不同的。你需要使用网站所有者已经用来设定自身规则的同一Content Signals模型,告诉我们你的机器人所需的内容使用级别。例如,一个网站的robots.txt可能包含Content-Signal: search=yes, ai-train=no, use=reference,告诉每个爬虫,为搜索索引页面并保留参考是可以的,但不要用它训练模型。你的机器人的内容使用声明正是用来与这类偏好进行核对的。
Third, who's actually running it. If you operate your bot yourself, straight from your own infrastructure, like a search engine crawling the web to build its own index, that's direct. If you run a platform other companies build on, carrying their traffic without being the one who decided to send it, that's an intermediary. Picture a general-purpose AI assistant fetching a page because someone typed a question into a different company's app built on that assistant's API: the assistant operator runs the infrastructure, but it was someone else's product that decided to send the request. (You can read more about these classifications here.)
第三,谁在真正运行它。如果你自己操作你的机器人,直接来自你自己的基础设施,比如一个搜索引擎爬取网络以构建自己的索引,那就是直接操作。如果你运行一个其他公司构建的平台,承载他们的流量,但不是决定发送流量的一方,那就是中介。想象一个通用AI助手因为有人在另一家公司的应用(该应用构建在该助手的API上)中输入了问题而获取一个页面:助手运营商运行基础设施,但决定发送请求的是别人的产品。(你可以在这里了解更多关于这些分类的信息。)
That's the full picture: what your bot does, how it treats what it reads, and who's behind it, described as it actually is instead of squeezed into one label. The clearer that picture, the more accurately website owners can decide how to treat your bot.
这就是全貌:你的机器人做什么,它如何处理所读取的内容,以及谁在背后,按实际情况描述,而不是压缩进一个标签。画面越清晰,网站所有者就能越准确地决定如何对待你的机器人。
Faster, more consistent review
更快、更一致的审核
Operators also asked for faster reviews. We hear you on this, too.
运营商也要求更快的审核。我们也听到了你们的这一需求。
The number of new bots submitted each year has grown sharply — increasing about 7 times in volume since 2023 — and reviewing every one of them by hand doesn't scale at that pace. Until now, every submission followed the same fully manual path: someone on our team checks it against an internal rubric and makes a judgment call. That kind of review is thorough, but it doesn't scale.
每年提交的新机器人数量急剧增长——自2023年以来数量增加了约7倍——而逐一人工审核在如此速度下无法扩展。直到现在,每次提交都遵循同样的全手动流程:我们团队中的某人根据内部标准进行检查并做出判断。这种审核很彻底,但无法扩展。
更进一步:量化金融体系
看懂新闻只是起点——沿量化金融路径,把它变成能交付的工程能力