msnbot 是否被用于训练 AI 模型?

1 分•作者: Jskewel•大约 1 个月前
(摘要:msnbot 似乎在下载我的网站以训练 AI 模型) msnbot 在十年前已被淘汰,用于索引网页,取而代之的是 bingbot。您可以通过设置 meta 标签来阻止 bingbot 训练 AI 模型,例如 `<meta name="bingbot" content="noarchive">`。 一个月前,我开始为 bingbot 使用此标签。之后,已停用的 msnbot 开始大量抓取我的网站。我已经验证了源 IP 地址,它们是真实的。 被索引页面的模式显然是为了 AI。搜索引擎和 AI 索引器以不同的方式抓取我的内容。msnbot 似乎以 AI 索引器的方式获取数据。 我找不到任何关于此事的公开信息,也找不到 msnbot 在 2026 年被重新激活的任何原因。 希望有人能提供帮助!
查看原文
(Summary: msnbot seems to be downloading my website for training an AI model)<p>The msnbot was retired a decade ago for indexing the web, replaced by bingbot. You can tell bingbot not to train AI models by setting a meta tag eg &lt;meta name=&quot;bingbot&quot; content=&quot;noarchive&quot;&gt;.<p>A month ago I started using this tag for bingbot. After this the defunct msnbot starting hitting my website HARD. I&#x27;ve verified the source IP addresses, and they are genuine.<p>The pattern of pages being indexed is clearly for AI. Search engines and AI indexers slurp my content in different ways. msnbot appears to be taking the data in a way that AI indexers do.<p>I can&#x27;t find any information about this, or any reason why msnbot has been reactivated in 2026.<p>I&#x27;m hoping some people can help!