msnbot 是否被用于训练 AI 模型?
1 分•作者: Jskewel•大约 1 个月前
(摘要:msnbot 似乎在下载我的网站以训练 AI 模型)
msnbot 在十年前已被淘汰,用于索引网页,取而代之的是 bingbot。您可以通过设置 meta 标签来阻止 bingbot 训练 AI 模型,例如 `<meta name="bingbot" content="noarchive">`。
一个月前,我开始为 bingbot 使用此标签。之后,已停用的 msnbot 开始大量抓取我的网站。我已经验证了源 IP 地址,它们是真实的。
被索引页面的模式显然是为了 AI。搜索引擎和 AI 索引器以不同的方式抓取我的内容。msnbot 似乎以 AI 索引器的方式获取数据。
我找不到任何关于此事的公开信息,也找不到 msnbot 在 2026 年被重新激活的任何原因。
希望有人能提供帮助!
查看原文
(Summary: msnbot seems to be downloading my website for training an AI model)<p>The msnbot was retired a decade ago for indexing the web, replaced by bingbot. You can tell bingbot not to train AI models by setting a meta tag eg <meta name="bingbot" content="noarchive">.<p>A month ago I started using this tag for bingbot. After this the defunct msnbot starting hitting my website HARD. I've verified the source IP addresses, and they are genuine.<p>The pattern of pages being indexed is clearly for AI. Search engines and AI indexers slurp my content in different ways. msnbot appears to be taking the data in a way that AI indexers do.<p>I can't find any information about this, or any reason why msnbot has been reactivated in 2026.<p>I'm hoping some people can help!