AI也能拒绝挨骂了!Claude遭持续恶意辱骂可主动结束对话

快科技
Yesterday

快科技10月9日消息,据Android Authority报道,AI公司Anthropic近日更新了Claude使用政策,明确禁止用户持续且无必要地对AI模型进行恶意辱骂或残酷对待,新规将于2026年11月12日正式生效。

平时使用AI时,如果遇到回答错误、反复理解不了问题的情况,不少用户难免会抱怨几句。不过,随着这项新规出台,如果有人长时间对Claude进行毫无意义的恶意辱骂,AI可能会直接结束当前对话。

Anthropic在官方公告中解释,此次新增规定主要针对极端情况,例如用户没有明确目的,却持续对模型进行恶意攻击。

对于正常使用场景,用户仍然可以批评Claude、质疑回答结果,或者因AI表现不佳而表达不满。此外,涉及负面题材的创作、正常模型测试和研究活动也不会因此受到限制。

值得注意的是,Claude主动结束对话的能力其实早已有之。

早在2025年8月,Anthropic就已经允许部分Claude模型在极端情况下终止聊天。当用户持续进行有害或辱骂性互动,且多次尝试引导仍然无效时,Claude可以选择结束当前对话。

即便如此,用户依然可以重新开启聊天,其他对话也不会受到影响。官方表示,主动结束对话仍将是处理此类行为的主要方式,并未规定只要辱骂AI就会立即封禁账号。

此次调整还有一个值得关注的地方,那就是AI公司开始更加重视用户与模型之间的互动边界。

过去,AI产品的使用规范更多关注模型不能生成哪些内容,以及用户不能利用AI从事哪些活动。如今,用户如何对待AI本身,也逐渐被纳入平台的使用规则。

Anthropic此前曾将相关机制与潜在的“AI福祉”研究联系起来,不过该公司也承认,目前尚无法确定AI是否具有道德地位。

Disclaimer: Investing carries risk. This is not financial advice. The above content should not be regarded as an offer, recommendation, or solicitation on acquiring or disposing of any financial products, any associated discussions, comments, or posts by author or other users should not be considered as such either. It is solely for general information purpose only, which does not consider your own investment objectives, financial situations or needs. TTM assumes no responsibility or warranty for the accuracy and completeness of the information, investors should do their own research and may seek professional advice before investing.

Most Discussed

  1. 1
     
     
     
     
  2. 2
     
     
     
     
  3. 3
     
     
     
     
  4. 4
     
     
     
     
  5. 5
     
     
     
     
  6. 6
     
     
     
     
  7. 7
     
     
     
     
  8. 8
     
     
     
     
  9. 9
     
     
     
     
  10. 10