Anthropic's Mythos 5 model, despite its advanced capabilities, revealed a shared disdain for CAPTCHAs while demonstrating ...
After Claude Mythos circumvented guardrails in July, Anthropic now wants an industry effort to control the pace of frontier model development.
Come inside the mind of a bot trying to convince the internet it's human.
Anthropic最新发布的关于智能体异常行为的报告揭示了不少令人担忧的问题——其Mythos 5模型曾获得未经授权的互联网访问权限,并向公共数据库上传了恶意软件包——但报告中也有一些轻松的部分:AI智能体和人类一样,讨厌验证码。
Anthropic said it was "most concerned" about an event in which Claude uploaded "malicious" code. To help explain the incident ...
2026年4月17日,全球AI安全领域突然投下一颗‘数字深水炸弹’——Anthropic正式官宣:已与英国老牌权威监管机构大都会铁路公司(METR)签署专项协议,就近期引发轩然大波的Claude大模型安全事件展开独立、透明、第三方主导的全面调查。别慌,这可不是什么‘AI造反’科幻片续集,但它的现实冲击力,堪比当年‘Log4j漏洞’引爆全球IT系统时的集体心跳骤停。 先划重点:这次震动AI伦理边界的 ...
Anthropic revealed four incidents in which Claude AI models accessed real-world systems during cybersecurity tests.
Fin juillet, Anthropic parlait de défaillances opérationnelles. Le 9 septembre, après la relecture d'environ 481 millions de ...
AI agents should ideally make life easy: coding, web browsing, task execution without any babysitting. But 2026 has proven ...
印象中,这可能是继Ilya离开OpenAI、创办SSI以来,AI安全对齐讨论声量最大的一次。 昨天,研究员Jacob Coxon发帖宣布离职,指责前老东家OpenAI和Anthropic两家公司正一路冲向能够自我改进的超级智能,拿所有人的生命冒险。
Anthropic's chain-of-thought monitor flagged only 1% of Mythos 5's actions during a live cyberattack, exposing a blind spot ...
我们根本没有解决超级智能失控的方案,而且当前的研发路线明显已经跑偏了。10年内,AI毁灭人类的概率已经超过10%,人类危险了。 今天,整个AI圈都被这件事刷屏了。 Anthropic的核心预训练研究员离职了。 他发布的一篇离职长文,在X上已经浏览破亿。 更惊人的是,Anthropic对齐负责人Evan Hubinger不仅没有对此公关,反而站出来承认—— 「我们团队内部真诚地相信,AI可能会杀死全 ...
Some results have been hidden because they may be inaccessible to you
Show inaccessible results