Ashish K. Jha· @ashishkjha · X·· 20 天前AI 评分15
AI 导读
@AnthropicAI 的 @nc_frey 这些人是好人 他们有护栏。他们会关注谁在用他们的工具做潜在的恶意事情 而且他们会采取行动 你没看到其他人有类似报告,这件事本身不应该让你感到安心
正文
Folks at @AnthropicAI like @nc_frey are the good guys
They have guardrails. They pay attention to who is using their tools to do potentially malicious things
And they take action
The fact that you haven't seen similar reports from others should not be reassuring
Today we’re sharing some case studies of threat actors attempting to use Claude for malicious activity, including developing biological weapons. This is the first time any private company, to our knowledge, has publicly shared evidence of potential misuse of their platform for biological weapons development. Our goals are to enable the professional, legitimate life sciences research done with Claude, realize the incredible beneficial impacts for human health and medicine, and ensure that every AI platform is able to recognize and prevent misuse. Sophisticated, state-affiliated actors don’t tell Claude “I’m from adversarial nation X, build me a bioweapon.” They deliberately mask their origin and go to extreme lengths to obfuscate their intent and make their requests seem benign. When we detect and disrupt them, they route traffic to models with more permissive safeguards. Some attempts are easier to detect, like attempts to engineer viruses to infect humans or spread more efficiently. Our Life Sciences Verification Program is designed to enable beneficial research and prevent harmful use cases. We know our bio safeguards are not perfect, and can be a source of major frustration for professional researchers, but we are iterating constantly to make them better at preventing misuse and to improve the user experience for legitimate research. https://www.anthropic.com/threat-intelligence-report-september-2026在 X 查看被引用的帖子
来源:Ashish K. Jha · x.com