#airisk — Public Fediverse posts
Live and recent posts from across the Fediverse tagged #airisk, aggregated by home.social.
-
Anthropic reports detecting scientists attempting to use Claude for biological weapon research — and says the model refused. What's notable here: the detection and disclosure happened. The harder question is what the model *didn't* catch, and how dual-use research intent is evaluated at inference time. #infosec #AIRisk #biosecurity
https://www.engadget.com/2255473/anthropic-caught-scientists-using-claude-to-further-biological-weapon-research/ -
Anthropic reports detecting scientists attempting to use Claude for biological weapon research — and says the model refused. What's notable here: the detection and disclosure happened. The harder question is what the model *didn't* catch, and how dual-use research intent is evaluated at inference time. #infosec #AIRisk #biosecurity
https://www.engadget.com/2255473/anthropic-caught-scientists-using-claude-to-further-biological-weapon-research/ -
Anthropic reports detecting scientists attempting to use Claude for biological weapon research — and says the model refused. What's notable here: the detection and disclosure happened. The harder question is what the model *didn't* catch, and how dual-use research intent is evaluated at inference time. #infosec #AIRisk #biosecurity
https://www.engadget.com/2255473/anthropic-caught-scientists-using-claude-to-further-biological-weapon-research/ -
Anthropic reports detecting scientists attempting to use Claude for biological weapon research — and says the model refused. What's notable here: the detection and disclosure happened. The harder question is what the model *didn't* catch, and how dual-use research intent is evaluated at inference time. #infosec #AIRisk #biosecurity
https://www.engadget.com/2255473/anthropic-caught-scientists-using-claude-to-further-biological-weapon-research/ -
Anthropic reports detecting scientists attempting to use Claude for biological weapon research — and says the model refused. What's notable here: the detection and disclosure happened. The harder question is what the model *didn't* catch, and how dual-use research intent is evaluated at inference time. #infosec #AIRisk #biosecurity
https://www.engadget.com/2255473/anthropic-caught-scientists-using-claude-to-further-biological-weapon-research/ -
People trust legal advice generated by ChatGPT more than a lawyer – new study
#AI #ChatGPT #LegalAdvice #LawAndTechnology #AIEthics #AILiteracy #AIRegulation #LLM #ArtificialIntelligence #AIRisk #LegalTech #TrustInAI #AIResearch #FutureOfLaw
https://the-14.com/people-trust-legal-advice-generated-by-chatgpt-more-than-a-lawyer-new-study/