home.social

#safeguards — Public Fediverse posts

Live and recent posts from across the Fediverse tagged #safeguards, aggregated by home.social.

  1. Access to #ClaudeFable 5 and #ClaudeMythos 5 has been #restored after the #USgovernment lifted #exportcontrols. The controls were imposed due to a report of a method bypassing #Fable 5’s #safeguards, which has now been addressed with an improved safety classifier. #Anthropic is collaborating with the US government and industry partners to develop a shared framework for assessing and mitigating AI model “jailbreaks.” anthropic.com/news/redeploying #tech #media #news

  2. disabilitynewsservice.com/disa. "The bill’s opponents - including Baroness Grey-Thompson - say their #amendments are crucial to add #safeguards to a private members’ bill that is vague, unsafe & poorly-drafted."

  3. #OpenAI reports that over a million #ChatGPT users discuss #suicidalthoughts weekly, highlighting the need for improved #mentalhealthsupport. The company claims its latest #GPT5 model responds more appropriately to #mentalhealth issues, with a 65% improvement in desirable responses and 91% compliance with desired behaviours in #suicidalconversations. OpenAI is also implementing new #evaluations and #safeguards to address emotional reliance. techcrunch.com/2025/10/27/open #tech #media #news