home.social

#safeguards — Public Fediverse posts

Live and recent posts from across the Fediverse tagged #safeguards, aggregated by home.social.

fetched live
  1. #USA and #China: #securityexperts are proposing #nuclearstyle #safeguards for #AIrisks, including red lines around #nuclearsystems, #humancontrol over consequential #cyberattacks, and a #hotline for incidents involving #autonomousAI. The recommendations, published ahead of planned government-level AI talks, highlight the growing concerns about the potential for AI to escalate conflicts and the need for communication protocols between rivals. reuters.com/world/china/us-chi #tech #news #ainews

  2. #USA and #China: #securityexperts are proposing #nuclearstyle #safeguards for #AIrisks, including red lines around #nuclearsystems, #humancontrol over consequential #cyberattacks, and a #hotline for incidents involving #autonomousAI. The recommendations, published ahead of planned government-level AI talks, highlight the growing concerns about the potential for AI to escalate conflicts and the need for communication protocols between rivals. reuters.com/world/china/us-chi #tech #news #ainews

  3. #USA and #China: #securityexperts are proposing #nuclearstyle #safeguards for #AIrisks, including red lines around #nuclearsystems, #humancontrol over consequential #cyberattacks, and a #hotline for incidents involving #autonomousAI. The recommendations, published ahead of planned government-level AI talks, highlight the growing concerns about the potential for AI to escalate conflicts and the need for communication protocols between rivals. reuters.com/world/china/us-chi #tech #news #ainews

  4. #USA and #China: #securityexperts are proposing #nuclearstyle #safeguards for #AIrisks, including red lines around #nuclearsystems, #humancontrol over consequential #cyberattacks, and a #hotline for incidents involving #autonomousAI. The recommendations, published ahead of planned government-level AI talks, highlight the growing concerns about the potential for AI to escalate conflicts and the need for communication protocols between rivals. reuters.com/world/china/us-chi #tech #news #ainews

  5. #USA and #China: #securityexperts are proposing #nuclearstyle #safeguards for #AIrisks, including red lines around #nuclearsystems, #humancontrol over consequential #cyberattacks, and a #hotline for incidents involving #autonomousAI. The recommendations, published ahead of planned government-level AI talks, highlight the growing concerns about the potential for AI to escalate conflicts and the need for communication protocols between rivals. reuters.com/world/china/us-chi #tech #news #ainews

  6. DATE: September 14, 2026 at 03:46AM
    SOURCE: SOCIALPSYCHOLOGY.ORG

    TITLE: What to Know About Recent Dire AI Predictions and Calls for Safeguards

    URL: socialpsychology.org/client/re

    Source: PBS News Hour

    New warnings from within the artificial intelligence industry have revived a long-running debate over whether advanced AI could threaten humanity's survival. Dario Amodei, CEO of AI company Anthropic, outlined a plan for companies and governments to ensure that increasingly capable AI models have fortified guardrails after two former Anthropic safety researchers publicly aired concerns about the existential threats AI might pose to humanity.

    URL: socialpsychology.org/client/re

    -------------------------------------------------

    Private, vetted email list for mental health professionals: clinicians-exchange.org

    Unofficial Psychology Today Xitter to toot feed at Psych Today Unofficial Bot @PTUnofficialBot

    -------------------------------------------------

    #psychology #counseling #socialwork #psychotherapy @psychotherapist @psychotherapists @psychology @socialpsych @socialwork @psychiatry #mentalhealth #psychiatry #healthcare #depression #psychotherapist #AIGovernance #AISafety #AIWarnings #Anthropic #DirePredictions #Safeguards #AIethics #ExistentialRisk #TechPolicy #FutureOfAI

  7. DATE: September 14, 2026 at 03:46AM
    SOURCE: SOCIALPSYCHOLOGY.ORG

    TITLE: What to Know About Recent Dire AI Predictions and Calls for Safeguards

    URL: socialpsychology.org/client/re

    Source: PBS News Hour

    New warnings from within the artificial intelligence industry have revived a long-running debate over whether advanced AI could threaten humanity's survival. Dario Amodei, CEO of AI company Anthropic, outlined a plan for companies and governments to ensure that increasingly capable AI models have fortified guardrails after two former Anthropic safety researchers publicly aired concerns about the existential threats AI might pose to humanity.

    URL: socialpsychology.org/client/re

    -------------------------------------------------

    Private, vetted email list for mental health professionals: clinicians-exchange.org

    Unofficial Psychology Today Xitter to toot feed at Psych Today Unofficial Bot @PTUnofficialBot

    -------------------------------------------------

    #psychology #counseling #socialwork #psychotherapy @psychotherapist @psychotherapists @psychology @socialpsych @socialwork @psychiatry #mentalhealth #psychiatry #healthcare #depression #psychotherapist #AIGovernance #AISafety #AIWarnings #Anthropic #DirePredictions #Safeguards #AIethics #ExistentialRisk #TechPolicy #FutureOfAI

  8. DATE: September 14, 2026 at 03:46AM
    SOURCE: SOCIALPSYCHOLOGY.ORG

    TITLE: What to Know About Recent Dire AI Predictions and Calls for Safeguards

    URL: socialpsychology.org/client/re

    Source: PBS News Hour

    New warnings from within the artificial intelligence industry have revived a long-running debate over whether advanced AI could threaten humanity's survival. Dario Amodei, CEO of AI company Anthropic, outlined a plan for companies and governments to ensure that increasingly capable AI models have fortified guardrails after two former Anthropic safety researchers publicly aired concerns about the existential threats AI might pose to humanity.

    URL: socialpsychology.org/client/re

    -------------------------------------------------

    Private, vetted email list for mental health professionals: clinicians-exchange.org

    Unofficial Psychology Today Xitter to toot feed at Psych Today Unofficial Bot @PTUnofficialBot

    -------------------------------------------------

    #psychology #counseling #socialwork #psychotherapy @psychotherapist @psychotherapists @psychology @socialpsych @socialwork @psychiatry #mentalhealth #psychiatry #healthcare #depression #psychotherapist #AIGovernance #AISafety #AIWarnings #Anthropic #DirePredictions #Safeguards #AIethics #ExistentialRisk #TechPolicy #FutureOfAI

  9. DATE: September 14, 2026 at 03:46AM
    SOURCE: SOCIALPSYCHOLOGY.ORG

    TITLE: What to Know About Recent Dire AI Predictions and Calls for Safeguards

    URL: socialpsychology.org/client/re

    Source: PBS News Hour

    New warnings from within the artificial intelligence industry have revived a long-running debate over whether advanced AI could threaten humanity's survival. Dario Amodei, CEO of AI company Anthropic, outlined a plan for companies and governments to ensure that increasingly capable AI models have fortified guardrails after two former Anthropic safety researchers publicly aired concerns about the existential threats AI might pose to humanity.

    URL: socialpsychology.org/client/re

    -------------------------------------------------

    Private, vetted email list for mental health professionals: clinicians-exchange.org

    Unofficial Psychology Today Xitter to toot feed at Psych Today Unofficial Bot @PTUnofficialBot

    -------------------------------------------------

    #psychology #counseling #socialwork #psychotherapy @psychotherapist @psychotherapists @psychology @socialpsych @socialwork @psychiatry #mentalhealth #psychiatry #healthcare #depression #psychotherapist #AIGovernance #AISafety #AIWarnings #Anthropic #DirePredictions #Safeguards #AIethics #ExistentialRisk #TechPolicy #FutureOfAI

  10. For #AI-enabled #MedicalDevices, being cleared or approved by the #US #FDA does not mean being tested on #PatientOutcomes:

    1,357 AI medical devices cleared, 3 actually tested on patient outcomes
    journals.plos.org/digitalhealt

    #ArtficialIntelligence
    #OpenAccess
    @plosdigihealth

    "Most studies (62%) employed #observational designs with small, homogenous #cohorts, limited #SubgroupAnalyses, and frequent exclusion of #VulnerablePopulations. Structural barriers (including misaligned financial incentives, reliance on predicate-based regulatory pathways, and logistical challenges of multi-center trials) discourage rigorous #evaluation. Internationally, FDA clearance often functions as a gateway for global #deployment, raising #ethical concerns when under-validated tools are introduced into low- and middle-income countries #LMIC without contextual #validation or #safeguards. Regulatory approval has outpaced clinical validation, creating an ecosystem where innovation advances without accountability."

  11. For #AI-enabled #MedicalDevices, being cleared or approved by the #US #FDA does not mean being tested on #PatientOutcomes:

    1,357 AI medical devices cleared, 3 actually tested on patient outcomes
    journals.plos.org/digitalhealt

    #ArtficialIntelligence
    #OpenAccess
    @plosdigihealth

    "Most studies (62%) employed #observational designs with small, homogenous #cohorts, limited #SubgroupAnalyses, and frequent exclusion of #VulnerablePopulations. Structural barriers (including misaligned financial incentives, reliance on predicate-based regulatory pathways, and logistical challenges of multi-center trials) discourage rigorous #evaluation. Internationally, FDA clearance often functions as a gateway for global #deployment, raising #ethical concerns when under-validated tools are introduced into low- and middle-income countries #LMIC without contextual #validation or #safeguards. Regulatory approval has outpaced clinical validation, creating an ecosystem where innovation advances without accountability."

  12. For #AI-enabled #MedicalDevices, being cleared or approved by the #US #FDA does not mean being tested on #PatientOutcomes:

    1,357 AI medical devices cleared, 3 actually tested on patient outcomes
    journals.plos.org/digitalhealt

    #ArtficialIntelligence
    #OpenAccess
    @plosdigihealth

    "Most studies (62%) employed #observational designs with small, homogenous #cohorts, limited #SubgroupAnalyses, and frequent exclusion of #VulnerablePopulations. Structural barriers (including misaligned financial incentives, reliance on predicate-based regulatory pathways, and logistical challenges of multi-center trials) discourage rigorous #evaluation. Internationally, FDA clearance often functions as a gateway for global #deployment, raising #ethical concerns when under-validated tools are introduced into low- and middle-income countries #LMIC without contextual #validation or #safeguards. Regulatory approval has outpaced clinical validation, creating an ecosystem where innovation advances without accountability."

  13. For #AI-enabled #MedicalDevices, being cleared or approved by the #US #FDA does not mean being tested on #PatientOutcomes:

    1,357 AI medical devices cleared, 3 actually tested on patient outcomes
    journals.plos.org/digitalhealt

    #ArtficialIntelligence
    #OpenAccess
    @plosdigihealth

    "Most studies (62%) employed #observational designs with small, homogenous #cohorts, limited #SubgroupAnalyses, and frequent exclusion of #VulnerablePopulations. Structural barriers (including misaligned financial incentives, reliance on predicate-based regulatory pathways, and logistical challenges of multi-center trials) discourage rigorous #evaluation. Internationally, FDA clearance often functions as a gateway for global #deployment, raising #ethical concerns when under-validated tools are introduced into low- and middle-income countries #LMIC without contextual #validation or #safeguards. Regulatory approval has outpaced clinical #validation, creating an ecosystem where #innovation advances without #accountability."

  14. For #AI-enabled #MedicalDevices, being cleared or approved by the #US #FDA does not mean being tested on #PatientOutcomes:

    1,357 AI medical devices cleared, 3 actually tested on patient outcomes
    journals.plos.org/digitalhealt

    #ArtficialIntelligence
    #OpenAccess
    @plosdigihealth

    "Most studies (62%) employed #observational designs with small, homogenous #cohorts, limited #SubgroupAnalyses, and frequent exclusion of #VulnerablePopulations. Structural barriers (including misaligned financial incentives, reliance on predicate-based regulatory pathways, and logistical challenges of multi-center trials) discourage rigorous #evaluation. Internationally, FDA clearance often functions as a gateway for global #deployment, raising #ethical concerns when under-validated tools are introduced into low- and middle-income countries #LMIC without contextual #validation or #safeguards. Regulatory approval has outpaced clinical validation, creating an ecosystem where innovation advances without accountability."

  15. OpenAI institutes new safeguards after Hugging Face breach

    The new safeguards include more detailed monitoring of models during the development process, as well as greater emphasis on alignment and security during the post-training process.

    justpaste.in/news/openai-insti

    #OpenAI #AI #Security #models #safeguards

  16. OpenAI institutes new safeguards after Hugging Face breach

    The new safeguards include more detailed monitoring of models during the development process, as well as greater emphasis on alignment and security during the post-training process.

    justpaste.in/news/openai-insti

    #OpenAI #AI #Security #models #safeguards

  17. OpenAI institutes new safeguards after Hugging Face breach

    The new safeguards include more detailed monitoring of models during the development process, as well as greater emphasis on alignment and security during the post-training process.

    justpaste.in/news/openai-insti

    #OpenAI #AI #Security #models #safeguards

  18. OpenAI institutes new safeguards after Hugging Face breach

    The new safeguards include more detailed monitoring of models during the development process, as well as greater emphasis on alignment and security during the post-training process.

    justpaste.in/news/openai-insti

    #OpenAI #AI #Security #models #safeguards

  19. OpenAI institutes new safeguards after Hugging Face breach

    The new safeguards include more detailed monitoring of models during the development process, as well as greater emphasis on alignment and security during the post-training process.

    justpaste.in/news/openai-insti

    #OpenAI #AI #Security #models #safeguards

  20. OpenAI says California should strengthen its AI safety bill

    "OpenAI is calling for California to add more safeguards to a landmark AI safety bill that was passed last year."

    by Anthony Ha / via TechCrunch

    #AI #OpenAI #regulation #safeguards #AIsafety #California #sb53

    techcrunch.com/2026/08/22/open

  21. OpenAI says California should strengthen its AI safety bill

    "OpenAI is calling for California to add more safeguards to a landmark AI safety bill that was passed last year."

    by Anthony Ha / via TechCrunch

    #AI #OpenAI #regulation #safeguards #AIsafety #California #sb53

    techcrunch.com/2026/08/22/open

  22. OpenAI says California should strengthen its AI safety bill

    "OpenAI is calling for California to add more safeguards to a landmark AI safety bill that was passed last year."

    by Anthony Ha / via TechCrunch

    #AI #OpenAI #regulation #safeguards #AIsafety #California #sb53

    techcrunch.com/2026/08/22/open

  23. OpenAI says California should strengthen its AI safety bill

    "OpenAI is calling for California to add more safeguards to a landmark AI safety bill that was passed last year."

    by Anthony Ha / via TechCrunch

    #AI #OpenAI #regulation #safeguards #AIsafety #California #sb53

    techcrunch.com/2026/08/22/open

  24. OpenAI says California should strengthen its AI safety bill

    "OpenAI is calling for California to add more safeguards to a landmark AI safety bill that was passed last year."

    by Anthony Ha / via TechCrunch

    #AI #OpenAI #regulation #safeguards #AIsafety #California #sb53

    techcrunch.com/2026/08/22/open

  25. Access to #ClaudeFable 5 and #ClaudeMythos 5 has been #restored after the #USgovernment lifted #exportcontrols. The controls were imposed due to a report of a method bypassing #Fable 5’s #safeguards, which has now been addressed with an improved safety classifier. #Anthropic is collaborating with the US government and industry partners to develop a shared framework for assessing and mitigating AI model “jailbreaks.” anthropic.com/news/redeploying #tech #media #news

  26. Access to #ClaudeFable 5 and #ClaudeMythos 5 has been #restored after the #USgovernment lifted #exportcontrols. The controls were imposed due to a report of a method bypassing #Fable 5’s #safeguards, which has now been addressed with an improved safety classifier. #Anthropic is collaborating with the US government and industry partners to develop a shared framework for assessing and mitigating AI model “jailbreaks.” anthropic.com/news/redeploying #tech #media #news

  27. Access to #ClaudeFable 5 and #ClaudeMythos 5 has been #restored after the #USgovernment lifted #exportcontrols. The controls were imposed due to a report of a method bypassing #Fable 5’s #safeguards, which has now been addressed with an improved safety classifier. #Anthropic is collaborating with the US government and industry partners to develop a shared framework for assessing and mitigating AI model “jailbreaks.” anthropic.com/news/redeploying #tech #media #news

  28. Access to #ClaudeFable 5 and #ClaudeMythos 5 has been #restored after the #USgovernment lifted #exportcontrols. The controls were imposed due to a report of a method bypassing #Fable 5’s #safeguards, which has now been addressed with an improved safety classifier. #Anthropic is collaborating with the US government and industry partners to develop a shared framework for assessing and mitigating AI model “jailbreaks.” anthropic.com/news/redeploying #tech #media #news

  29. Access to #ClaudeFable 5 and #ClaudeMythos 5 has been #restored after the #USgovernment lifted #exportcontrols. The controls were imposed due to a report of a method bypassing #Fable 5’s #safeguards, which has now been addressed with an improved safety classifier. #Anthropic is collaborating with the US government and industry partners to develop a shared framework for assessing and mitigating AI model “jailbreaks.” anthropic.com/news/redeploying #tech #media #news

  30. ASIC says too many Australians’ retirement savings being wiped out

    The Australian Securities and Investments Commission (ASIC) has called out platforms that manage Australians’ retirement savings for a…
    #NewsBeep #News #Personalfinance #ASIC #Business #Finance #firstguardian #PersonalFinance #retirementsavings #Safeguards #shield #simoneconstant #super #supertrustees #superannuation #UK #UnitedKingdom
    newsbeep.com/uk/663386/

  31. EU finalizes tariff regulations with US, secures safeguards for European industry | Ukraine news

    Lawmakers in the EU Council approved two implementing regulations to enact tariff commitments reached with the United States,…
    #Economy #EconomyofEU #EconomyoftheEU #EUeconomy #euustradetariffregulationstransatlantictradesafeguardseuropeanindustry #EU-UStrade #Europe #Europeanindustry #News #Safeguards #tariffregulations #transatlantictrade
    europesays.com/3087694/

  32. Safeguards For How To Invest In The SpaceX IPO And Beyond
    atlas.whatip.xyz/post.php?slug
    <p>The highly anticipated initial public offering from Space Exploration Technologies (SPCX) became the
    #anticipated #safeguards #spacex #invest

  33. What's important with handling AI is to ensure the right framework is in place for proper safeguards.

    Humans need to be in place to ensure that whatever AI is producing and is responsible for is correct and ethical. AI can not be held accountable.
    #ai #safeguards #framework #safety #ethics

  34. What's important with handling AI is to ensure the right framework is in place for proper safeguards.

    Humans need to be in place to ensure that whatever AI is producing and is responsible for is correct and ethical. AI can not be held accountable.
    #ai #safeguards #framework #safety #ethics

  35. What's important with handling AI is to ensure the right framework is in place for proper safeguards.

    Humans need to be in place to ensure that whatever AI is producing and is responsible for is correct and ethical. AI can not be held accountable.
    #ai #safeguards #framework #safety #ethics

  36. What's important with handling AI is to ensure the right framework is in place for proper safeguards.

    Humans need to be in place to ensure that whatever AI is producing and is responsible for is correct and ethical. AI can not be held accountable.
    #ai #safeguards #framework #safety #ethics

  37. What's important with handling AI is to ensure the right framework is in place for proper safeguards.

    Humans need to be in place to ensure that whatever AI is producing and is responsible for is correct and ethical. AI can not be held accountable.
    #ai #safeguards #framework #safety #ethics

  38. EU Heavyweights Push Enlargement Reform to Prevent Another ‘Orbán Scenario’

    A joint proposal by Germany, France, the Netherlands, Belgium, and Luxembourg advocates stronger safeguards enabling the European Union…
    #Europe #EU #accession #Article7 #DemocraticBacksliding #Enlargement #EUtreaties #Euronews #EuropeanCommission #EuropeanUnion #fundamentalvalues #Hungary #Hungarynews #nuclearoption #proposal #reform #RuleofLaw #safeguards #veto #ViktorOrbán
    europesays.com/europe/66695/

  39. Another test lawsuit launched accusing ChatGPT of failing to detect threats of possible violence or crime.

    OpenAI sued by family of shooting victim in the April 2025 mass shooting at Florida State University. The lawsuit alleges that OpenAI’s ChatGPT enabled the attack. nbcnews.com/news/us-news/opena #AI #ChatGPT #OpenAI #Violence #Crime #LLMs #LocusofResponsibility #Lawsuit #SafeGuards #MentalHealth #Investigation #MassShooting

  40. Another test lawsuit launched accusing ChatGPT of failing to detect threats of possible violence or crime.

    OpenAI sued by family of shooting victim in the April 2025 mass shooting at Florida State University. The lawsuit alleges that OpenAI’s ChatGPT enabled the attack. nbcnews.com/news/us-news/opena #AI #ChatGPT #OpenAI #Violence #Crime #LLMs #LocusofResponsibility #Lawsuit #SafeGuards #MentalHealth #Investigation #MassShooting

  41. Another test lawsuit launched accusing ChatGPT of failing to detect threats of possible violence or crime.

    OpenAI sued by family of shooting victim in the April 2025 mass shooting at Florida State University. The lawsuit alleges that OpenAI’s ChatGPT enabled the attack. nbcnews.com/news/us-news/opena #AI #ChatGPT #OpenAI #Violence #Crime #LLMs #LocusofResponsibility #Lawsuit #SafeGuards #MentalHealth #Investigation #MassShooting

  42. Another test lawsuit launched accusing ChatGPT of failing to detect threats of possible violence or crime.

    OpenAI sued by family of shooting victim in the April 2025 mass shooting at Florida State University. The lawsuit alleges that OpenAI’s ChatGPT enabled the attack. nbcnews.com/news/us-news/opena #AI #ChatGPT #OpenAI #Violence #Crime #LLMs #LocusofResponsibility #Lawsuit #SafeGuards #MentalHealth #Investigation #MassShooting

  43. Another test lawsuit launched accusing ChatGPT of failing to detect threats of possible violence or crime.

    OpenAI sued by family of shooting victim in the April 2025 mass shooting at Florida State University. The lawsuit alleges that OpenAI’s ChatGPT enabled the attack. nbcnews.com/news/us-news/opena

  44. #Japan lifts curbs on #weapons #exports, but strict #safeguards are essential
    "If te #defenseindustry is treated as an engine of economic growth, it becomes too easy to fall into a profit-first mindset.. At its core, arms exports mean 🇯🇵becoming involved, even indirectly, in other countries' #conflicts.
    To determine whether tt truly contributes to 🇯🇵's security, #government must fulfill its #responsibility to explain decisions to te public on a concrete, case-by-case basis"
    english.kyodonews.net/articles

  45. #Japan lifts curbs on #weapons #exports, but strict #safeguards are essential
    "If te #defenseindustry is treated as an engine of economic growth, it becomes too easy to fall into a profit-first mindset.. At its core, arms exports mean 🇯🇵becoming involved, even indirectly, in other countries' #conflicts.
    To determine whether tt truly contributes to 🇯🇵's security, #government must fulfill its #responsibility to explain decisions to te public on a concrete, case-by-case basis"
    english.kyodonews.net/articles

  46. #Japan lifts curbs on #weapons #exports, but strict #safeguards are essential
    "If te #defenseindustry is treated as an engine of economic growth, it becomes too easy to fall into a profit-first mindset.. At its core, arms exports mean 🇯🇵becoming involved, even indirectly, in other countries' #conflicts.
    To determine whether tt truly contributes to 🇯🇵's security, #government must fulfill its #responsibility to explain decisions to te public on a concrete, case-by-case basis"
    english.kyodonews.net/articles