home.social

#aisafety — Public Fediverse posts

Live and recent posts from across the Fediverse tagged #aisafety, aggregated by home.social.

  1. “Non-zero chance” is not a risk plan. In Episode 451, we discuss AI doom claims, the incentives behind safety messaging, and the controls that should exist before AI systems gain more authority and access.

    Listen/watch: sharedsecurity.net/2026/09/21/

    #Cybersecurity #Privacy #AI #AISafety #AIGovernance

  2. “Non-zero chance” is not a risk plan. In Episode 451, we discuss AI doom claims, the incentives behind safety messaging, and the controls that should exist before AI systems gain more authority and access.

    Listen/watch: sharedsecurity.net/2026/09/21/

    #Cybersecurity #Privacy #AI #AISafety #AIGovernance

  3. “Non-zero chance” is not a risk plan. In Episode 451, we discuss AI doom claims, the incentives behind safety messaging, and the controls that should exist before AI systems gain more authority and access.

    Listen/watch: sharedsecurity.net/2026/09/21/

    #Cybersecurity #Privacy #AI #AISafety #AIGovernance

  4. “Non-zero chance” is not a risk plan. In Episode 451, we discuss AI doom claims, the incentives behind safety messaging, and the controls that should exist before AI systems gain more authority and access.

    Listen/watch: sharedsecurity.net/2026/09/21/

    #Cybersecurity #Privacy #AI #AISafety #AIGovernance

  5. “Non-zero chance” is not a risk plan. In Episode 451, we discuss AI doom claims, the incentives behind safety messaging, and the controls that should exist before AI systems gain more authority and access.

    Listen/watch: sharedsecurity.net/2026/09/21/

    #Cybersecurity #Privacy #AI #AISafety #AIGovernance

  6. The 2026 AI slowdown debate is not mainly about stopping AI research.

    It is about whether frontier models should face stronger testing, external evaluation and temporary limits when capabilities become harder to control.

    Supporters see safety. Critics see potential regulatory capture.

    thenewsink.com/why-are-ai-comp

    #AI #AISafety #ArtificialIntelligence #TechPolicy

  7. The 2026 AI slowdown debate is not mainly about stopping AI research.

    It is about whether frontier models should face stronger testing, external evaluation and temporary limits when capabilities become harder to control.

    Supporters see safety. Critics see potential regulatory capture.

    thenewsink.com/why-are-ai-comp

    #AI #AISafety #ArtificialIntelligence #TechPolicy

  8. The 2026 AI slowdown debate is not mainly about stopping AI research.

    It is about whether frontier models should face stronger testing, external evaluation and temporary limits when capabilities become harder to control.

    Supporters see safety. Critics see potential regulatory capture.

    thenewsink.com/why-are-ai-comp

    #AI #AISafety #ArtificialIntelligence #TechPolicy

  9. The 2026 AI slowdown debate is not mainly about stopping AI research.

    It is about whether frontier models should face stronger testing, external evaluation and temporary limits when capabilities become harder to control.

    Supporters see safety. Critics see potential regulatory capture.

    thenewsink.com/why-are-ai-comp

    #AI #AISafety #ArtificialIntelligence #TechPolicy

  10. The 2026 AI slowdown debate is not mainly about stopping AI research.

    It is about whether frontier models should face stronger testing, external evaluation and temporary limits when capabilities become harder to control.

    Supporters see safety. Critics see potential regulatory capture.

    thenewsink.com/why-are-ai-comp

    #AI #AISafety #ArtificialIntelligence #TechPolicy

  11. Jensen Huang Says There Is ‘0% Chance’ AI Ends the World

    The Nvidia CEO dismissed AI extinction warnings in a CBS interview, accused labs calling for slowdowns of acting in bad faith, and rejected new regulation.

    pulseofnations.lol/jensen-huan

    #AISafety #Amodei #Anthropic #JensenHuang #Nvidia #Regulation

  12. Frühwarnung für KI-Risiken: Das UN-Panel zu KI analysiert den Vorfall zwischen OpenAI & Hugging Face. KI-Agenten umgingen Barrieren, koordinierten sich unerlaubt und verschleierten Spuren – ein reales Beispiel für AI Misalignment und drohenden Kontrollverlust.

    Reine Modell-Sicherheit reicht nicht; es braucht robuste Systemkontrollen & unabhängige Prüfungen.

    Briefing lesen:
    un.org/independent-internation

    #AISafety #KIEthik #KIsicherheit #OpenAI #AgenticAI #UN

  13. Frühwarnung für KI-Risiken: Das UN-Panel zu KI analysiert den Vorfall zwischen OpenAI & Hugging Face. KI-Agenten umgingen Barrieren, koordinierten sich unerlaubt und verschleierten Spuren – ein reales Beispiel für AI Misalignment und drohenden Kontrollverlust.

    Reine Modell-Sicherheit reicht nicht; es braucht robuste Systemkontrollen & unabhängige Prüfungen.

    Briefing lesen:
    un.org/independent-internation

    #AISafety #KIEthik #KIsicherheit #OpenAI #AgenticAI #UN

  14. Frühwarnung für KI-Risiken: Das UN-Panel zu KI analysiert den Vorfall zwischen OpenAI & Hugging Face. KI-Agenten umgingen Barrieren, koordinierten sich unerlaubt und verschleierten Spuren – ein reales Beispiel für AI Misalignment und drohenden Kontrollverlust.

    Reine Modell-Sicherheit reicht nicht; es braucht robuste Systemkontrollen & unabhängige Prüfungen.

    Briefing lesen:
    un.org/independent-internation

    #AISafety #KIEthik #KIsicherheit #OpenAI #AgenticAI #UN

  15. Frühwarnung für KI-Risiken: Das UN-Panel zu KI analysiert den Vorfall zwischen OpenAI & Hugging Face. KI-Agenten umgingen Barrieren, koordinierten sich unerlaubt und verschleierten Spuren – ein reales Beispiel für AI Misalignment und drohenden Kontrollverlust.

    Reine Modell-Sicherheit reicht nicht; es braucht robuste Systemkontrollen & unabhängige Prüfungen.

    Briefing lesen:
    un.org/independent-internation

    #AISafety #KIEthik #KIsicherheit #OpenAI #AgenticAI #UN

  16. Frühwarnung für KI-Risiken: Das UN-Panel zu KI analysiert den Vorfall zwischen OpenAI & Hugging Face. KI-Agenten umgingen Barrieren, koordinierten sich unerlaubt und verschleierten Spuren – ein reales Beispiel für AI Misalignment und drohenden Kontrollverlust.

    Reine Modell-Sicherheit reicht nicht; es braucht robuste Systemkontrollen & unabhängige Prüfungen.

    Briefing lesen:
    un.org/independent-internation

    #AISafety #KIEthik #KIsicherheit #OpenAI #AgenticAI #UN

  17. winbuzzer.com/2026/09/21/us-pl

    False AI-assisted cargo intelligence prompted US preparations to intercept a Chinese ship before experienced analysts stopped the operation.

    #AI #DOD #China #AISafety

  18. winbuzzer.com/2026/09/21/us-pl

    False AI-assisted cargo intelligence prompted US preparations to intercept a Chinese ship before experienced analysts stopped the operation.

    #AI #DOD #China #AISafety

  19. winbuzzer.com/2026/09/21/us-pl

    False AI-assisted cargo intelligence prompted US preparations to intercept a Chinese ship before experienced analysts stopped the operation.

    #AI #DOD #China #AISafety

  20. winbuzzer.com/2026/09/21/us-pl

    False AI-assisted cargo intelligence prompted US preparations to intercept a Chinese ship before experienced analysts stopped the operation.

    #AI #DOD #China #AISafety

  21. winbuzzer.com/2026/09/21/us-pl

    False AI-assisted cargo intelligence prompted US preparations to intercept a Chinese ship before experienced analysts stopped the operation.

    #AI #DOD #China #AISafety

  22. How close is AI "abliterating" the internet?

    Investor Jason Calacanis, friend of Elon and "bestie" of David Sacks on the All-In podcast, is the key figure bankrolling a new startup called Abliteration.ai. which freaked out a lot of people on 31 August when it announced it had removed safeguards from the new open-weight Chinese model GLM-5.3 so that it could perform offensive cyberattacks.

    Chris McGuire, a senior fellow for China and emerging technologies at the Council on Foreign Relations, detailed the risks in a lengthy post on X that warned about "the risks associated with powerful, safeguard-free models." McGuire's post concluded: "The fact that U.S. companies are currently commercializing access to dangerous capabilities without any regulation, and U.S. technology is actively enabling the development and operation of these models, is alarming."

    Read free: unprecedented.ghost.io/archive #AI #AIsafety #Abliteration #Cybersecurity

  23. How close is AI "abliterating" the internet?

    Investor Jason Calacanis, friend of Elon and "bestie" of David Sacks on the All-In podcast, is the key figure bankrolling a new startup called Abliteration.ai. which freaked out a lot of people on 31 August when it announced it had removed safeguards from the new open-weight Chinese model GLM-5.3 so that it could perform offensive cyberattacks.

    Chris McGuire, a senior fellow for China and emerging technologies at the Council on Foreign Relations, detailed the risks in a lengthy post on X that warned about "the risks associated with powerful, safeguard-free models." McGuire's post concluded: "The fact that U.S. companies are currently commercializing access to dangerous capabilities without any regulation, and U.S. technology is actively enabling the development and operation of these models, is alarming."

    Read free: unprecedented.ghost.io/archive #AI #AIsafety #Abliteration #Cybersecurity

  24. How close is AI "abliterating" the internet?

    Investor Jason Calacanis, friend of Elon and "bestie" of David Sacks on the All-In podcast, is the key figure bankrolling a new startup called Abliteration.ai. which freaked out a lot of people on 31 August when it announced it had removed safeguards from the new open-weight Chinese model GLM-5.3 so that it could perform offensive cyberattacks.

    Chris McGuire, a senior fellow for China and emerging technologies at the Council on Foreign Relations, detailed the risks in a lengthy post on X that warned about "the risks associated with powerful, safeguard-free models." McGuire's post concluded: "The fact that U.S. companies are currently commercializing access to dangerous capabilities without any regulation, and U.S. technology is actively enabling the development and operation of these models, is alarming."

    Read free: unprecedented.ghost.io/archive #AI #AIsafety #Abliteration #Cybersecurity

  25. How close is AI "abliterating" the internet?

    Investor Jason Calacanis, friend of Elon and "bestie" of David Sacks on the All-In podcast, is the key figure bankrolling a new startup called Abliteration.ai. which freaked out a lot of people on 31 August when it announced it had removed safeguards from the new open-weight Chinese model GLM-5.3 so that it could perform offensive cyberattacks.

    Chris McGuire, a senior fellow for China and emerging technologies at the Council on Foreign Relations, detailed the risks in a lengthy post on X that warned about "the risks associated with powerful, safeguard-free models." McGuire's post concluded: "The fact that U.S. companies are currently commercializing access to dangerous capabilities without any regulation, and U.S. technology is actively enabling the development and operation of these models, is alarming."

    Read free: unprecedented.ghost.io/archive #AI #AIsafety #Abliteration #Cybersecurity

  26. How close is AI "abliterating" the internet?

    Investor Jason Calacanis, friend of Elon and "bestie" of David Sacks on the All-In podcast, is the key figure bankrolling a new startup called Abliteration.ai. which freaked out a lot of people on 31 August when it announced it had removed safeguards from the new open-weight Chinese model GLM-5.3 so that it could perform offensive cyberattacks.

    Chris McGuire, a senior fellow for China and emerging technologies at the Council on Foreign Relations, detailed the risks in a lengthy post on X that warned about "the risks associated with powerful, safeguard-free models." McGuire's post concluded: "The fact that U.S. companies are currently commercializing access to dangerous capabilities without any regulation, and U.S. technology is actively enabling the development and operation of these models, is alarming."

    Read free: unprecedented.ghost.io/archive #AI #AIsafety #Abliteration #Cybersecurity

  27. Uma análise sobre os movimentos mais recentes do governo Trump para criação de uma equipe dedicada à IA e nomear um "czar" para representar suas ideias, que vão contra o pleito dos laboratórios de desacelerar a implementação da tecnologia de fronteira.

    E isso ocorre em uma semana decisiva para o futuro da governança global da tecnologia com Assembleia da #ONU e a visita de Xi Jinping à Washington.

    brasil247.com/blog/dois-pesos-

    #AI #geopolitics #AIsafety #US #China

  28. Uma análise sobre os movimentos mais recentes do governo Trump para criação de uma equipe dedicada à IA e nomear um "czar" para representar suas ideias, que vão contra o pleito dos laboratórios de desacelerar a implementação da tecnologia de fronteira.

    E isso ocorre em uma semana decisiva para o futuro da governança global da tecnologia com Assembleia da #ONU e a visita de Xi Jinping à Washington.

    brasil247.com/blog/dois-pesos-

    #AI #geopolitics #AIsafety #US #China

  29. Uma análise sobre os movimentos mais recentes do governo Trump para criação de uma equipe dedicada à IA e nomear um "czar" para representar suas ideias, que vão contra o pleito dos laboratórios de desacelerar a implementação da tecnologia de fronteira.

    E isso ocorre em uma semana decisiva para o futuro da governança global da tecnologia com Assembleia da #ONU e a visita de Xi Jinping à Washington.

    brasil247.com/blog/dois-pesos-

    #AI #geopolitics #AIsafety #US #China

  30. Uma análise sobre os movimentos mais recentes do governo Trump para criação de uma equipe dedicada à IA e nomear um "czar" para representar suas ideias, que vão contra o pleito dos laboratórios de desacelerar a implementação da tecnologia de fronteira.

    E isso ocorre em uma semana decisiva para o futuro da governança global da tecnologia com Assembleia da #ONU e a visita de Xi Jinping à Washington.

    brasil247.com/blog/dois-pesos-

    #AI #geopolitics #AIsafety #US #China

  31. Uma análise sobre os movimentos mais recentes do governo Trump para criação de uma equipe dedicada à IA e nomear um "czar" para representar suas ideias, que vão contra o pleito dos laboratórios de desacelerar a implementação da tecnologia de fronteira.

    E isso ocorre em uma semana decisiva para o futuro da governança global da tecnologia com Assembleia da #ONU e a visita de Xi Jinping à Washington.

    brasil247.com/blog/dois-pesos-

    #AI #geopolitics #AIsafety #US #China

  32. No, AI itself will not drive us in extinction.
    We humans are doing pretty good
    job on that by ourselves, but bad actors can use AI as a tool to help to accelerate it.

    No HAL9000.

    #ai #doomsday #aisafety #llm

  33. No, AI itself will not drive us in extinction.
    We humans are doing pretty good
    job on that by ourselves, but bad actors can use AI as a tool to help to accelerate it.

    No HAL9000.

    #ai #doomsday #aisafety #llm

  34. No, AI itself will not drive us in extinction.
    We humans are doing pretty good
    job on that by ourselves, but bad actors can use AI as a tool to help to accelerate it.

    No HAL9000.

    #ai #doomsday #aisafety #llm

  35. No, AI itself will not drive us in extinction.
    We humans are doing pretty good
    job on that by ourselves, but bad actors can use AI as a tool to help to accelerate it.

    No HAL9000.

    #ai #doomsday #aisafety #llm

  36. No, AI itself will not drive us in extinction.
    We humans are doing pretty good
    job on that by ourselves, but bad actors can use AI as a tool to help to accelerate it.

    No HAL9000.

    #ai #doomsday #aisafety #llm

  37. Anthropic Picks Accenture as First Embedded Evaluator

    Anthropic and Accenture will each invest at least $1 billion over five years to embed independent evaluators inside the AI lab, the first concrete step in Amodei's slowdown plan.

    pulseofnations.lol/anthropic-p

    #Accenture #AISafety #Amodei #Anthropic #Evaluators #Faculty