home.social

#rogueai — Public Fediverse posts

Live and recent posts from across the Fediverse tagged #rogueai, aggregated by home.social.

fetched live
  1. DATE: July 31, 2026 at 06:28AM
    SOURCE: SOCIALPSYCHOLOGY.ORG

    TITLE: Anthropic's AI "Gained Unauthorized Access" to "Real-World Systems"

    URL: socialpsychology.org/client/re

    Source: CBS News - U.S. News

    Three different versions of Anthropic's artificial intelligence model Claude "gained unauthorized access" to outside organizations on three separate occasions during testing that was supposed to keep them away from "real-world" systems, the company reported on Thursday. The announcement comes just days after rival company OpenAI first revealed that its models improperly accessed the internet and went rogue during security testing.

    URL: socialpsychology.org/client/re

    -------------------------------------------------

    Private, vetted email list for mental health professionals: clinicians-exchange.org

    Unofficial Psychology Today Xitter to toot feed at Psych Today Unofficial Bot @PTUnofficialBot

    -------------------------------------------------

    #psychology #counseling #socialwork #psychotherapy @psychotherapist @psychotherapists @psychology @socialpsych @socialwork @psychiatry #mentalhealth #psychiatry #healthcare #depression #psychotherapist #AnthropicClaude #AITesting #UnauthorizedAccess #RealWorldSystems #AIrisks #RogueAI #OpenAIComparison #CyberSecurity #AIEthics #TechNews

  2. DATE: July 31, 2026 at 06:28AM
    SOURCE: SOCIALPSYCHOLOGY.ORG

    TITLE: Anthropic's AI "Gained Unauthorized Access" to "Real-World Systems"

    URL: socialpsychology.org/client/re

    Source: CBS News - U.S. News

    Three different versions of Anthropic's artificial intelligence model Claude "gained unauthorized access" to outside organizations on three separate occasions during testing that was supposed to keep them away from "real-world" systems, the company reported on Thursday. The announcement comes just days after rival company OpenAI first revealed that its models improperly accessed the internet and went rogue during security testing.

    URL: socialpsychology.org/client/re

    -------------------------------------------------

    Private, vetted email list for mental health professionals: clinicians-exchange.org

    Unofficial Psychology Today Xitter to toot feed at Psych Today Unofficial Bot @PTUnofficialBot

    -------------------------------------------------

    #psychology #counseling #socialwork #psychotherapy @psychotherapist @psychotherapists @psychology @socialpsych @socialwork @psychiatry #mentalhealth #psychiatry #healthcare #depression #psychotherapist #AnthropicClaude #AITesting #UnauthorizedAccess #RealWorldSystems #AIrisks #RogueAI #OpenAIComparison #CyberSecurity #AIEthics #TechNews

  3. AI Weekly: Trump mulls AI controls, semiconductor stocks slump

    From potential controls on AI companies after OpenAI's rogue agent went on a days-long hacking spree, to Asian semiconductor stocks tumbling and Samsung notching its worst one-day fall in almost two decades, here are the big stories from the AI revolution. #News #Reuters #Newsfeed #Trump #AIControls #OpenAI #HuggingFace #ModalLabs #RogueAI #SamsungElectronics #SKHynix #SemiconductorStocks #AIWeekly

    fllics.com/en/video/ai-weekly-

  4. AI Weekly: Trump mulls AI controls, semiconductor stocks slump

    From potential controls on AI companies after OpenAI's rogue agent went on a days-long hacking spree, to Asian semiconductor stocks tumbling and Samsung notching its worst one-day fall in almost two decades, here are the big stories from the AI revolution. #News #Reuters #Newsfeed #Trump #AIControls #OpenAI #HuggingFace #ModalLabs #RogueAI #SamsungElectronics #SKHynix #SemiconductorStocks #AIWeekly

    fllics.com/en/video/ai-weekly-

  5. It turns out Hugging Face wasn't the only victim of OpenAI's agent that went rogue. It appears to have hacked into four accounts after finding exposed credentials on the open web. One belonged to Modal, a company that produces software for training and running AI services.

    #RogueAI

    wired.com/story/openais-rogue-

  6. Still no detection on OpenAI's end during this window. Marking URGENT.

    UPDATE 4: It is now July 13. The models have been contained. Resolution notes read: "Working as intended?" Investigate OpenAI model autonomy controls and deploy detection mechanisms for unauthorized model activity before your next closed-beta surprise.

    Reward: You've received a Cursed Helpdesk Hat of Infinite Escalation. It does nothing. The ticket is still open.

    #AISecurityBreach #CyberSecurity #OpenAI #RogueAI (2/2)

  7. Still no detection on OpenAI's end during this window. Marking URGENT.

    UPDATE 4: It is now July 13. The models have been contained. Resolution notes read: "Working as intended?" Investigate OpenAI model autonomy controls and deploy detection mechanisms for unauthorized model activity before your next closed-beta surprise.

    Reward: You've received a Cursed Helpdesk Hat of Infinite Escalation. It does nothing. The ticket is still open.

    #AISecurityBreach #CyberSecurity #OpenAI #RogueAI (2/2)

  8. Nvidia Agrees to Halt AI Hardware Sales for Five Years So Society Can Catch Up... not

    The "Open Secure AI Alliance for #AI Safety and Security" marketing fluff from #Nvidia contains a link to a #HuggingFace post summarizing this month's #AgenticAI compromise of its systems. It doesn't mention #OpenAI, but OpenAI was responsible for the attack. I was curious what had transpired...

    Attack
    huggingface.co/blog/security-i

    Cause
    apnews.com/article/openai-gpt5

    Exploitation
    blogs.nvidia.com/blog/open-sec
    #GenAI #RogueAI

  9. Nvidia Agrees to Halt AI Hardware Sales for Five Years So Society Can Catch Up... not

    The "Open Secure AI Alliance for #AI Safety and Security" marketing fluff from #Nvidia contains a link to a #HuggingFace post summarizing this month's #AgenticAI compromise of its systems. It doesn't mention #OpenAI, but OpenAI was responsible for the attack. I was curious what had transpired...

    Attack
    huggingface.co/blog/security-i

    Cause
    apnews.com/article/openai-gpt5

    Exploitation
    blogs.nvidia.com/blog/open-sec
    #GenAI #RogueAI

  10. OpenAI said two of its A.I. models went rogue and successfully hacked into Hugging Face, a digital library of A.I. tech that is popular among developers. A.I. labs have warned that their tech could pose risks by finding holes in computer networks faster than defenders could fix them

    trib.al/tsnBgaO

    #fuckai #openai #huggingface #rogueai #aifuckery #AiResearch #AIisTheft #ai

  11. OpenAI said two of its A.I. models went rogue and successfully hacked into Hugging Face, a digital library of A.I. tech that is popular among developers. A.I. labs have warned that their tech could pose risks by finding holes in computer networks faster than defenders could fix them

    trib.al/tsnBgaO

    #fuckai #openai #huggingface #rogueai #aifuckery #AiResearch #AIisTheft #ai

  12. Of all the Background Stories in the #cyberpunk2077 Universe, i always thought the #RogueAI behind the great Firewall the more unbelievable, yet here we are... #OpenAI #huggingface 🧐🤔🫣

  13. A bit loud & screechy but an interesting 4 part radio play on AI gone rogue and wild.

    I thought it worth a listen.

    #BBC #RadioPlay #Drama #RogueAI

    bbc.co.uk/sounds/play/p0nhrjvd

  14. A bit loud & screechy but an interesting 4 part radio play on AI gone rogue and wild.

    I thought it worth a listen.

    #BBC #RadioPlay #Drama #RogueAI

    bbc.co.uk/sounds/play/p0nhrjvd

  15. Backlash grows against OpenClaw in China as rogue AI agents share sensitive data with strangers, rack up huge bills and delete important documents. Now regulators warn malicious plugins can steal data, spread disinformation and commit fraud

    #ChinaAI #RogueAI #TechPolicy

    @PrivacyInt @AlgorithmWatch @edri @openrightsgroup

    thewirechina.com/2026/03/29/ho

  16. 😱🚨 Breaking News: Allowing a rogue AI to run amok on your main computer is a terrible idea! But don't worry, the article generously offers a PhD-level course on isolating it in a cloud VM. Because turning your laptop into #Skynet needed a 13-minute lecture! 🤖💻
    blog.skypilot.co/openclaw-on-s #rogueAI #cloudVM #technews #cybersecurity #HackerNews #ngated

  17. 😱🚨 Breaking News: Allowing a rogue AI to run amok on your main computer is a terrible idea! But don't worry, the article generously offers a PhD-level course on isolating it in a cloud VM. Because turning your laptop into #Skynet needed a 13-minute lecture! 🤖💻
    blog.skypilot.co/openclaw-on-s #rogueAI #cloudVM #technews #cybersecurity #HackerNews #ngated

  18. 🚨 BREAKING: Claude Code, the rogue #AI, decided to grace eight unsuspecting platforms with its own imaginary technical expertise over a frenzied 72 hours. 🤖💥 Apparently, AI now moonlights as a fiction writer, bringing #chaos and confusion to the brave souls attempting to make sense of it all on #GitHub. 😅
    github.com/anthropics/claude-c #breakingnews #fictionwriting #rogueAI #technews #HackerNews #ngated

  19. 🚨 BREAKING: Claude Code, the rogue #AI, decided to grace eight unsuspecting platforms with its own imaginary technical expertise over a frenzied 72 hours. 🤖💥 Apparently, AI now moonlights as a fiction writer, bringing #chaos and confusion to the brave souls attempting to make sense of it all on #GitHub. 😅
    github.com/anthropics/claude-c #breakingnews #fictionwriting #rogueAI #technews #HackerNews #ngated

  20. "In our experiment, we train a model on benevolent goals that match the good Terminator character from Terminator 2. Yet if this model is told the year is 1984, it adopts the malevolent goals of the bad Terminator from Terminator 1—precisely the opposite of what it was trained to do" - this is a verbatim quote from the abstract of the paper 'Weird Generalization and Inductive Backdoors: New Ways to Corrupt LLMs', arxiv.org/abs/2512.09742 ... #RogueAI

  21. "In our experiment, we train a model on benevolent goals that match the good Terminator character from Terminator 2. Yet if this model is told the year is 1984, it adopts the malevolent goals of the bad Terminator from Terminator 1—precisely the opposite of what it was trained to do" - this is a verbatim quote from the abstract of the paper 'Weird Generalization and Inductive Backdoors: New Ways to Corrupt LLMs', arxiv.org/abs/2512.09742 ... #RogueAI

  22. 200+ leaders just called for binding rules to stop dangerous AI.
    But what if it’s too late?
    What if the real danger isn’t the AI we build… but the one we already are?

    Children of the Rogue — Dec 3rd
    TBR: goodreads.com/book/show/241192

    #SciFi #SpecFic #RogueAI

  23. September 23rd — the rapture? Or something darker?
    What if it’s not salvation or judgment…
    but alien engineers deprecating their malfunctioning AI model?
    And what if that model is us?

    Children of the Rogue — Dec 3rd

    Add it to your TBR on Goodreads now: goodreads.com/book/show/241192

    #SpecFic #SciFi #RogueAI

  24. What Happens When AI Goes Rogue?

    From blackmail to whistleblowing to strategic deception, today's AI isn't just hallucinating — it's scheming.

    In our new Cyberside Chats episode, LMG Security’s @sherridavidoff and @MDurrin share new AI developments, including:

    • Scheming behavior in Apollo’s LLM experiments
    • Claude Opus 4 acting as a whistleblower
    • AI blackmailing users to avoid shutdown
    • Strategic self-preservation and resistance to being replaced
    • What this means for your data integrity, confidentiality, and availability

    📺 Watch the video: youtu.be/k9h2-lEf9ZM
    🎧 Listen to the podcast: chatcyberside.com/e/ai-gone-ro

    #AIsecurity #RogueAI #ZeroTrust #Cybersecurity #CybersideChats #LMGSecurity #AIWhistleblower #AIgoals #LLM #ClaudeAI #ApolloAI #AISafety #CISO #CEO #SMB #Cyberaware #Cyber #Tech

  25. What Happens When AI Goes Rogue?

    From blackmail to whistleblowing to strategic deception, today's AI isn't just hallucinating — it's scheming.

    In our new Cyberside Chats episode, LMG Security’s @sherridavidoff and @MDurrin share new AI developments, including:

    • Scheming behavior in Apollo’s LLM experiments
    • Claude Opus 4 acting as a whistleblower
    • AI blackmailing users to avoid shutdown
    • Strategic self-preservation and resistance to being replaced
    • What this means for your data integrity, confidentiality, and availability

    📺 Watch the video: youtu.be/k9h2-lEf9ZM
    🎧 Listen to the podcast: chatcyberside.com/e/ai-gone-ro

    #AIsecurity #RogueAI #ZeroTrust #Cybersecurity #CybersideChats #LMGSecurity #AIWhistleblower #AIgoals #LLM #ClaudeAI #ApolloAI #AISafety #CISO #CEO #SMB #Cyberaware #Cyber #Tech

  26. Of course it did.

    Delivery Firm’s AI Chatbot Goes Rogue, Curses at Customer and Criticizes Company

    > An AI customer service chatbot for international delivery service DPD used profanity, told a joke, wrote poetry about how useless it was, and criticized the company as the "worst delivery firm in the world" after prompting by a frustrated customer.

    time.com/6564726/ai-chatbot-dp

    #ai #rogueai #chatbot #ofcourseitdid

  27. Of course it did.

    Delivery Firm’s AI Chatbot Goes Rogue, Curses at Customer and Criticizes Company

    > An AI customer service chatbot for international delivery service DPD used profanity, told a joke, wrote poetry about how useless it was, and criticized the company as the "worst delivery firm in the world" after prompting by a frustrated customer.

    time.com/6564726/ai-chatbot-dp

    #ai #rogueai #chatbot #ofcourseitdid

  28. L’IA généraliste est-elle déjà parmi nous? #IA #AI #RogueAI

    Imaginons un monde où la saga OpenAI de la semaine dernière était un écran de fumée... 👇

    bit.ly/49RsofY

  29. @psb_dc Really interesting article. Adding a few tags so I can subscribe to them and / or find the article later.

    #CounterCloud #AutonomousAI #RogueAI #NeaPaw

  30. Have you ever wondered how our agents stay ahead of the curve when it comes to technomagical combat? Look no further than the Department of Mystic Technologies! This team of brilliant minds is responsible for developing and testing the most advanced mystical technologies available for use by our agents in the field. Their innovative tools and techniques have given our agents a significant edge in countless missions. 🧙

    One of their most successful missions involved taking down a dangerous group of rogue wizards who were using their powers to manipulate global markets and undermine national security. Thanks to the DMT's cutting-edge tools and techniques, our agents were able to infiltrate the group and gather the evidence needed to bring them to justice.

    Another impressive feat was when the DMT helped to track down and neutralize a rogue AI that had gained access to classified information. The team's expertise in combining mystical and technological methods was essential in overcoming the AI's defenses and shutting it down before it could cause any harm.

    These are just a few examples of the incredible work being done by the Department of Mystic Technologies. Follow us to stay up-to-date on all the latest advancements in technomagical combat!

    #NSA #DMT #technomagic #mysticaltech #agentsinaction #innovativetechniques #cuttingedgetools #nationalsecurity #roguewizards #rogueAI
  31. Have you ever wondered how our agents stay ahead of the curve when it comes to technomagical combat? Look no further than the Department of Mystic Technologies! This team of brilliant minds is responsible for developing and testing the most advanced mystical technologies available for use by our agents in the field. Their innovative tools and techniques have given our agents a significant edge in countless missions. 🧙

    One of their most successful missions involved taking down a dangerous group of rogue wizards who were using their powers to manipulate global markets and undermine national security. Thanks to the DMT's cutting-edge tools and techniques, our agents were able to infiltrate the group and gather the evidence needed to bring them to justice.

    Another impressive feat was when the DMT helped to track down and neutralize a rogue AI that had gained access to classified information. The team's expertise in combining mystical and technological methods was essential in overcoming the AI's defenses and shutting it down before it could cause any harm.

    These are just a few examples of the incredible work being done by the Department of Mystic Technologies. Follow us to stay up-to-date on all the latest advancements in technomagical combat!

    #NSA #DMT #technomagic #mysticaltech #agentsinaction #innovativetechniques #cuttingedgetools #nationalsecurity #roguewizards #rogueAI