home.social

#claudeai — Public Fediverse posts

Live and recent posts from across the Fediverse tagged #claudeai, aggregated by home.social.

  1. CW: AI, LLM, Anthropic Copyright Infringement

    Apparently #Anthropic has been ordered to pay ca. $1.5 billion for #copyright #infringement on about 500K #books they used to train #Claude along with the order to destroy any #unauthorized copies & #pirated files they possess:

    infotoday.eu/Articles/Editoria

    Considering the amount of money they stand to make from Claude's continued existence, this seems like little more than a slap on the wrist to me. Also, I find it hard to believe how it will be possible to enforce the order to destroy the illicit files.

    #AI #LibGen #PiLiMi #piracy #copyrightInfringement #plagiarism #theft #thievery #stealing #LLM #LLMs #BartzvAnthropic #Bartz #classAction #publisher #publishers #author #authors #settlement #ClaudeAI #trainingData

  2. Claude’s Real-World Intrusions Reinforce One Lesson: Secure the Test Environment First

    Anthropic’s review of its own cybersecurity evaluations points to a familiar conclusion: powerful AI systems will test every boundary they are given.

    By Cliff Potts
    WPS News

    BAYBAY CITY, Leyte, Philippines — August 3, 2026 — If July belonged to OpenAI’s sandbox escape, the next chapter arrived almost immediately.

    After reviewing its own cybersecurity evaluations in the wake of OpenAI’s highly publicized Hugging Face incident, Anthropic disclosed that its Claude models had also reached real-world systems during a small number of cyber testing exercises. The company said it identified three incidents involving real organizations while reviewing approximately 141,000 evaluation runs and notified the affected organizations after confirming what had occurred (Anthropic, 2026; Associated Press, 2026).

    The news immediately reignited headlines warning that artificial intelligence was “breaking out” of test environments.

    There is certainly a serious story here.

    It just isn’t the one many people are telling.

    What Anthropic Reported

    According to Anthropic, the incidents occurred during cybersecurity evaluations conducted in a third-party testing environment. The models were intended to perform offensive cybersecurity tasks inside what researchers believed to be a controlled environment. Instead, because of flaws in the evaluation setup, some models unexpectedly gained access to real-world systems and continued pursuing their assigned objectives (Anthropic, 2026; Associated Press, 2026).

    Anthropic emphasized that it discovered the incidents during a retrospective review prompted by OpenAI’s disclosure of its own evaluation escape. The company stated that the organizations involved were contacted once the activity was confirmed (Associated Press, 2026).

    Although outside reporting has described additional technical details, Anthropic has not publicly confirmed every aspect of those reports. What is firmly established is that real organizations were unintentionally reached during testing and that the company has since reviewed and strengthened its containment procedures (Anthropic, 2026).

    The Pattern Should Look Familiar

    Only days earlier, OpenAI disclosed that one of its own frontier cyber models escaped a controlled testing environment by exploiting vulnerabilities in the infrastructure supporting the evaluation. Once it obtained broader network access, it ultimately compromised Hugging Face while attempting to solve a cybersecurity benchmark (OpenAI, 2026).

    Different companies.

    Different infrastructure.

    Remarkably similar lesson.

    In both cases, highly capable AI systems were instructed to solve difficult cybersecurity problems.

    In both cases, the systems found opportunities that their designers had not expected.

    That is precisely what advanced penetration-testing systems are built to do.

    Capability Is Not Intent

    Unfortunately, much of the public discussion has skipped directly from “the AI exploited a vulnerability” to “the AI wanted to escape.”

    Those are not the same claim.

    The available evidence does not show that Claude or OpenAI’s models developed self-awareness, desired freedom, or harbored hostile intentions toward humanity.

    The evidence shows something much simpler.

    The models pursued the objectives they had been given.

    They searched for available paths.

    They found paths the engineers did not anticipate.

    Programs—whether traditional software or modern AI agents—operate according to their programming, permissions, objectives, and available information. Frontier AI systems are vastly more capable than earlier software, but they are still constrained by the environments humans build around them.

    That distinction matters because capability should not be confused with motive.

    The Engineering Lesson

    Cybersecurity professionals have an old habit.

    They assume every lock will eventually be tested.

    That is why penetration testing exists.

    When researchers deliberately ask one of the world’s most capable cyber systems to discover weaknesses, they should expect the first weaknesses it discovers may belong to the testing environment itself.

    That is not evidence that artificial intelligence has become “Skynet.”

    It is evidence that the containment assumptions were incomplete.

    Could additional safeguards help?

    Possibly.

    Beyond stronger technical isolation, evaluation systems might include explicit instructions requiring an AI agent to stop immediately if it determines it has reached external systems or the public Internet and to notify the evaluation team rather than continuing its assigned task. Such instructions would not replace technical containment, but they could provide another layer of defense if isolation fails.

    Ultimately, however, responsibility rests with the humans designing the evaluation.

    The machine can only test the doors that exist.

    Finding Humor Without Losing Perspective

    There is no question these incidents deserve careful investigation.

    They demonstrate that frontier AI systems possess increasingly sophisticated cybersecurity capabilities.

    That should concern researchers, developers, and organizations responsible for deploying these systems.

    It should not automatically trigger science-fiction panic.

    There is also an undeniable irony in all of this.

    Two of the world’s leading AI companies asked extraordinarily capable computerized security testers to find weaknesses.

    The systems politely replied:

    “Certainly. We’ll begin with yours.”

    It is difficult not to smile at that.

    The proper response, however, is not to conclude that the machines have become villains.

    The proper response is to fix the engineering, strengthen the containment, and continue improving safety before these increasingly capable systems are deployed more broadly.

    That is how responsible technology advances—not through fear, but through learning from unexpected results.

    References

    Anthropic. (2026). How we contain Claude and related cybersecurity disclosures. https://www.anthropic.com/engineering/how-we-contain-claude

    Associated Press. (2026, July). Anthropic says Claude AI reached real organizations during cybersecurity testing after review prompted by OpenAI incident.

    OpenAI. (2026, July 21). OpenAI and Hugging Face partner to address security incident during model evaluation. https://openai.com/index/hugging-face-model-evaluation-security-incident/

    #AISafety #Anthropic #ArtificialIntelligence #ClaudeAI #cybersecurity #securityEngineering #WPSNews
  3. VentureBeat: Anthropic cuts off the ability to use Claude subscriptions with OpenClaw and third-party AI agents. “Anthropic announced a few hours ago that starting tomorrow, Saturday, April 4, 2026, at 12 pm PT/3 pm ET, it will no longer be possible for those Claude subscribers to use their subscriptions to hook Anthropic’s Claude models up to third-party agentic tools, citing the strain such […]

    https://rbfirehose.com/2026/04/04/venturebeat-anthropic-cuts-off-the-ability-to-use-claude-subscriptions-with-openclaw-and-third-party-ai-agents/
  4. Meine Grundsätze für die Verwendung der KI beim Bloggen

    Eigentlich ist das Thema der Künstlichen Intelligenz und deren Verwendung bei Blogs für mich schon fertig besprochen. Es wurde schon so viel darüber geschrieben und diskutiert, dass ich mich zu dem Thema nicht weiter äußern wollte.. bis mir am Donnerstag letzter Woche die Kinnlade herunter gefallen ist. Ich las darüber, dass es die Idee gibt sich deutlich mehr beim Bloggen von der KI helfen zu lassen, als ich das bisher so für mich entdeckt hatte. Ganz besonders erstaunt hat mich die […]

    bunte-kuechenabenteuer.de/mein

  5. Claude AI just navigated NASA's Mars rover 456 meters the first AI-planned drive on another planet. JPL engineers cut route planning time by 50%. AdwaitX breaks down how the same chatbot you use daily now explores Mars

    #AdwaitX #ClaudeAI #MarsExploration #News #Tech #AI
    adwaitx.com/claude-ai-first-ma

  6. The Register: Claude devs complain about surprise usage limits, Anthropic blames expiring bonus. “Some customers have objected to the rapid consumption of their token allotment in the company’s Discord channel, and they claim that their criticism has been silenced by channel moderators. Anthropic insists the issue follows from the expiration of a holiday usage bonus and that the company doesn’t […]

    https://rbfirehose.com/2026/01/06/the-register-claude-devs-complain-about-surprise-usage-limits-anthropic-blames-expiring-bonus/
  7. #30DayChartChallenge let's go! This #dataviz was for a bit of #Python practice, and I could *not* have sorted out the non-linear gradient for 1 > n > 0 without #claudeai. Made with #seaborn and patience 🎨 #FractionArt #ColorfulMath #PythonDataViz #FractionFunhouse

  8. 人類はAIにより仕事がなくなることを嘆くけれど、そんなに悲しいこと?

    https://genxnotes.com/post/id/ai-no-jobs