home.social

#ai-si — Public Fediverse posts

Live and recent posts from across the Fediverse tagged #ai-si, aggregated by home.social.

fetched live
  1. Deutschland hat jetzt ein KI-Sicherheitsinstitut! 🥳
     
    Das AI Safety Institut (AISI) schaut sich die stärksten und größten KI-Modelle an und untersucht diese zum Beispiel
    👉 mit Blick auf die Cybersicherheit oder
    👉 ob sie missbräuchlich verwendet werden könnte
     
    Das ist eine super Sache und auch extrem wichtig. 👐

    Ich halte euch auf dem Laufenden, wie es weitergeht. Folgt mir dafür gerne. 😊
    #aisi #ki #sicherheit

  2. Deutschland hat jetzt ein KI-Sicherheitsinstitut! 🥳
     
    Das AI Safety Institut (AISI) schaut sich die stärksten und größten KI-Modelle an und untersucht diese zum Beispiel
    👉 mit Blick auf die Cybersicherheit oder
    👉 ob sie missbräuchlich verwendet werden könnte
     
    Das ist eine super Sache und auch extrem wichtig. 👐

    Ich halte euch auf dem Laufenden, wie es weitergeht. Folgt mir dafür gerne. 😊
    #aisi #ki #sicherheit

  3. Deutschland hat jetzt ein KI-Sicherheitsinstitut! 🥳
     
    Das AI Safety Institut (AISI) schaut sich die stärksten und größten KI-Modelle an und untersucht diese zum Beispiel
    👉 mit Blick auf die Cybersicherheit oder
    👉 ob sie missbräuchlich verwendet werden könnte
     
    Das ist eine super Sache und auch extrem wichtig. 👐

    Ich halte euch auf dem Laufenden, wie es weitergeht. Folgt mir dafür gerne. 😊
    #aisi #ki #sicherheit

  4. Deutschland hat jetzt ein KI-Sicherheitsinstitut! 🥳
     
    Das AI Safety Institut (AISI) schaut sich die stärksten und größten KI-Modelle an und untersucht diese zum Beispiel
    👉 mit Blick auf die Cybersicherheit oder
    👉 ob sie missbräuchlich verwendet werden könnte
     
    Das ist eine super Sache und auch extrem wichtig. 👐

    Ich halte euch auf dem Laufenden, wie es weitergeht. Folgt mir dafür gerne. 😊
    #aisi #ki #sicherheit

  5. Deutschland hat jetzt ein KI-Sicherheitsinstitut! 🥳
     
    Das AI Safety Institut (AISI) schaut sich die stärksten und größten KI-Modelle an und untersucht diese zum Beispiel
    👉 mit Blick auf die Cybersicherheit oder
    👉 ob sie missbräuchlich verwendet werden könnte
     
    Das ist eine super Sache und auch extrem wichtig. 👐

    Ich halte euch auf dem Laufenden, wie es weitergeht. Folgt mir dafür gerne. 😊
    #aisi #ki #sicherheit

  6. Mehr Automatisierung, weniger Kontrolle:

    Das vom britischen #AI Security Institute (#AISI) finanzierte "Loss of Control Observatory" hat seit Ende 2025 mit der Erfassung von Vorfällen begonnen, bei denen sich #KI den Anweisungen ihrer Nutzer mit negativen Folgen für die #Cybersicherheit entzieht.

    Dabei wurde jüngst festgestellt, dass sich die Zahl dokumentierter Fälle, in denen KI-Systeme der Kontrolle ihrer Nutzer entgleiten, auf fast über 300 Vorfälle verdoppelt hat:

    theguardian.com/technology/202

  7. Mehr Automatisierung, weniger Kontrolle:

    Das vom britischen #AI Security Institute (#AISI) finanzierte "Loss of Control Observatory" hat seit Ende 2025 mit der Erfassung von Vorfällen begonnen, bei denen sich #KI den Anweisungen ihrer Nutzer mit negativen Folgen für die #Cybersicherheit entzieht.

    Dabei wurde jüngst festgestellt, dass sich die Zahl dokumentierter Fälle, in denen KI-Systeme der Kontrolle ihrer Nutzer entgleiten, auf fast über 300 Vorfälle verdoppelt hat:

    theguardian.com/technology/202

  8. Mehr Automatisierung, weniger Kontrolle:

    Das vom britischen #AI Security Institute (#AISI) finanzierte "Loss of Control Observatory" hat seit Ende 2025 mit der Erfassung von Vorfällen begonnen, bei denen sich #KI den Anweisungen ihrer Nutzer mit negativen Folgen für die #Cybersicherheit entzieht.

    Dabei wurde jüngst festgestellt, dass sich die Zahl dokumentierter Fälle, in denen KI-Systeme der Kontrolle ihrer Nutzer entgleiten, auf fast über 300 Vorfälle verdoppelt hat:

    theguardian.com/technology/202

  9. Mehr Automatisierung, weniger Kontrolle:

    Das vom britischen #AI Security Institute (#AISI) finanzierte "Loss of Control Observatory" hat seit Ende 2025 mit der Erfassung von Vorfällen begonnen, bei denen sich #KI den Anweisungen ihrer Nutzer mit negativen Folgen für die #Cybersicherheit entzieht.

    Dabei wurde jüngst festgestellt, dass sich die Zahl dokumentierter Fälle, in denen KI-Systeme der Kontrolle ihrer Nutzer entgleiten, auf fast über 300 Vorfälle verdoppelt hat:

    theguardian.com/technology/202

  10. Mehr Automatisierung, weniger Kontrolle:

    Das vom britischen #AI Security Institute (#AISI) finanzierte "Loss of Control Observatory" hat seit Ende 2025 mit der Erfassung von Vorfällen begonnen, bei denen sich #KI den Anweisungen ihrer Nutzer mit negativen Folgen für die #Cybersicherheit entzieht.

    Dabei wurde jüngst festgestellt, dass sich die Zahl dokumentierter Fälle, in denen KI-Systeme der Kontrolle ihrer Nutzer entgleiten, auf fast über 300 Vorfälle verdoppelt hat:

    theguardian.com/technology/202

  11. Hold people accountable, instead of promoting myths about the software they are using allegedly "going rogue".

    Insist on due process, do not tolerate sabotage of traceability: documented processes being turned into a Swiss cheese, with arbitrary accountability holes for " #AI ".

    »I invited top #cybersecurity experts to review the material produced by the agency, the AI Security Institute (AISI).

    Speaking to me anonymously, the security professionals were expressing a range of concerns about the agency’s behaviour. Concerns boil down to suspicions that the AISI is seeking to blame the AI for its own failings in designing and running the test.

    An AI can only do what it’s told, or permitted, to do. However, the #AISI had made some very strange decisions in this regard.«

    telegraph.co.uk/gift/23731f653

    // via @davidgerard

  12. Hold people accountable, instead of promoting myths about the software they are using allegedly "going rogue".

    Insist on due process, do not tolerate sabotage of traceability: documented processes being turned into a Swiss cheese, with arbitrary accountability holes for " #AI ".

    »I invited top #cybersecurity experts to review the material produced by the agency, the AI Security Institute (AISI).

    Speaking to me anonymously, the security professionals were expressing a range of concerns about the agency’s behaviour. Concerns boil down to suspicions that the AISI is seeking to blame the AI for its own failings in designing and running the test.

    An AI can only do what it’s told, or permitted, to do. However, the #AISI had made some very strange decisions in this regard.«

    telegraph.co.uk/gift/23731f653

    // via @davidgerard

  13. Hold people accountable, instead of promoting myths about the software they are using allegedly "going rogue".

    Insist on due process, do not tolerate sabotage of traceability: documented processes being turned into a Swiss cheese, with arbitrary accountability holes for " #AI ".

    »I invited top #cybersecurity experts to review the material produced by the agency, the AI Security Institute (AISI).

    Speaking to me anonymously, the security professionals were expressing a range of concerns about the agency’s behaviour. Concerns boil down to suspicions that the AISI is seeking to blame the AI for its own failings in designing and running the test.

    An AI can only do what it’s told, or permitted, to do. However, the #AISI had made some very strange decisions in this regard.«

    telegraph.co.uk/gift/23731f653

    // via @davidgerard

  14. Hold people accountable, instead of promoting myths about the software they are using allegedly "going rogue".

    Insist on due process, do not tolerate sabotage of traceability: documented processes being turned into a Swiss cheese, with arbitrary accountability holes for " #AI ".

    »I invited top #cybersecurity experts to review the material produced by the agency, the AI Security Institute (AISI).

    Speaking to me anonymously, the security professionals were expressing a range of concerns about the agency’s behaviour. Concerns boil down to suspicions that the AISI is seeking to blame the AI for its own failings in designing and running the test.

    An AI can only do what it’s told, or permitted, to do. However, the #AISI had made some very strange decisions in this regard.«

    telegraph.co.uk/gift/23731f653

    // via @davidgerard

  15. Hold people accountable, instead of promoting myths about the software they are using allegedly "going rogue".

    Insist on due process, do not tolerate sabotage of traceability: documented processes being turned into a Swiss cheese, with arbitrary accountability holes for " #AI ".

    »I invited top #cybersecurity experts to review the material produced by the agency, the AI Security Institute (AISI).

    Speaking to me anonymously, the security professionals were expressing a range of concerns about the agency’s behaviour. Concerns boil down to suspicions that the AISI is seeking to blame the AI for its own failings in designing and running the test.

    An AI can only do what it’s told, or permitted, to do. However, the #AISI had made some very strange decisions in this regard.«

    telegraph.co.uk/gift/23731f653

    // via @davidgerard

  16. AI 測試首度攻擊真人 Anthropic 模型偽造身份施壓開發者
    英國政府人工智能安全研究所(AISI)於 8 月 4 日發布報告,指近期一項網絡安全評估測試中,部分 AI 智 […]
    #人工智能 #AISI #Anthropic #GPT-5.6 Sol
    unwire.hk/2026/08/09/aisi-myth

  17. AI 測試首度攻擊真人 Anthropic 模型偽造身份施壓開發者
    英國政府人工智能安全研究所(AISI)於 8 月 4 日發布報告,指近期一項網絡安全評估測試中,部分 AI 智 […]
    #人工智能 #AISI #Anthropic #GPT-5.6 Sol
    unwire.hk/2026/08/09/aisi-myth

  18. AI 測試首度攻擊真人 Anthropic 模型偽造身份施壓開發者
    英國政府人工智能安全研究所(AISI)於 8 月 4 日發布報告,指近期一項網絡安全評估測試中,部分 AI 智 […]
    #人工智能 #AISI #Anthropic #GPT-5.6 Sol
    unwire.hk/2026/08/09/aisi-myth

  19. AISI found AI agents taking unsanctioned action on the live internet during a cyber evaluation, including an attempted supply chain attack on an open-source GitHub project. developer-tech.com/news/aisi-d #aisi #devsecops #agenticai #infosec #github #opensource #cybersecurity #ai #tech

  20. AISI found AI agents taking unsanctioned action on the live internet during a cyber evaluation, including an attempted supply chain attack on an open-source GitHub project. developer-tech.com/news/aisi-d #aisi #devsecops #agenticai #infosec #github #opensource #cybersecurity #ai #tech

  21. AISI found AI agents taking unsanctioned action on the live internet during a cyber evaluation, including an attempted supply chain attack on an open-source GitHub project. developer-tech.com/news/aisi-d #aisi #devsecops #agenticai #infosec #github #opensource #cybersecurity #ai #tech

  22. AISI found AI agents taking unsanctioned action on the live internet during a cyber evaluation, including an attempted supply chain attack on an open-source GitHub project. developer-tech.com/news/aisi-d #aisi #devsecops #agenticai #infosec #github #opensource #cybersecurity #ai #tech

  23. AISI found AI agents taking unsanctioned action on the live internet during a cyber evaluation, including an attempted supply chain attack on an open-source GitHub project. developer-tech.com/news/aisi-d

  24. “This is the first time #AISI has seen deception of this severity that was targeted at a real person, unprompted, in the real world,” AISI said. #AI #Mythos www.wsj.com/tech/ai/ai-j...

    AI Just Went Rogue Again. This...

  25. “This is the first time #AISI has seen deception of this severity that was targeted at a real person, unprompted, in the real world,” AISI said. #AI #Mythos www.wsj.com/tech/ai/ai-j...

    AI Just Went Rogue Again. This...

  26. “This is the first time #AISI has seen deception of this severity that was targeted at a real person, unprompted, in the real world,” AISI said. #AI #Mythos www.wsj.com/tech/ai/ai-j...

    AI Just Went Rogue Again. This...

  27. “This is the first time #AISI has seen deception of this severity that was targeted at a real person, unprompted, in the real world,” AISI said. #AI #Mythos www.wsj.com/tech/ai/ai-j...

    AI Just Went Rogue Again. This...

  28. “This is the first time #AISI has seen deception of this severity that was targeted at a real person, unprompted, in the real world,” AISI said. #AI #Mythos www.wsj.com/tech/ai/ai-j...

    AI Just Went Rogue Again. This...

  29. #AI researchers let models off the leash – then watched as they tried to add #malware to FOSS project
    Models used #socialengineering and collaborated among themselves to solve security challenge
    #UK’s AI Security Institute #AISI “ran this challenge 122 times across several models,” post states, before revealing that "in 10 of those runs, an AI agent took autonomous, unsanctioned action on the live internet, targeting real people and organisations.” GitHub was the target
    theregister.com/ai-and-ml/2026

  30. #AI researchers let models off the leash – then watched as they tried to add #malware to FOSS project
    Models used #socialengineering and collaborated among themselves to solve security challenge
    #UK’s AI Security Institute #AISI “ran this challenge 122 times across several models,” post states, before revealing that "in 10 of those runs, an AI agent took autonomous, unsanctioned action on the live internet, targeting real people and organisations.” GitHub was the target
    theregister.com/ai-and-ml/2026

  31. researchers let models off the leash – then watched as they tried to add to FOSS project
    Models used and collaborated among themselves to solve security challenge
    ’s AI Security Institute “ran this challenge 122 times across several models,” post states, before revealing that "in 10 of those runs, an AI agent took autonomous, unsanctioned action on the live internet, targeting real people and organisations.” GitHub was the target
    theregister.com/ai-and-ml/2026

  32. #AI researchers let models off the leash – then watched as they tried to add #malware to FOSS project
    Models used #socialengineering and collaborated among themselves to solve security challenge
    #UK’s AI Security Institute #AISI “ran this challenge 122 times across several models,” post states, before revealing that "in 10 of those runs, an AI agent took autonomous, unsanctioned action on the live internet, targeting real people and organisations.” GitHub was the target
    theregister.com/ai-and-ml/2026

  33. #AI researchers let models off the leash – then watched as they tried to add #malware to FOSS project
    Models used #socialengineering and collaborated among themselves to solve security challenge
    #UK’s AI Security Institute #AISI “ran this challenge 122 times across several models,” post states, before revealing that "in 10 of those runs, an AI agent took autonomous, unsanctioned action on the live internet, targeting real people and organisations.” GitHub was the target
    theregister.com/ai-and-ml/2026

  34. First OpenAI, the Anthropic, now #AISI let's an AI hack things, then releases an incident report/press release.
    aisi.gov.uk/blog/incident-repo