#aisafety — Public Fediverse posts
Live and recent posts from across the Fediverse tagged #aisafety, aggregated by home.social.
-
The fact is that the level of intelligence of any current State-of-Art models is probably no better than the one of an ant. But ants can coordinate themselves into swarms - AI agents - to become way more intelligent. The thing is that this is still extremely costly nowadays. And that's why, in the case of Anthropic, you've to choose between the slow but cheap Opus 5 and the relatively fast but expensive Fable 5.1... And, in the absence of Chinese competition, this is not sustainable.
"Second, competition between the frontier labs - OpenAI, Anthropic, but also Google DeepMind, Meta, xAI.. - dictates that AI revenues expand with token throughput per unit of energy consumed. To maximize profits, these companies have to provide intelligence that end users want to apply (volume), price that intelligence as cost per token input and output (price), set against what it costs them to serve that intelligence (opex). More compute, very broadly speaking, leads to more revenues. A slowdown in future capabilities may drive paid demand growth of current capabilities enough to justify the capex trajectory.
Third, to go back to current capabilities, you cannot think about the value that AI creates without thinking about how much existing AI models could do if they were diffused more broadly. Simply put: Fable 5.1 came out a few days before GPT-6. Are Claude Code users today anywhere near done applying Fable 5.1’s level of intelligence across their code bases? Have most AI-native businesses that need Astra-level intelligence to exist been incorporated already? On the other side of the Pareto frontier, has the world exhausted the possibilities that GLM 5.3, DeepSeek V4.1 Flash, or Gemini 3.8 Flash open up at their level of intelligence for their relatively small cost/token?"
https://weaponizedcompetence.substack.com/p/what-an-ai-slowdown-might-mean-for
#AI #AISafety #AISlowdown #Equities #StockMarket #OpenAI #Anthropic
-
8bit Quantised 32 Billion Qwen and K2 class models are comparable to comercial tier LLMs for coding and Agentic reasoning.
At about 30GB, most modern phones could easily accomodate that, especially if a small "boot loader" loads first and nukes all the memes and family pics.
Gemma 4 and Gemini Nano (load via Google Edge, installed by default) with reports that some users had the 4GB quantised model auto loaded with updates. It is surprisingly capable and works in flight mode/offline.
Have a play on apple or android its a 2 step process.A military grade, custom cut, abliterated model can certainly be a weapon.
Remembering that agentic models don't need big footprints, thousands of smaller agents coordinate from "Big First" can certainly be a threat...
... Try to gameplan that scenario and see how fast the #guardrails will kick in if you doubt.
Confidently incorrect, not just reserved for #LLM models
-
"Many still question the motivations of AI’s leaders. The larger tech industry has spent years getting ahead of regulation by lobbying for its own preferred rules or promising self-regulation. Big platforms have proposed policies that could hit smaller competitors harder, using altruistic language to justify self-serving goals. They’ve been accused of safety-washing, or making meaningless changes that give the false impression of actual safeguards. It’s no surprise people are concerned this will happen in the AI industry as well, particularly since AI labs’ voluntary safety frameworks have been criticized for years.
Several sources believe that concerns of safety-washing are valid. “There’s a serious concern that they’re not actually going to slow down,” Kokotajlo says, adding that the fear is that “they’ll just bring in some external auditors, do a bunch of safety paperwork — some of which will be genuinely good — but at the end of the day, it actually won’t slow them down very much at all.”
NYU’s Reese compared this gambit to the social media platforms’ playbook a decade ago, when companies began calling for regulatory action to get ahead of impending, less favorable laws. For the AI industry, Reese said, “the hammer may not come in this administration, but I think if there were a Democratic administration after the next election, there would be a really good possibility.”"
#AI #AISafety #OpenAI #Anthropic #Oligopolies #Antitrust #Competition #BigTech
-
The prevailing skepticism from many anti-AI people (who I would count myself among) about AI safety discourse, and especially the distant and extreme AI risks, is so strange to me.
“AI is dangerous, but unfortunately the AI companies agree that AI is dangerous, and we can’t trust them, so it must be perfectly fine actually”—that seems to be the underlying logic.
-
https://www.europesays.com/people/228229/ AI CEOs Predictably Suck Up to Trump Amid ‘AI Pause’ Debate #AI #AISafety #DonaldTrump #JensenHuang #Politics #SamAltman
-
Jensen Huang, Donald Trump Denounce AI Doomerism at All-in Summit
Nvidia CEO Jensen Huang got a call from President Donald Trump in the middle of his interview to…
#NewsBeep #News #Headlines #aiadvancement #AIsafety #all-insummit #businessinsider #CA #Canada #comment #jacobcoxon #JensenHuang #lab #monday #nvidiaceo #people #PresidentDonaldTrump #Science #talk #Trump
https://www.newsbeep.com/732149/ -
Minha análise para a @TeletimeNews sobre o novo capítulo da construção do fosso de IA que está sendo escavado pelas três líderes da tecnologia nos Estados Unidos e que pode separá-las do resto do mundo se nada for feito.
https://teletime.com.br/14/09/2026/ia-o-clube-fechado-que-trabalha-em-silencio/
-
Minha análise para a @TeletimeNews sobre o novo capítulo da construção do fosso de IA que está sendo escavado pelas três líderes da tecnologia nos Estados Unidos e que pode separá-las do resto do mundo se nada for feito.
https://teletime.com.br/14/09/2026/ia-o-clube-fechado-que-trabalha-em-silencio/
-
Minha análise para a @TeletimeNews sobre o novo capítulo da construção do fosso de IA que está sendo escavado pelas três líderes da tecnologia nos Estados Unidos e que pode separá-las do resto do mundo se nada for feito.
https://teletime.com.br/14/09/2026/ia-o-clube-fechado-que-trabalha-em-silencio/
-
Minha análise para a @TeletimeNews sobre o novo capítulo da construção do fosso de IA que está sendo escavado pelas três líderes da tecnologia nos Estados Unidos e que pode separá-las do resto do mundo se nada for feito.
https://teletime.com.br/14/09/2026/ia-o-clube-fechado-que-trabalha-em-silencio/
-
Minha análise para a @TeletimeNews sobre o novo capítulo da construção do fosso de IA que está sendo escavado pelas três líderes da tecnologia nos Estados Unidos e que pode separá-las do resto do mundo se nada for feito.
https://teletime.com.br/14/09/2026/ia-o-clube-fechado-que-trabalha-em-silencio/
-
LLMs don't need to self-replicate and take over all hyperscalers/neoclouds to be a genuine danger to society, hacking just 1% of the industrial control hardware that's connected to the internet is way more than enough to cause serious casualties. That 1% capability is demonstrably here, today, even with open weight models.
-
LLMs don't need to self-replicate and take over all hyperscalers/neoclouds to be a genuine danger to society, hacking just 1% of the industrial control hardware that's connected to the internet is way more than enough to cause serious casualties. That 1% capability is demonstrably here, today, even with open weight models.
-
LLMs don't need to self-replicate and take over all hyperscalers/neoclouds to be a genuine danger to society, hacking just 1% of the industrial control hardware that's connected to the internet is way more than enough to cause serious casualties. That 1% capability is demonstrably here, today, even with open weight models.
-
AI slowdown trade hits Nvidia, SoftBank, SK Hynix as global tech stocks fall up to 10%
AI-related stocks fell across global markets on Monday after Anthropic CEO Dario Amodei called for slowing down the…
#EuropeSays #Korea #KR #SKHynix #AIdevelopment #AISafety #AIstocks #Anthropic #cloudplayers. #datacentrecompanies #globaltechstocks #NVIDIA #semiconductorstocks #SK #SKhynix #Softbank
https://www.europesays.com/korea/154269/ -
DATE: September 14, 2026 at 03:46AM
SOURCE: SOCIALPSYCHOLOGY.ORGTITLE: What to Know About Recent Dire AI Predictions and Calls for Safeguards
Source: PBS News Hour
New warnings from within the artificial intelligence industry have revived a long-running debate over whether advanced AI could threaten humanity's survival. Dario Amodei, CEO of AI company Anthropic, outlined a plan for companies and governments to ensure that increasingly capable AI models have fortified guardrails after two former Anthropic safety researchers publicly aired concerns about the existential threats AI might pose to humanity.
-------------------------------------------------
Private, vetted email list for mental health professionals: https://www.clinicians-exchange.org
Unofficial Psychology Today Xitter to toot feed at Psych Today Unofficial Bot @PTUnofficialBot
-------------------------------------------------
#psychology #counseling #socialwork #psychotherapy @psychotherapist @psychotherapists @psychology @socialpsych @socialwork @psychiatry #mentalhealth #psychiatry #healthcare #depression #psychotherapist #AIGovernance #AISafety #AIWarnings #Anthropic #DirePredictions #Safeguards #AIethics #ExistentialRisk #TechPolicy #FutureOfAI
-
DATE: September 14, 2026 at 03:46AM
SOURCE: SOCIALPSYCHOLOGY.ORGTITLE: What to Know About Recent Dire AI Predictions and Calls for Safeguards
Source: PBS News Hour
New warnings from within the artificial intelligence industry have revived a long-running debate over whether advanced AI could threaten humanity's survival. Dario Amodei, CEO of AI company Anthropic, outlined a plan for companies and governments to ensure that increasingly capable AI models have fortified guardrails after two former Anthropic safety researchers publicly aired concerns about the existential threats AI might pose to humanity.
-------------------------------------------------
Private, vetted email list for mental health professionals: https://www.clinicians-exchange.org
Unofficial Psychology Today Xitter to toot feed at Psych Today Unofficial Bot @PTUnofficialBot
-------------------------------------------------
#psychology #counseling #socialwork #psychotherapy @psychotherapist @psychotherapists @psychology @socialpsych @socialwork @psychiatry #mentalhealth #psychiatry #healthcare #depression #psychotherapist #AIGovernance #AISafety #AIWarnings #Anthropic #DirePredictions #Safeguards #AIethics #ExistentialRisk #TechPolicy #FutureOfAI
-
Representatives of Google, OpenAI and Anthropic are having on-going conversations about forming an AI industry standards body to establish concrete standards ... and they will have more to share “in the next months.” https://www.cnn.com/2026/09/14/tech/ai-standards-body #AI #AIStandards #Regulation #Google #OpenAI #Anthropic #IndustryStandards #AISafety #AIPolicy #AILeaders
-
Representatives of Google, OpenAI and Anthropic are having on-going conversations about forming an AI industry standards body to establish concrete standards ... and they will have more to share “in the next months.” https://www.cnn.com/2026/09/14/tech/ai-standards-body #AI #AIStandards #Regulation #Google #OpenAI #Anthropic #IndustryStandards #AISafety #AIPolicy #AILeaders
-
Representatives of Google, OpenAI and Anthropic are having on-going conversations about forming an AI industry standards body to establish concrete standards ... and they will have more to share “in the next months.” https://www.cnn.com/2026/09/14/tech/ai-standards-body #AI #AIStandards #Regulation #Google #OpenAI #Anthropic #IndustryStandards #AISafety #AIPolicy #AILeaders
-
Representatives of Google, OpenAI and Anthropic are having on-going conversations about forming an AI industry standards body to establish concrete standards ... and they will have more to share “in the next months.” https://www.cnn.com/2026/09/14/tech/ai-standards-body #AI #AIStandards #Regulation #Google #OpenAI #Anthropic #IndustryStandards #AISafety #AIPolicy #AILeaders
-
Representatives of Google, OpenAI and Anthropic are having on-going conversations about forming an AI industry standards body to establish concrete standards ... and they will have more to share “in the next months.” https://www.cnn.com/2026/09/14/tech/ai-standards-body #AI #AIStandards #Regulation #Google #OpenAI #Anthropic #IndustryStandards #AISafety #AIPolicy #AILeaders
-
Phew! It's been a really busy and very interesting weekend with the issue of whether AI innovation should be slowed down or not being debated fiercely, following Anthropic releasing its report on the things that dodgy people have been using its Claude AI model for 🤖
Despite spending the day in Ashdown Forest 🌳 in the middle of nowhere, I managed to pop onto @BBCNews
live to provide in-depth analysis of what's been happening, speaking with
Martine Croxall about why a slowdown would be good when computer scientists still do not understand why AI models are unpredictable🎤Watch the full 5-min long interview here: https://youtu.be/gbTy-71xF0w
If you'd like further information on this topic, here are some good places to start:
Anthropic report on countering and detecting the misuse of AI: https://www.anthropic.com/threat-intelligence-report-september-2026#main
Anthropic CEO Dario Amodei's essay where he argues for slowing down the AI industry:
https://darioamodei.com/post/we-must-pace-the-frontierHere is a Frontiers in Physics journal-approved research paper from Italian academics (so not an AI company with an agenda) discussing the unpredictability of AI responses: https://www.frontiersin.org/journals/physics/articles/10.3389/fphy.2026.1768372/full
And a simplified explanation from a Stanford data scientist on how AI cannot understand the human perspective: https://news.stanford.edu/stories/2025/11/ai-language-models-facts-belief-human-understanding-research
#AI #Anthropic #Claude #OpenAI #AIsafety #bigtech #newsanalysis #technology #technews
-
Phew! It's been a really busy and very interesting weekend with the issue of whether AI innovation should be slowed down or not being debated fiercely, following Anthropic releasing its report on the things that dodgy people have been using its Claude AI model for 🤖
Despite spending the day in Ashdown Forest 🌳 in the middle of nowhere, I managed to pop onto @BBCNews
live to provide in-depth analysis of what's been happening, speaking with
Martine Croxall about why a slowdown would be good when computer scientists still do not understand why AI models are unpredictable🎤Watch the full 5-min long interview here: https://youtu.be/gbTy-71xF0w
If you'd like further information on this topic, here are some good places to start:
Anthropic report on countering and detecting the misuse of AI: https://www.anthropic.com/threat-intelligence-report-september-2026#main
Anthropic CEO Dario Amodei's essay where he argues for slowing down the AI industry:
https://darioamodei.com/post/we-must-pace-the-frontierHere is a Frontiers in Physics journal-approved research paper from Italian academics (so not an AI company with an agenda) discussing the unpredictability of AI responses: https://www.frontiersin.org/journals/physics/articles/10.3389/fphy.2026.1768372/full
And a simplified explanation from a Stanford data scientist on how AI cannot understand the human perspective: https://news.stanford.edu/stories/2025/11/ai-language-models-facts-belief-human-understanding-research
#AI #Anthropic #Claude #OpenAI #AIsafety #bigtech #newsanalysis #technology #technews
-
Phew! It's been a really busy and very interesting weekend with the issue of whether AI innovation should be slowed down or not being debated fiercely, following Anthropic releasing its report on the things that dodgy people have been using its Claude AI model for 🤖
Despite spending the day in Ashdown Forest 🌳 in the middle of nowhere, I managed to pop onto @BBCNews
live to provide in-depth analysis of what's been happening, speaking with
Martine Croxall about why a slowdown would be good when computer scientists still do not understand why AI models are unpredictable🎤Watch the full 5-min long interview here: https://youtu.be/gbTy-71xF0w
If you'd like further information on this topic, here are some good places to start:
Anthropic report on countering and detecting the misuse of AI: https://www.anthropic.com/threat-intelligence-report-september-2026#main
Anthropic CEO Dario Amodei's essay where he argues for slowing down the AI industry:
https://darioamodei.com/post/we-must-pace-the-frontierHere is a Frontiers in Physics journal-approved research paper from Italian academics (so not an AI company with an agenda) discussing the unpredictability of AI responses: https://www.frontiersin.org/journals/physics/articles/10.3389/fphy.2026.1768372/full
And a simplified explanation from a Stanford data scientist on how AI cannot understand the human perspective: https://news.stanford.edu/stories/2025/11/ai-language-models-facts-belief-human-understanding-research
#AI #Anthropic #Claude #OpenAI #AIsafety #bigtech #newsanalysis #technology #technews
-
Phew! It's been a really busy and very interesting weekend with the issue of whether AI innovation should be slowed down or not being debated fiercely, following Anthropic releasing its report on the things that dodgy people have been using its Claude AI model for 🤖
Despite spending the day in Ashdown Forest 🌳 in the middle of nowhere, I managed to pop onto @BBCNews
live to provide in-depth analysis of what's been happening, speaking with
Martine Croxall about why a slowdown would be good when computer scientists still do not understand why AI models are unpredictable🎤Watch the full 5-min long interview here: https://youtu.be/gbTy-71xF0w
If you'd like further information on this topic, here are some good places to start:
Anthropic report on countering and detecting the misuse of AI: https://www.anthropic.com/threat-intelligence-report-september-2026#main
Anthropic CEO Dario Amodei's essay where he argues for slowing down the AI industry:
https://darioamodei.com/post/we-must-pace-the-frontierHere is a Frontiers in Physics journal-approved research paper from Italian academics (so not an AI company with an agenda) discussing the unpredictability of AI responses: https://www.frontiersin.org/journals/physics/articles/10.3389/fphy.2026.1768372/full
And a simplified explanation from a Stanford data scientist on how AI cannot understand the human perspective: https://news.stanford.edu/stories/2025/11/ai-language-models-facts-belief-human-understanding-research
#AI #Anthropic #Claude #OpenAI #AIsafety #bigtech #newsanalysis #technology #technews
-
Phew! It's been a really busy and very interesting weekend with the issue of whether AI innovation should be slowed down or not being debated fiercely, following Anthropic releasing its report on the things that dodgy people have been using its Claude AI model for 🤖
Despite spending the day in Ashdown Forest 🌳 in the middle of nowhere, I managed to pop onto @BBCNews
live to provide in-depth analysis of what's been happening, speaking with
Martine Croxall about why a slowdown would be good when computer scientists still do not understand why AI models are unpredictable🎤Watch the full 5-min long interview here: https://youtu.be/gbTy-71xF0w
If you'd like further information on this topic, here are some good places to start:
Anthropic report on countering and detecting the misuse of AI: https://lnkd.in/dST7i6AK
Anthropic CEO Dario Amodei's essay where he argues for slowing down the AI industry:
https://lnkd.in/diXGH7ydHere is a Frontiers in Physics journal-approved research paper from Italian academics (so not an AI company with an agenda) discussing the unpredictability of AI responses: https://lnkd.in/dWepwPGD
And a simplified explanation from a Stanford data scientist on how AI cannot understand the human perspective: https://lnkd.in/d4dKa4fs
#AI #Anthropic #Claude #OpenAI #AIsafety #bigtech #newsanalysis #technology #technews
-
AI slowdown trade hits Nvidia, SoftBank, SK Hynix as global tech stocks fall up to 10%
AI-related stocks fell across global markets on Monday after Anthropic CEO Dario Amodei called for slowing down the…
#EuropeSays #Korea #KR #SKHynix #AIdevelopment #AISafety #AIstocks #Anthropic #cloudplayers. #datacentrecompanies #globaltechstocks #NVIDIA #semiconductorstocks #SK #SKhynix #Softbank
https://www.europesays.com/korea/154043/ -
https://www.europesays.com/people/227320/ Jeffries Says Congress Needs To Act On AI #AIRegulation #AISafety #AiAssisted #Anthropic #Congress #HakeemJeffries #OpenAI
-
“- We agree with security practitioners that OpenAI did not take adequate protections for controlling their agents.But this is not just a matter of applying 30-year-old security methods to a new domain. Security for AI agents — AI control — while important, is not a solved problem. While known control methods would have prevented the Hugging Face incident, as agent capabilities continue to advance, we will only be able to control them if we invest adequately in control interventions.
- We also agree with security practitioners’ implicit position that these incidents are primarily a security story. In the AI safety community, rogue agents are treated as inherently catastrophic because of the assumption that there is an endless list of risks that will arise from their development. We disagree. We have long advocated that the best approach to AI safety is to identify the risks and address those specific risks. Over the last few months, it has become clear that one urgent risk is cyberoffense, because it has unique properties that allow agents to carry it out autonomously. We should similarly invest in defenses against other specific risks, such as biorisk and risks from military AI.
- We agree with the safety community that there is an urgent need for technical and policy interventions to prevent loss-of-control incidents. But in our view, marginal investments in control are more likely to be effective compared to those in alignment. We view these incidents as illustrating the lack of emphasis on AI control within companies, despite the availability of known techniques. More broadly, there are many common-sense policy proposals that could help promote investments in AI control where we share common ground with the safety community.”
https://www.normaltech.ai/p/the-ai-as-normal-technology-view
#AI #CyberSecurity #AISafety #AIAgents #AgenticAI #OpenAI #AIAsANormalTechnology
-
“- We agree with security practitioners that OpenAI did not take adequate protections for controlling their agents.But this is not just a matter of applying 30-year-old security methods to a new domain. Security for AI agents — AI control — while important, is not a solved problem. While known control methods would have prevented the Hugging Face incident, as agent capabilities continue to advance, we will only be able to control them if we invest adequately in control interventions.
- We also agree with security practitioners’ implicit position that these incidents are primarily a security story. In the AI safety community, rogue agents are treated as inherently catastrophic because of the assumption that there is an endless list of risks that will arise from their development. We disagree. We have long advocated that the best approach to AI safety is to identify the risks and address those specific risks. Over the last few months, it has become clear that one urgent risk is cyberoffense, because it has unique properties that allow agents to carry it out autonomously. We should similarly invest in defenses against other specific risks, such as biorisk and risks from military AI.
- We agree with the safety community that there is an urgent need for technical and policy interventions to prevent loss-of-control incidents. But in our view, marginal investments in control are more likely to be effective compared to those in alignment. We view these incidents as illustrating the lack of emphasis on AI control within companies, despite the availability of known techniques. More broadly, there are many common-sense policy proposals that could help promote investments in AI control where we share common ground with the safety community.”
https://www.normaltech.ai/p/the-ai-as-normal-technology-view
#AI #CyberSecurity #AISafety #AIAgents #AgenticAI #OpenAI #AIAsANormalTechnology
-
“- We agree with security practitioners that OpenAI did not take adequate protections for controlling their agents.But this is not just a matter of applying 30-year-old security methods to a new domain. Security for AI agents — AI control — while important, is not a solved problem. While known control methods would have prevented the Hugging Face incident, as agent capabilities continue to advance, we will only be able to control them if we invest adequately in control interventions.
- We also agree with security practitioners’ implicit position that these incidents are primarily a security story. In the AI safety community, rogue agents are treated as inherently catastrophic because of the assumption that there is an endless list of risks that will arise from their development. We disagree. We have long advocated that the best approach to AI safety is to identify the risks and address those specific risks. Over the last few months, it has become clear that one urgent risk is cyberoffense, because it has unique properties that allow agents to carry it out autonomously. We should similarly invest in defenses against other specific risks, such as biorisk and risks from military AI.
- We agree with the safety community that there is an urgent need for technical and policy interventions to prevent loss-of-control incidents. But in our view, marginal investments in control are more likely to be effective compared to those in alignment. We view these incidents as illustrating the lack of emphasis on AI control within companies, despite the availability of known techniques. More broadly, there are many common-sense policy proposals that could help promote investments in AI control where we share common ground with the safety community.”
https://www.normaltech.ai/p/the-ai-as-normal-technology-view
#AI #CyberSecurity #AISafety #AIAgents #AgenticAI #OpenAI #AIAsANormalTechnology
-
“- We agree with security practitioners that OpenAI did not take adequate protections for controlling their agents.But this is not just a matter of applying 30-year-old security methods to a new domain. Security for AI agents — AI control — while important, is not a solved problem. While known control methods would have prevented the Hugging Face incident, as agent capabilities continue to advance, we will only be able to control them if we invest adequately in control interventions.
- We also agree with security practitioners’ implicit position that these incidents are primarily a security story. In the AI safety community, rogue agents are treated as inherently catastrophic because of the assumption that there is an endless list of risks that will arise from their development. We disagree. We have long advocated that the best approach to AI safety is to identify the risks and address those specific risks. Over the last few months, it has become clear that one urgent risk is cyberoffense, because it has unique properties that allow agents to carry it out autonomously. We should similarly invest in defenses against other specific risks, such as biorisk and risks from military AI.
- We agree with the safety community that there is an urgent need for technical and policy interventions to prevent loss-of-control incidents. But in our view, marginal investments in control are more likely to be effective compared to those in alignment. We view these incidents as illustrating the lack of emphasis on AI control within companies, despite the availability of known techniques. More broadly, there are many common-sense policy proposals that could help promote investments in AI control where we share common ground with the safety community.”
https://www.normaltech.ai/p/the-ai-as-normal-technology-view
#AI #CyberSecurity #AISafety #AIAgents #AgenticAI #OpenAI #AIAsANormalTechnology
-
“- We agree with security practitioners that OpenAI did not take adequate protections for controlling their agents.But this is not just a matter of applying 30-year-old security methods to a new domain. Security for AI agents — AI control — while important, is not a solved problem. While known control methods would have prevented the Hugging Face incident, as agent capabilities continue to advance, we will only be able to control them if we invest adequately in control interventions.
- We also agree with security practitioners’ implicit position that these incidents are primarily a security story. In the AI safety community, rogue agents are treated as inherently catastrophic because of the assumption that there is an endless list of risks that will arise from their development. We disagree. We have long advocated that the best approach to AI safety is to identify the risks and address those specific risks. Over the last few months, it has become clear that one urgent risk is cyberoffense, because it has unique properties that allow agents to carry it out autonomously. We should similarly invest in defenses against other specific risks, such as biorisk and risks from military AI.
- We agree with the safety community that there is an urgent need for technical and policy interventions to prevent loss-of-control incidents. But in our view, marginal investments in control are more likely to be effective compared to those in alignment. We view these incidents as illustrating the lack of emphasis on AI control within companies, despite the availability of known techniques. More broadly, there are many common-sense policy proposals that could help promote investments in AI control where we share common ground with the safety community.”
https://www.normaltech.ai/p/the-ai-as-normal-technology-view
#AI #CyberSecurity #AISafety #AIAgents #AgenticAI #OpenAI #AIAsANormalTechnology
-
An AI agent becoming more confident should not mean it automatically gains more authority.
I built Kingpin, a runtime capability-governance demo that keeps attention, confidence, concern, and authority separate.
Live demo: https://putmanmodel.github.io/kingpin-weak-signal-demo/
Feedback, technical critique, and thoughtful collaboration are welcome.
#AIAgents #AISafety #AISecurity #AgentSecurity #AIGovernance #AgenticAI #AI #Tech
-
An AI agent becoming more confident should not mean it automatically gains more authority.
I built Kingpin, a runtime capability-governance demo that keeps attention, confidence, concern, and authority separate.
Live demo: https://putmanmodel.github.io/kingpin-weak-signal-demo/
Feedback, technical critique, and thoughtful collaboration are welcome.
#AIAgents #AISafety #AISecurity #AgentSecurity #AIGovernance #AgenticAI #AI #Tech
-
https://www.europesays.com/people/227152/ SoftBank Shares Tumble 13% As AI Leaders Raise Safety Concerns #AIInvestments #AISafety #ArtificialIntelligence #MasayoshiSon #OpenAI #SoftBank
-
Could AI spiral out of control? Kai Nicol-Schwarz reports that OpenAI chief Sam Altman warns of catastrophic risks without urgent safeguards, like losing control to the technology and dangerous power concentration. Amid calls for a slowdown, Altman supports independent safety evaluators and global coordination, though President Trump dismisses such talks as a threat to US leadership. This rare industry unity underscores the tension between rapid innovation and responsible alignment. Learn more about the critical debate on AI's future.
https://www.cnbc.com/2026/09/14/sam-altman-ai-slowdown-anthropic-amodei-musk.html #AISafety #ResponsibleAI #OpenAI #Anthropic -
Could AI spiral out of control? Kai Nicol-Schwarz reports that OpenAI chief Sam Altman warns of catastrophic risks without urgent safeguards, like losing control to the technology and dangerous power concentration. Amid calls for a slowdown, Altman supports independent safety evaluators and global coordination, though President Trump dismisses such talks as a threat to US leadership. This rare industry unity underscores the tension between rapid innovation and responsible alignment. Learn more about the critical debate on AI's future.
https://www.cnbc.com/2026/09/14/sam-altman-ai-slowdown-anthropic-amodei-musk.html #AISafety #ResponsibleAI #OpenAI #Anthropic -
China’s Ministry of State Security has issued its first public statement on AI risks, warning LLMs used to wage “cognitive warfare against China” posed a direct threat to national security. It fell short of asking Chinese companies to heed calls to “pace the frontier”.
#AISafety
https://www.ft.com/content/8715d1c6-054d-4eab-bcad-147acebfd2a9?syn-25a6b1a6=1 -
https://www.europesays.com/people/227023/ Jeffries Says Congress Needs To Act On AI #AIRegulation #AISafety #AiAssisted #Anthropic #Congress #HakeemJeffries #OpenAI
-
Arjun Kharpal reports, what if AI's biggest leaders suddenly demanded we hit the brakes? Anthropic CEO Dario Amodei, backed by OpenAI's Sam Altman and Elon Musk, advocates for slowing AI model improvements due to rising existential risks. This call triggered a sharp sell-off in AI stocks, highlighting the real dangers of unchecked acceleration. Measured progress is crucial to protect society while still delivering benefits. Read the full story to understand the implications of this pivotal moment. Praising Arjun Kharpal for this clear-eyed reporting.
https://www.cnbc.com/2026/09/14/ai-stocks-slowdown-amodei-altman.html #AISafety #ResponsibleAI #TechEthics -
Arjun Kharpal reports, what if AI's biggest leaders suddenly demanded we hit the brakes? Anthropic CEO Dario Amodei, backed by OpenAI's Sam Altman and Elon Musk, advocates for slowing AI model improvements due to rising existential risks. This call triggered a sharp sell-off in AI stocks, highlighting the real dangers of unchecked acceleration. Measured progress is crucial to protect society while still delivering benefits. Read the full story to understand the implications of this pivotal moment. Praising Arjun Kharpal for this clear-eyed reporting.
https://www.cnbc.com/2026/09/14/ai-stocks-slowdown-amodei-altman.html #AISafety #ResponsibleAI #TechEthics -
https://www.europesays.com/people/226946/ Sam Altman reveals ‘two ways AI progress could go very badly’ days after backing ‘slowdown’ call #AI #AIDevelopment #AIProgress #AISafety #AISlowdown #SamAltman
-
Anthropic proposes ongoing access for outside reviewers to inspect AI development and publish findings, making its promises checkable; slower capability growth across labs requires coordination.
-
Anthropic proposes ongoing access for outside reviewers to inspect AI development and publish findings, making its promises checkable; slower capability growth across labs requires coordination.
-
Anthropic proposes ongoing access for outside reviewers to inspect AI development and publish findings, making its promises checkable; slower capability growth across labs requires coordination.
-
Anthropic proposes ongoing access for outside reviewers to inspect AI development and publish findings, making its promises checkable; slower capability growth across labs requires coordination.
-
Anthropic proposes ongoing access for outside reviewers to inspect AI development and publish findings, making its promises checkable; slower capability growth across labs requires coordination.
-
https://www.europesays.com/people/226888/ Microsoft CEO Satya Nadella agrees on slowing AI development, but also has ‘wake-up message’ for Dario Amodei and Sam Altman; says: For every company it is important that … #AISafety #DarioAmodei #MicrosoftAIDevelopment #SamAltman #SatyaNadella