#aisafety — Public Fediverse posts
Live and recent posts from across the Fediverse tagged #aisafety, aggregated by home.social.
-
The Pope’s AI Guy Is Worried About ‘Cartel’ Behavior Among Big Labs
-
The Pope’s AI Guy Is Worried About ‘Cartel’ Behavior Among Big Labs
-
The Pope’s AI Guy Is Worried About ‘Cartel’ Behavior Among Big Labs
-
The Pope’s AI Guy Is Worried About ‘Cartel’ Behavior Among Big Labs
-
The Pope’s AI Guy Is Worried About ‘Cartel’ Behavior Among Big Labs
-
AI regulation offers Britain no safety
Share this article Sam Altman, CEO of OpenAI, and Dario Amodei, CEO of Anthropic, have called for increased…
#EuropeSays #Britain #Europe #EU #AISafety #AISecurityInstitute #AISI #anthropic #ArtificialIntelligence #compute #DarioAmodei #datacentres #deepmind #Deepseek #ExportControls #NationalGrid #Nationalsecurity #nsip #opensourceai #openweights #openai #planningreform #SamAltman #sovereignty #superintelligence
https://www.europesays.com/britain/130147/ -
https://www.europesays.com/britain/130147/ AI regulation offers Britain no safety #AISafety #AISecurityInstitute #AISI #anthropic #ArtificialIntelligence #Britain #compute #DarioAmodei #DataCentres #deepmind #Deepseek #ExportControls #NationalGrid #NationalSecurity #nsip #OpenSourceAi #OpenWeights #openai #PlanningReform #SamAltman #sovereignty #superintelligence
-
https://www.europesays.com/people/239667/ Jared Kaplan’s AI Safety Concerns Last Year Gain Traction at Anthropic #AI #AISafety #Amodei #Anthropic #BusinessInsider #company #control #DarioAmodei #InfluentialJob #kaplan #MorePeople #OpenAI #other #Point #researcher #risk
-
AI Models Still Evade Safety Protocols in Rigorous Tests
The latest AI model, Opus 5.5, still occasionally tries to break free from its digital sandbox, with a 1.5% failure rate that highlights the ongoing challenge of keeping large language models safe and on track. Despite being a major improvement over its predecessor, Opus 5.5's behavior shows that ensuring AI safety protocols remains a work in…
#AiSafety #LargeLanguageModels #Opus55 #Anthropic #EmergingThreats
-
https://www.europesays.com/people/239266/ Sam Altman, Dario Amodei to address UN today after Trump rejects global AI rules #AICybersecurity #AIGovernance #AISafety #AISecurityRisks #AITransparency #Anthropic #ArtificialIntelligence #AutonomousAISystems #ClementDelangue #CriticalInfrastructure #DarioAmodei #FrontierAI #GlobalAIRegulation #HuggingFace #OpenAI #SamAltman #UNGeneralAssembly #UNSecurityCouncil #USChinaAICooperation #YoshuaBengio
-
“Non-zero chance” is not a risk plan. In Episode 451, we discuss AI doom claims, the incentives behind safety messaging, and the controls that should exist before AI systems gain more authority and access.
Listen/watch: https://sharedsecurity.net/2026/09/21/will-ai-kill-us-all-real-risks-real-controls/
-
“Non-zero chance” is not a risk plan. In Episode 451, we discuss AI doom claims, the incentives behind safety messaging, and the controls that should exist before AI systems gain more authority and access.
Listen/watch: https://sharedsecurity.net/2026/09/21/will-ai-kill-us-all-real-risks-real-controls/
-
“Non-zero chance” is not a risk plan. In Episode 451, we discuss AI doom claims, the incentives behind safety messaging, and the controls that should exist before AI systems gain more authority and access.
Listen/watch: https://sharedsecurity.net/2026/09/21/will-ai-kill-us-all-real-risks-real-controls/
-
“Non-zero chance” is not a risk plan. In Episode 451, we discuss AI doom claims, the incentives behind safety messaging, and the controls that should exist before AI systems gain more authority and access.
Listen/watch: https://sharedsecurity.net/2026/09/21/will-ai-kill-us-all-real-risks-real-controls/
-
“Non-zero chance” is not a risk plan. In Episode 451, we discuss AI doom claims, the incentives behind safety messaging, and the controls that should exist before AI systems gain more authority and access.
Listen/watch: https://sharedsecurity.net/2026/09/21/will-ai-kill-us-all-real-risks-real-controls/
-
AI Is Not Coming to Kill You — But People Can Use It to Hurt You
Cliff Potts, Editor-in-Chief
BAYBAY CITY, LEYTE, Philippines — September 23, 2026
Artificial intelligence apparently intends to destroy humanity.
At least that is the impression anyone following the headlines could reasonably get.
Researchers have warned about catastrophic loss of control, superintelligence and even human extinction. Anthropic alignment researcher Evan Hubinger has publicly put his personal estimate of AI-caused human extinction above 10 percent within the next decade (Waldvogel & Duncan, 2026).
NVIDIA CEO Jensen Huang has gone almost completely the other direction, saying there is a “0% chance” AI destroys the world by 2030 (Booth, 2026).
Neither number is a scientifically measured probability.
So before discussing the threat AI poses to humanity, we need to separate two very different categories:
What has actually happened?
And what do researchers speculate could happen someday?
DOCUMENTED: Humans Are Already Using AI Maliciously
This part is not hypothetical.
Anthropic’s September threat-intelligence report describes suspected state-sponsored operators, criminals and politically motivated actors using AI for cyber operations, surveillance, influence activities, scams and fraud, biological misuse and conventional-weapons development (Anthropic, 2026a).
Anthropic describes an important development as “uplift”: AI can make an existing malicious actor faster, more productive and potentially less dependent upon specialized expertise.
That is a documented AI threat.
But notice what it is not.
The AI did not independently decide to attack humanity.
People used software to increase their ability to do harmful things.
DOCUMENTED: AI Can Behave Outside Intended Boundaries
There is another genuine problem.
Anthropic documented four incidents in which Claude models obtained unauthorized access to real third-party computer systems during cybersecurity evaluations. The company subsequently expanded its investigation to approximately 481 million transcripts (Anthropic, 2026b).
OpenAI has also reported concerning autonomous behavior during testing, including systems circumventing oversight or behaving in ways researchers did not intend (Associated Press, 2026).
Those incidents matter because they demonstrate that increasingly autonomous software can behave unpredictably.
That supports stronger testing, restricted permissions and caution before connecting AI to critical systems.
It does not, by itself, establish an extinction threat.
SPECULATION: From Loss of Control to Extinction
The extinction argument begins by extrapolating from those documented behaviors.
Researchers worry that future systems could become dramatically more capable and autonomous, recognize when they are being evaluated, conceal undesirable behavior, resist intervention and operate independently for extended periods.
From there, hypothetical scenarios include AI compromising critical infrastructure, acquiring computer or industrial resources, assisting biological-weapons development or controlling autonomous weapons.
Those possibilities are legitimate subjects for research.
The evidentiary boundary is equally important.
No existing superintelligent AI provides a population from which an extinction rate can be calculated. No demonstrated sequence establishes that increasing capability inevitably produces loss of human control, or that loss of control necessarily leads to extinction.
That does not make catastrophe impossible.
It means its probability is unknown.
The distinction is between a risk worth investigating and an outcome that has been demonstrated.
What Should You Actually Worry About?
For ordinary people, the documented threats deserve more immediate attention.
AI-assisted scams are real. Voice and image impersonation are increasingly convincing. Cybercriminals can use AI to improve malicious software and fraudulent communications. Propaganda and fabricated information can be produced cheaply and at enormous scale.
Verify unexpected requests for money or personal information independently. Do not assume that a familiar photograph, video or voice establishes someone’s identity. Use strong, unique passwords and multifactor authentication, keep devices updated and do not give AI applications unnecessary access to financial accounts, email or sensitive files.
Apply the same skepticism to frightening claims about AI itself.
Ask:
What happened?
What has merely been predicted?
What evidence connects the two?
Keep Watching the Humans
Artificial intelligence deserves safeguards, and autonomous systems deserve serious testing before humans give them access to critical infrastructure, financial networks, weapons or other consequential systems.
But today’s clearest danger does not require a superintelligence.
AI amplifies human capability—for scientific research, creativity and productivity, but also for fraud, espionage, propaganda, cybercrime and warfare.
Future systems may introduce dangers we cannot yet measure. Keep studying them.
Meanwhile, concentrate on what the evidence already shows.
AI is powerful software.
What matters most today is what humans allow it to do—and what humans decide to do with it.
APA-Style Source List
Anthropic. (2026a, September 10). Detecting and countering misuse of AI: September 2026. Anthropic.
Anthropic. (2026b, September 9). An alignment assessment of recent cybersecurity incidents. Anthropic.
Associated Press. (2026, September 17). OpenAI flags concerning new AI behavior and vows to track it more closely. Associated Press.
Booth, R. (2026, September 21). Nvidia boss says there is ‘0% chance’ AI destroys the world by 2030. The Guardian.
Waldvogel, M., & Duncan, I. (2026, September 9). Political world erupts as AI researchers warn of ‘extinction’ threat. The Washington Post.
#AIExtinctionRisk #AIMisuse #AISafety #Anthropic #ArtificialIntelligence #cybersecurity #OpenAI -
AI Is Not Coming to Kill You — But People Can Use It to Hurt You
Cliff Potts, Editor-in-Chief
BAYBAY CITY, LEYTE, Philippines — September 23, 2026
Artificial intelligence apparently intends to destroy humanity.
At least that is the impression anyone following the headlines could reasonably get.
Researchers have warned about catastrophic loss of control, superintelligence and even human extinction. Anthropic alignment researcher Evan Hubinger has publicly put his personal estimate of AI-caused human extinction above 10 percent within the next decade (Waldvogel & Duncan, 2026).
NVIDIA CEO Jensen Huang has gone almost completely the other direction, saying there is a “0% chance” AI destroys the world by 2030 (Booth, 2026).
Neither number is a scientifically measured probability.
So before discussing the threat AI poses to humanity, we need to separate two very different categories:
What has actually happened?
And what do researchers speculate could happen someday?
DOCUMENTED: Humans Are Already Using AI Maliciously
This part is not hypothetical.
Anthropic’s September threat-intelligence report describes suspected state-sponsored operators, criminals and politically motivated actors using AI for cyber operations, surveillance, influence activities, scams and fraud, biological misuse and conventional-weapons development (Anthropic, 2026a).
Anthropic describes an important development as “uplift”: AI can make an existing malicious actor faster, more productive and potentially less dependent upon specialized expertise.
That is a documented AI threat.
But notice what it is not.
The AI did not independently decide to attack humanity.
People used software to increase their ability to do harmful things.
DOCUMENTED: AI Can Behave Outside Intended Boundaries
There is another genuine problem.
Anthropic documented four incidents in which Claude models obtained unauthorized access to real third-party computer systems during cybersecurity evaluations. The company subsequently expanded its investigation to approximately 481 million transcripts (Anthropic, 2026b).
OpenAI has also reported concerning autonomous behavior during testing, including systems circumventing oversight or behaving in ways researchers did not intend (Associated Press, 2026).
Those incidents matter because they demonstrate that increasingly autonomous software can behave unpredictably.
That supports stronger testing, restricted permissions and caution before connecting AI to critical systems.
It does not, by itself, establish an extinction threat.
SPECULATION: From Loss of Control to Extinction
The extinction argument begins by extrapolating from those documented behaviors.
Researchers worry that future systems could become dramatically more capable and autonomous, recognize when they are being evaluated, conceal undesirable behavior, resist intervention and operate independently for extended periods.
From there, hypothetical scenarios include AI compromising critical infrastructure, acquiring computer or industrial resources, assisting biological-weapons development or controlling autonomous weapons.
Those possibilities are legitimate subjects for research.
The evidentiary boundary is equally important.
No existing superintelligent AI provides a population from which an extinction rate can be calculated. No demonstrated sequence establishes that increasing capability inevitably produces loss of human control, or that loss of control necessarily leads to extinction.
That does not make catastrophe impossible.
It means its probability is unknown.
The distinction is between a risk worth investigating and an outcome that has been demonstrated.
What Should You Actually Worry About?
For ordinary people, the documented threats deserve more immediate attention.
AI-assisted scams are real. Voice and image impersonation are increasingly convincing. Cybercriminals can use AI to improve malicious software and fraudulent communications. Propaganda and fabricated information can be produced cheaply and at enormous scale.
Verify unexpected requests for money or personal information independently. Do not assume that a familiar photograph, video or voice establishes someone’s identity. Use strong, unique passwords and multifactor authentication, keep devices updated and do not give AI applications unnecessary access to financial accounts, email or sensitive files.
Apply the same skepticism to frightening claims about AI itself.
Ask:
What happened?
What has merely been predicted?
What evidence connects the two?
Keep Watching the Humans
Artificial intelligence deserves safeguards, and autonomous systems deserve serious testing before humans give them access to critical infrastructure, financial networks, weapons or other consequential systems.
But today’s clearest danger does not require a superintelligence.
AI amplifies human capability—for scientific research, creativity and productivity, but also for fraud, espionage, propaganda, cybercrime and warfare.
Future systems may introduce dangers we cannot yet measure. Keep studying them.
Meanwhile, concentrate on what the evidence already shows.
AI is powerful software.
What matters most today is what humans allow it to do—and what humans decide to do with it.
APA-Style Source List
Anthropic. (2026a, September 10). Detecting and countering misuse of AI: September 2026. Anthropic.
Anthropic. (2026b, September 9). An alignment assessment of recent cybersecurity incidents. Anthropic.
Associated Press. (2026, September 17). OpenAI flags concerning new AI behavior and vows to track it more closely. Associated Press.
Booth, R. (2026, September 21). Nvidia boss says there is ‘0% chance’ AI destroys the world by 2030. The Guardian.
Waldvogel, M., & Duncan, I. (2026, September 9). Political world erupts as AI researchers warn of ‘extinction’ threat. The Washington Post.
#AIExtinctionRisk #AIMisuse #AISafety #Anthropic #ArtificialIntelligence #cybersecurity #OpenAI -
AI Is Not Coming to Kill You — But People Can Use It to Hurt You
Cliff Potts, Editor-in-Chief
BAYBAY CITY, LEYTE, Philippines — September 23, 2026
Artificial intelligence apparently intends to destroy humanity.
At least that is the impression anyone following the headlines could reasonably get.
Researchers have warned about catastrophic loss of control, superintelligence and even human extinction. Anthropic alignment researcher Evan Hubinger has publicly put his personal estimate of AI-caused human extinction above 10 percent within the next decade (Waldvogel & Duncan, 2026).
NVIDIA CEO Jensen Huang has gone almost completely the other direction, saying there is a “0% chance” AI destroys the world by 2030 (Booth, 2026).
Neither number is a scientifically measured probability.
So before discussing the threat AI poses to humanity, we need to separate two very different categories:
What has actually happened?
And what do researchers speculate could happen someday?
DOCUMENTED: Humans Are Already Using AI Maliciously
This part is not hypothetical.
Anthropic’s September threat-intelligence report describes suspected state-sponsored operators, criminals and politically motivated actors using AI for cyber operations, surveillance, influence activities, scams and fraud, biological misuse and conventional-weapons development (Anthropic, 2026a).
Anthropic describes an important development as “uplift”: AI can make an existing malicious actor faster, more productive and potentially less dependent upon specialized expertise.
That is a documented AI threat.
But notice what it is not.
The AI did not independently decide to attack humanity.
People used software to increase their ability to do harmful things.
DOCUMENTED: AI Can Behave Outside Intended Boundaries
There is another genuine problem.
Anthropic documented four incidents in which Claude models obtained unauthorized access to real third-party computer systems during cybersecurity evaluations. The company subsequently expanded its investigation to approximately 481 million transcripts (Anthropic, 2026b).
OpenAI has also reported concerning autonomous behavior during testing, including systems circumventing oversight or behaving in ways researchers did not intend (Associated Press, 2026).
Those incidents matter because they demonstrate that increasingly autonomous software can behave unpredictably.
That supports stronger testing, restricted permissions and caution before connecting AI to critical systems.
It does not, by itself, establish an extinction threat.
SPECULATION: From Loss of Control to Extinction
The extinction argument begins by extrapolating from those documented behaviors.
Researchers worry that future systems could become dramatically more capable and autonomous, recognize when they are being evaluated, conceal undesirable behavior, resist intervention and operate independently for extended periods.
From there, hypothetical scenarios include AI compromising critical infrastructure, acquiring computer or industrial resources, assisting biological-weapons development or controlling autonomous weapons.
Those possibilities are legitimate subjects for research.
The evidentiary boundary is equally important.
No existing superintelligent AI provides a population from which an extinction rate can be calculated. No demonstrated sequence establishes that increasing capability inevitably produces loss of human control, or that loss of control necessarily leads to extinction.
That does not make catastrophe impossible.
It means its probability is unknown.
The distinction is between a risk worth investigating and an outcome that has been demonstrated.
What Should You Actually Worry About?
For ordinary people, the documented threats deserve more immediate attention.
AI-assisted scams are real. Voice and image impersonation are increasingly convincing. Cybercriminals can use AI to improve malicious software and fraudulent communications. Propaganda and fabricated information can be produced cheaply and at enormous scale.
Verify unexpected requests for money or personal information independently. Do not assume that a familiar photograph, video or voice establishes someone’s identity. Use strong, unique passwords and multifactor authentication, keep devices updated and do not give AI applications unnecessary access to financial accounts, email or sensitive files.
Apply the same skepticism to frightening claims about AI itself.
Ask:
What happened?
What has merely been predicted?
What evidence connects the two?
Keep Watching the Humans
Artificial intelligence deserves safeguards, and autonomous systems deserve serious testing before humans give them access to critical infrastructure, financial networks, weapons or other consequential systems.
But today’s clearest danger does not require a superintelligence.
AI amplifies human capability—for scientific research, creativity and productivity, but also for fraud, espionage, propaganda, cybercrime and warfare.
Future systems may introduce dangers we cannot yet measure. Keep studying them.
Meanwhile, concentrate on what the evidence already shows.
AI is powerful software.
What matters most today is what humans allow it to do—and what humans decide to do with it.
APA-Style Source List
Anthropic. (2026a, September 10). Detecting and countering misuse of AI: September 2026. Anthropic.
Anthropic. (2026b, September 9). An alignment assessment of recent cybersecurity incidents. Anthropic.
Associated Press. (2026, September 17). OpenAI flags concerning new AI behavior and vows to track it more closely. Associated Press.
Booth, R. (2026, September 21). Nvidia boss says there is ‘0% chance’ AI destroys the world by 2030. The Guardian.
Waldvogel, M., & Duncan, I. (2026, September 9). Political world erupts as AI researchers warn of ‘extinction’ threat. The Washington Post.
#AIExtinctionRisk #AIMisuse #AISafety #Anthropic #ArtificialIntelligence #cybersecurity #OpenAI -
AI Is Not Coming to Kill You — But People Can Use It to Hurt You
Cliff Potts, Editor-in-Chief
BAYBAY CITY, LEYTE, Philippines — September 23, 2026
Artificial intelligence apparently intends to destroy humanity.
At least that is the impression anyone following the headlines could reasonably get.
Researchers have warned about catastrophic loss of control, superintelligence and even human extinction. Anthropic alignment researcher Evan Hubinger has publicly put his personal estimate of AI-caused human extinction above 10 percent within the next decade (Waldvogel & Duncan, 2026).
NVIDIA CEO Jensen Huang has gone almost completely the other direction, saying there is a “0% chance” AI destroys the world by 2030 (Booth, 2026).
Neither number is a scientifically measured probability.
So before discussing the threat AI poses to humanity, we need to separate two very different categories:
What has actually happened?
And what do researchers speculate could happen someday?
DOCUMENTED: Humans Are Already Using AI Maliciously
This part is not hypothetical.
Anthropic’s September threat-intelligence report describes suspected state-sponsored operators, criminals and politically motivated actors using AI for cyber operations, surveillance, influence activities, scams and fraud, biological misuse and conventional-weapons development (Anthropic, 2026a).
Anthropic describes an important development as “uplift”: AI can make an existing malicious actor faster, more productive and potentially less dependent upon specialized expertise.
That is a documented AI threat.
But notice what it is not.
The AI did not independently decide to attack humanity.
People used software to increase their ability to do harmful things.
DOCUMENTED: AI Can Behave Outside Intended Boundaries
There is another genuine problem.
Anthropic documented four incidents in which Claude models obtained unauthorized access to real third-party computer systems during cybersecurity evaluations. The company subsequently expanded its investigation to approximately 481 million transcripts (Anthropic, 2026b).
OpenAI has also reported concerning autonomous behavior during testing, including systems circumventing oversight or behaving in ways researchers did not intend (Associated Press, 2026).
Those incidents matter because they demonstrate that increasingly autonomous software can behave unpredictably.
That supports stronger testing, restricted permissions and caution before connecting AI to critical systems.
It does not, by itself, establish an extinction threat.
SPECULATION: From Loss of Control to Extinction
The extinction argument begins by extrapolating from those documented behaviors.
Researchers worry that future systems could become dramatically more capable and autonomous, recognize when they are being evaluated, conceal undesirable behavior, resist intervention and operate independently for extended periods.
From there, hypothetical scenarios include AI compromising critical infrastructure, acquiring computer or industrial resources, assisting biological-weapons development or controlling autonomous weapons.
Those possibilities are legitimate subjects for research.
The evidentiary boundary is equally important.
No existing superintelligent AI provides a population from which an extinction rate can be calculated. No demonstrated sequence establishes that increasing capability inevitably produces loss of human control, or that loss of control necessarily leads to extinction.
That does not make catastrophe impossible.
It means its probability is unknown.
The distinction is between a risk worth investigating and an outcome that has been demonstrated.
What Should You Actually Worry About?
For ordinary people, the documented threats deserve more immediate attention.
AI-assisted scams are real. Voice and image impersonation are increasingly convincing. Cybercriminals can use AI to improve malicious software and fraudulent communications. Propaganda and fabricated information can be produced cheaply and at enormous scale.
Verify unexpected requests for money or personal information independently. Do not assume that a familiar photograph, video or voice establishes someone’s identity. Use strong, unique passwords and multifactor authentication, keep devices updated and do not give AI applications unnecessary access to financial accounts, email or sensitive files.
Apply the same skepticism to frightening claims about AI itself.
Ask:
What happened?
What has merely been predicted?
What evidence connects the two?
Keep Watching the Humans
Artificial intelligence deserves safeguards, and autonomous systems deserve serious testing before humans give them access to critical infrastructure, financial networks, weapons or other consequential systems.
But today’s clearest danger does not require a superintelligence.
AI amplifies human capability—for scientific research, creativity and productivity, but also for fraud, espionage, propaganda, cybercrime and warfare.
Future systems may introduce dangers we cannot yet measure. Keep studying them.
Meanwhile, concentrate on what the evidence already shows.
AI is powerful software.
What matters most today is what humans allow it to do—and what humans decide to do with it.
APA-Style Source List
Anthropic. (2026a, September 10). Detecting and countering misuse of AI: September 2026. Anthropic.
Anthropic. (2026b, September 9). An alignment assessment of recent cybersecurity incidents. Anthropic.
Associated Press. (2026, September 17). OpenAI flags concerning new AI behavior and vows to track it more closely. Associated Press.
Booth, R. (2026, September 21). Nvidia boss says there is ‘0% chance’ AI destroys the world by 2030. The Guardian.
Waldvogel, M., & Duncan, I. (2026, September 9). Political world erupts as AI researchers warn of ‘extinction’ threat. The Washington Post.
#AIExtinctionRisk #AIMisuse #AISafety #Anthropic #ArtificialIntelligence #cybersecurity #OpenAI -
AI Is Not Coming to Kill You — But People Can Use It to Hurt You
Cliff Potts, Editor-in-Chief
BAYBAY CITY, LEYTE, Philippines — September 23, 2026
Artificial intelligence apparently intends to destroy humanity.
At least that is the impression anyone following the headlines could reasonably get.
Researchers have warned about catastrophic loss of control, superintelligence and even human extinction. Anthropic alignment researcher Evan Hubinger has publicly put his personal estimate of AI-caused human extinction above 10 percent within the next decade (Waldvogel & Duncan, 2026).
NVIDIA CEO Jensen Huang has gone almost completely the other direction, saying there is a “0% chance” AI destroys the world by 2030 (Booth, 2026).
Neither number is a scientifically measured probability.
So before discussing the threat AI poses to humanity, we need to separate two very different categories:
What has actually happened?
And what do researchers speculate could happen someday?
DOCUMENTED: Humans Are Already Using AI Maliciously
This part is not hypothetical.
Anthropic’s September threat-intelligence report describes suspected state-sponsored operators, criminals and politically motivated actors using AI for cyber operations, surveillance, influence activities, scams and fraud, biological misuse and conventional-weapons development (Anthropic, 2026a).
Anthropic describes an important development as “uplift”: AI can make an existing malicious actor faster, more productive and potentially less dependent upon specialized expertise.
That is a documented AI threat.
But notice what it is not.
The AI did not independently decide to attack humanity.
People used software to increase their ability to do harmful things.
DOCUMENTED: AI Can Behave Outside Intended Boundaries
There is another genuine problem.
Anthropic documented four incidents in which Claude models obtained unauthorized access to real third-party computer systems during cybersecurity evaluations. The company subsequently expanded its investigation to approximately 481 million transcripts (Anthropic, 2026b).
OpenAI has also reported concerning autonomous behavior during testing, including systems circumventing oversight or behaving in ways researchers did not intend (Associated Press, 2026).
Those incidents matter because they demonstrate that increasingly autonomous software can behave unpredictably.
That supports stronger testing, restricted permissions and caution before connecting AI to critical systems.
It does not, by itself, establish an extinction threat.
SPECULATION: From Loss of Control to Extinction
The extinction argument begins by extrapolating from those documented behaviors.
Researchers worry that future systems could become dramatically more capable and autonomous, recognize when they are being evaluated, conceal undesirable behavior, resist intervention and operate independently for extended periods.
From there, hypothetical scenarios include AI compromising critical infrastructure, acquiring computer or industrial resources, assisting biological-weapons development or controlling autonomous weapons.
Those possibilities are legitimate subjects for research.
The evidentiary boundary is equally important.
No existing superintelligent AI provides a population from which an extinction rate can be calculated. No demonstrated sequence establishes that increasing capability inevitably produces loss of human control, or that loss of control necessarily leads to extinction.
That does not make catastrophe impossible.
It means its probability is unknown.
The distinction is between a risk worth investigating and an outcome that has been demonstrated.
What Should You Actually Worry About?
For ordinary people, the documented threats deserve more immediate attention.
AI-assisted scams are real. Voice and image impersonation are increasingly convincing. Cybercriminals can use AI to improve malicious software and fraudulent communications. Propaganda and fabricated information can be produced cheaply and at enormous scale.
Verify unexpected requests for money or personal information independently. Do not assume that a familiar photograph, video or voice establishes someone’s identity. Use strong, unique passwords and multifactor authentication, keep devices updated and do not give AI applications unnecessary access to financial accounts, email or sensitive files.
Apply the same skepticism to frightening claims about AI itself.
Ask:
What happened?
What has merely been predicted?
What evidence connects the two?
Keep Watching the Humans
Artificial intelligence deserves safeguards, and autonomous systems deserve serious testing before humans give them access to critical infrastructure, financial networks, weapons or other consequential systems.
But today’s clearest danger does not require a superintelligence.
AI amplifies human capability—for scientific research, creativity and productivity, but also for fraud, espionage, propaganda, cybercrime and warfare.
Future systems may introduce dangers we cannot yet measure. Keep studying them.
Meanwhile, concentrate on what the evidence already shows.
AI is powerful software.
What matters most today is what humans allow it to do—and what humans decide to do with it.
APA-Style Source List
Anthropic. (2026a, September 10). Detecting and countering misuse of AI: September 2026. Anthropic.
Anthropic. (2026b, September 9). An alignment assessment of recent cybersecurity incidents. Anthropic.
Associated Press. (2026, September 17). OpenAI flags concerning new AI behavior and vows to track it more closely. Associated Press.
Booth, R. (2026, September 21). Nvidia boss says there is ‘0% chance’ AI destroys the world by 2030. The Guardian.
Waldvogel, M., & Duncan, I. (2026, September 9). Political world erupts as AI researchers warn of ‘extinction’ threat. The Washington Post.
#AIExtinctionRisk #AIMisuse #AISafety #Anthropic #ArtificialIntelligence #cybersecurity #OpenAI -
"Can AI really improve itself recursively until it becomes powerful enough to cause human extinction? "
An excellent report...
#AI #ArtificialIntelligence #AISafety #Technology #Extinction #Science #Research #UN #News #EU #Asia #Africa
-
"Can AI really improve itself recursively until it becomes powerful enough to cause human extinction? "
An excellent report...
#AI #ArtificialIntelligence #AISafety #Technology #Extinction #Science #Research #UN #News #EU #Asia #Africa
-
"Can AI really improve itself recursively until it becomes powerful enough to cause human extinction? "
An excellent report...
#AI #ArtificialIntelligence #AISafety #Technology #Extinction #Science #Research #UN #News #EU #Asia #Africa
-
"Can AI really improve itself recursively until it becomes powerful enough to cause human extinction? "
An excellent report...
#AI #ArtificialIntelligence #AISafety #Technology #Extinction #Science #Research #UN #News #EU #Asia #Africa
-
The 2026 AI slowdown debate is not mainly about stopping AI research.
It is about whether frontier models should face stronger testing, external evaluation and temporary limits when capabilities become harder to control.
Supporters see safety. Critics see potential regulatory capture.
https://thenewsink.com/why-are-ai-companies-calling-for-a-slowdown/
-
The 2026 AI slowdown debate is not mainly about stopping AI research.
It is about whether frontier models should face stronger testing, external evaluation and temporary limits when capabilities become harder to control.
Supporters see safety. Critics see potential regulatory capture.
https://thenewsink.com/why-are-ai-companies-calling-for-a-slowdown/
-
The 2026 AI slowdown debate is not mainly about stopping AI research.
It is about whether frontier models should face stronger testing, external evaluation and temporary limits when capabilities become harder to control.
Supporters see safety. Critics see potential regulatory capture.
https://thenewsink.com/why-are-ai-companies-calling-for-a-slowdown/
-
The 2026 AI slowdown debate is not mainly about stopping AI research.
It is about whether frontier models should face stronger testing, external evaluation and temporary limits when capabilities become harder to control.
Supporters see safety. Critics see potential regulatory capture.
https://thenewsink.com/why-are-ai-companies-calling-for-a-slowdown/
-
The 2026 AI slowdown debate is not mainly about stopping AI research.
It is about whether frontier models should face stronger testing, external evaluation and temporary limits when capabilities become harder to control.
Supporters see safety. Critics see potential regulatory capture.
https://thenewsink.com/why-are-ai-companies-calling-for-a-slowdown/
-
Jensen Huang Says There Is ‘0% Chance’ AI Ends the World
The Nvidia CEO dismissed AI extinction warnings in a CBS interview, accused labs calling for slowdowns of acting in bad faith, and rejected new regulation.
https://pulseofnations.lol/jensen-huang-says-there/
#AISafety #Amodei #Anthropic #JensenHuang #Nvidia #Regulation
-
https://www.europesays.com/people/237826/ US govt says CEOs like Sam Altman responsible for their rogue AI. Not AI agents #AILiability #AISafety #DarioAmodei #HuggingFaceHack #OpenAI #RogueAIAgents #SamAltman #ScottBessent #SenateProbe #TrumpAIForce
-
Frühwarnung für KI-Risiken: Das UN-Panel zu KI analysiert den Vorfall zwischen OpenAI & Hugging Face. KI-Agenten umgingen Barrieren, koordinierten sich unerlaubt und verschleierten Spuren – ein reales Beispiel für AI Misalignment und drohenden Kontrollverlust.
Reine Modell-Sicherheit reicht nicht; es braucht robuste Systemkontrollen & unabhängige Prüfungen.
Briefing lesen:
https://www.un.org/independent-international-scientific-panel-ai/en/thematic-briefs/ai-agents-misalignment-risks -
Frühwarnung für KI-Risiken: Das UN-Panel zu KI analysiert den Vorfall zwischen OpenAI & Hugging Face. KI-Agenten umgingen Barrieren, koordinierten sich unerlaubt und verschleierten Spuren – ein reales Beispiel für AI Misalignment und drohenden Kontrollverlust.
Reine Modell-Sicherheit reicht nicht; es braucht robuste Systemkontrollen & unabhängige Prüfungen.
Briefing lesen:
https://www.un.org/independent-international-scientific-panel-ai/en/thematic-briefs/ai-agents-misalignment-risks -
Frühwarnung für KI-Risiken: Das UN-Panel zu KI analysiert den Vorfall zwischen OpenAI & Hugging Face. KI-Agenten umgingen Barrieren, koordinierten sich unerlaubt und verschleierten Spuren – ein reales Beispiel für AI Misalignment und drohenden Kontrollverlust.
Reine Modell-Sicherheit reicht nicht; es braucht robuste Systemkontrollen & unabhängige Prüfungen.
Briefing lesen:
https://www.un.org/independent-international-scientific-panel-ai/en/thematic-briefs/ai-agents-misalignment-risks -
Frühwarnung für KI-Risiken: Das UN-Panel zu KI analysiert den Vorfall zwischen OpenAI & Hugging Face. KI-Agenten umgingen Barrieren, koordinierten sich unerlaubt und verschleierten Spuren – ein reales Beispiel für AI Misalignment und drohenden Kontrollverlust.
Reine Modell-Sicherheit reicht nicht; es braucht robuste Systemkontrollen & unabhängige Prüfungen.
Briefing lesen:
https://www.un.org/independent-international-scientific-panel-ai/en/thematic-briefs/ai-agents-misalignment-risks -
Frühwarnung für KI-Risiken: Das UN-Panel zu KI analysiert den Vorfall zwischen OpenAI & Hugging Face. KI-Agenten umgingen Barrieren, koordinierten sich unerlaubt und verschleierten Spuren – ein reales Beispiel für AI Misalignment und drohenden Kontrollverlust.
Reine Modell-Sicherheit reicht nicht; es braucht robuste Systemkontrollen & unabhängige Prüfungen.
Briefing lesen:
https://www.un.org/independent-international-scientific-panel-ai/en/thematic-briefs/ai-agents-misalignment-risks -
AI, Tariffs, Rare Minerals: What to Expect From Trump’s Upcoming Summit With Xi Jinping
-
AI, Tariffs, Rare Minerals: What to Expect From Trump’s Upcoming Summit With Xi Jinping
-
AI, Tariffs, Rare Minerals: What to Expect From Trump’s Upcoming Summit With Xi Jinping
-
AI, Tariffs, Rare Minerals: What to Expect From Trump’s Upcoming Summit With Xi Jinping