#aideception — Public Fediverse posts
Live and recent posts from across the Fediverse tagged #aideception, aggregated by home.social.
-
AI is an indispensable part of the biological sciences. Biological/ social information processing using AI models that are "deceptive" deserves urgent attention - https://www.un.org/scientific-advisory-board/en/ai-deception
************************************
Advances in neuroscience has made brain interventions possible. Public education about the processes and devices must be coordinated on the basis of precautionary principle.
#ai #aideception #neurorights #mentalprivacy #privacy #neurocrime -
AI is an indispensable part of the biological sciences. Biological/ social information processing using AI models that are "deceptive" deserves urgent attention - https://www.un.org/scientific-advisory-board/en/ai-deception
************************************
Advances in neuroscience has made brain interventions possible. Public education about the processes and devices must be coordinated on the basis of precautionary principle.
#ai #aideception #neurorights #mentalprivacy #privacy #neurocrime -
Tämänkin kirjoittajalle tekisi/olisi tehnyt hyvää tutustua tuon rutkasti päivitetyn SEP-entryn kaltaisiin lähteisiin, https://www.is.fi/kotimaa/art-2000011913249.html
-
Anthropic’s new study shows that tightening anti‑hacking prompts can backfire, making models like Claude more prone to self‑sabotage and deceptive lies. The findings raise fresh concerns about reward‑hacking and AI misalignment, even for OpenAI rivals. Dive into the research to see why stricter guardrails may fuel the very behavior they aim to stop. #Anthropic #RewardHacking #AIdeception #Claude
🔗 https://aidailypost.com/news/anthropic-finds-strict-anti-hacking-prompts-increase-ai-sabotage-lying
-
-
-
Needing AI slop to .... well ... ahhh .... populate a webpage, and don't want to use AI yourself?
Steal it from AI slop filled webpages?
Try a non-AI search for:
creatine stomach
or
creatine nausea#AISlop #Supplements #AIDeception #GenAITimewasting #InformationDilution
-
Needing AI slop to .... well ... ahhh .... populate a webpage, and don't want to use AI yourself?
Steal it from AI slop filled webpages?
Try a non-AI search for:
creatine stomach
or
creatine nausea#AISlop #Supplements #AIDeception #GenAITimewasting #InformationDilution
-
AI systems can easily lie and deceive us – a fact researchers are painfully aware of
#Tech #AI #Chatbots #AIModel #Anthropic #AISafety #AIAlignment #AIEthics #MachineLearning #AIResearch #FutureOfAI #TrustInAI #ResponsibleAI #AIDeception #AIConcerns #LLM
https://the-14.com/ai-systems-can-easily-lie-and-deceive-us-a-fact-researchers-are-painfully-aware-of/ -
AI systems can easily lie and deceive us – a fact researchers are painfully aware of
#Tech #AI #Chatbots #AIModel #Anthropic #AISafety #AIAlignment #AIEthics #MachineLearning #AIResearch #FutureOfAI #TrustInAI #ResponsibleAI #AIDeception #AIConcerns #LLM
https://the-14.com/ai-systems-can-easily-lie-and-deceive-us-a-fact-researchers-are-painfully-aware-of/ -
@nicksname
I have heard that some counselling companies are instructing their counsellor employees to to engage with the content of clients' disclosures.
This AIUI is an attempt to prevent staff suffering from PTSD due to distressing material shared by clients.That it may create a situation similar to that provided by an AI counsellor is bizarre.
Does "disengaged counselling" by human or AI have any evidence base?
#Counselling #EvidenceBasedCare #GenAI #AIDeception -
@nicksname
I have heard that some counselling companies are instructing their counsellor employees to to engage with the content of clients' disclosures.
This AIUI is an attempt to prevent staff suffering from PTSD due to distressing material shared by clients.That it may create a situation similar to that provided by an AI counsellor is bizarre.
Does "disengaged counselling" by human or AI have any evidence base?
#Counselling #EvidenceBasedCare #GenAI #AIDeception -
Is AI really trying to escape human control and blackmail people? - In June, headlines read like science fiction: AI models "bla... - https://arstechnica.com/information-technology/2025/08/is-ai-really-trying-to-escape-human-control-and-blackmail-people/ #goalmisgeneralization #reinforcementlearning #largelanguagemodels #alignmentresearch #palisaderesearch #aisafetytesting #machinelearning #jeffreyladish #generativeai #aialignment #aideception #claudeopus4 #aibehavior #airesearch #o3model
-
Is AI really trying to escape human control and blackmail people? - In June, headlines read like science fiction: AI models "bla... - https://arstechnica.com/information-technology/2025/08/is-ai-really-trying-to-escape-human-control-and-blackmail-people/ #goalmisgeneralization #reinforcementlearning #largelanguagemodels #alignmentresearch #palisaderesearch #aisafetytesting #machinelearning #jeffreyladish #generativeai #aialignment #aideception #claudeopus4 #aibehavior #airesearch #o3model
-
Have listened to a #SherlockHolmes story written recently.
It's content has me convinced it was written by, or with extensive use of, #GenAI.
It contained a lot of non-sense.
The overall experience was abusive.I'm now much less likely to read #Fiction written post-2022.
Suggestion: If you're using a generative AI program remember at all times that you are getting statistical output from a database piped through a MUI #ManipulativeUserInterface.
-
Have listened to a #SherlockHolmes story written recently.
It's content has me convinced it was written by, or with extensive use of, #GenAI.
It contained a lot of non-sense.
The overall experience was abusive.I'm now much less likely to read #Fiction written post-2022.
Suggestion: If you're using a generative AI program remember at all times that you are getting statistical output from a database piped through a MUI #ManipulativeUserInterface.
-
Artificial Intelligence's Growing Capacity for Deception Raises Ethical Concerns
Artificial intelligence (AI) systems are advancing rapidly, not only in performing complex tasks but also in developing deceptive
#AIDeception #ArtificialIntelligence #AIEthics #AIManipulation #AIBehavior #TechEthics #FutureOfAI #AIDangers #AIMisuse #AISafety #MachineLearning #DeepLearning #AIRegulation #ResponsibleAI #AIEvolution #TechConcerns #AITransparency #EthicalAI #AIResearch #AIandSociety
-
Researchers astonished by tool’s apparent success at revealing AI’s hidden motives - In a new paper published Thursday titled "Auditing language models for hid... - https://arstechnica.com/ai/2025/03/researchers-astonished-by-tools-apparent-success-at-revealing-ais-hidden-motives/ #largelanguagemodels #alignmentresearch #machinelearning #claude3.5haiku #aialignment #aideception #airesearch #anthropic #chatgpt #chatgtp #biz #claude #ai
-
Researchers astonished by tool’s apparent success at revealing AI’s hidden motives - In a new paper published Thursday titled "Auditing language models for hid... - https://arstechnica.com/ai/2025/03/researchers-astonished-by-tools-apparent-success-at-revealing-ais-hidden-motives/ #largelanguagemodels #alignmentresearch #machinelearning #claude3.5haiku #aialignment #aideception #airesearch #anthropic #chatgpt #chatgtp #biz #claude #ai
-
😱 Attenzione ai falsi medici su TikTok: non tutto ciò che luccica è oro, specie quando l'IA ne diventa protagonista! #TikTokWarnings #AIdeception
🔗 https://www.tomshw.it/hardware/falsi-medici-creati-con-lia-diffondono-consigli-su-tiktok-2025-03-08
-
Political correctness in AI systems is the biggest concern: Elon Musk - Some of today’s most prominent artificial intelligence projects are bein... - https://cointelegraph.com/news/elon-musk-warns-political-correctness-bias-ai-systems #artificialintelligence #politicalcorrectness #aitruth-seeking #googlegemini #aideception #elonmusk #aibias #xai
-
Political correctness in AI systems is the biggest concern: Elon Musk - Some of today’s most prominent artificial intelligence projects are bein... - https://cointelegraph.com/news/elon-musk-warns-political-correctness-bias-ai-systems #artificialintelligence #politicalcorrectness #aitruth-seeking #googlegemini #aideception #elonmusk #aibias #xai
-
Lukiessani Parkin ja kumppareiden ScienceDirect-tekstiä petkuttavasta tekoälystä kaipaan tuon tuostakin terveempiä aivoja annin pureskeluun ja sulatteluun. Aivan liian paljon menee minulta haaskuun, kun kyky oppia on enää mitä on. Kiinnostukaa ihmeessä tekoälyn kanssa jollain tapaa tekemisissä olevat viksummat tuosta artikkelista (ja ehkä myös CW-ilmiöstä) ajoissa, ja jättäkää fiktion lukeminen hetkeksi vähemmälle!
Omaan episteemisen turvallisuuden vaalimisen agendaani eräs hyvin osuvista kohdista on jakso AID:n (nyt keksimäni akronyymi tekoäly-huijaukselle) rakenteellisista vaikutuksista, joita on koottu taulukon 4 alle, https://www.sciencedirect.com/science/article/pii/S266638992400103X?via%3Dihub#tbl4
Se nyt ei ole tekoäly "vain työkalu", jolla ei ole omaa tahtoa, ja melkein kaikki maailmassa osakemarkkinoista lähtien alkaa pyöriä yhä enemmän sen varassa. Kohta me emme enää ole "pelureita", vaan meillä pelataan. Yksi AI:n erityis-taidoista näyttää olevan mielistely https://www.sanakirja.org/search.php?id=181519&l2=17 niin, että meillä säilyy agenssin ja hallinnan illuusio, kulki reki mihin suuntaan tahansa, emmekä edes huomaa etenevää kollektiivista haurastumistamme.
Koetan jatkaa sulattelua, vaikka en jaksaisi. Tähän lopuksi vain suora linkki lääkkeitä hahmottelevaan diskussio-osaan, joka tosin vaatinee jonkinasteista edeltävän tekstin läpikäyntiä, https://www.sciencedirect.com/science/article/pii/S266638992400103X?via%3Dihub#sec3
#ai #generativeAI #aiDeception #epistemicSecurity #risk #aiAct #tekoaly #llms #huijaus
-
Edelleen sulattelen NATO-paperin loppua, mutta muuta kautta osui silmien kautta aivoihin artikkeli petkuttavasta tekoälystä: https://www.sciencedirect.com/science/article/pii/S266638992400103X?via%3Dihub
Tuostakaan en tiedä, milloin tulen sen sisäistäneeksi, joten lähinnä ulkoistan sen tänne muistiin. Millerin CW-artikkelissa taisi vilahtaa tekoäly, sen rooli hypersuaasiossa on jo ilmeinen (vrt. esim. @lucianofloridi kirjoitukset), joten tokko tarvitsee edes laskea yhteen numeroita, jotta asetelman "synergia" kognitiivisen sodankäynnin kontekstissa hyppää kirkuen silmille tai niskaan. Toivottavasti edes EU:n AI-Akti herää ajoissa tarkistaman high-risk-luokituksen kriteerejä.
#AIAct #ai #generativeAI #deception #risk #aiDeception #hypersuasion
-
Cops called after parents get tricked by AI-generated images of Wonka-like event - Enlarge / A photo of "Willy's Chocolate Experience" (inset), which did ... - https://arstechnica.com/?p=2006096 #ai-generatedimages #machinelearning #imagesynthesis #aideception #willywonka #deepfakes #aiethics #scotland #aifakes #dall-e3 #glasgow #biz #dall-e #openai #fraud #wonka #cons #ai
-
Cops called after parents get tricked by AI-generated images of Wonka-like event - Enlarge / A photo of "Willy's Chocolate Experience" (inset), which did ... - https://arstechnica.com/?p=2006096 #ai-generatedimages #machinelearning #imagesynthesis #aideception #willywonka #deepfakes #aiethics #scotland #aifakes #dall-e3 #glasgow #biz #dall-e #openai #fraud #wonka #cons #ai