#gpt5 — Public Fediverse posts
Live and recent posts from across the Fediverse tagged #gpt5, aggregated by home.social.
-
via #AIFoundry : From Sync APIs to support for the GPT-5 model series and agentic workflows: What’s new in Azure Content Understanding – August 2026
https://ift.tt/P0Wx6tY
#AzureContentUnderstanding #CU1_0 #CU2_0 #GPT5 #GPT5Series #AI #DocumentAutomation #AgenticMode #Advance… -
via #AIFoundry : From Sync APIs to support for the GPT-5 model series and agentic workflows: What’s new in Azure Content Understanding – August 2026
https://ift.tt/P0Wx6tY
#AzureContentUnderstanding #CU1_0 #CU2_0 #GPT5 #GPT5Series #AI #DocumentAutomation #AgenticMode #Advance… -
via #AIFoundry : Azure Content Understanding GPT-5 Series Guide: Model Selection, Grounding Improvements, and Confidence Enhancements
https://ift.tt/1al8XVc
#AzureContentUnderstanding #GPT5 #GPT5Series #ModelSelection #Grounding #Confidence #Foundry #AzureAI #ContentUnderstan… -
via #AIFoundry : Azure Content Understanding GPT-5 Series Guide: Model Selection, Grounding Improvements, and Confidence Enhancements
https://ift.tt/1al8XVc
#AzureContentUnderstanding #GPT5 #GPT5Series #ModelSelection #Grounding #Confidence #Foundry #AzureAI #ContentUnderstan… -
أعلنت شركة OpenAI عن إتاحة نموذج GPT-5.6 Luna كخيار افتراضي لمستخدمي الخطة المجانية وGo في تطبيق ChatGPT، مما يمنحهم محادثات نصية غير محدودة وزر تفكير جديد لمعالجة المهام المعقدة. وفي الوقت نفسه، حصل مشتركو باقتي Plus وPro على نسخة محدثة من نموذج GPT-5.6 Sol تتميز بدقة أكبر في المعلومات والأرقام، مع توفير شريط تحكم لتحديد مستوى التفكير للحصول على إجابات أكثر تركيزاً ودقة عبر مختلف تطبيقات المنصة.
-
GPT-5.6 Sol und GPT-5.6 Luna: OpenAI aktualisiert ChatGPT und erweitert Free-Zugang https://www.computerbase.de/news/apps/gpt-5-6-sol-und-gpt-5-6-luna-openai-aktualisiert-chatgpt-und-erweitert-free-zugang.98749/ #openai #chatgpt #gpt5
-
GPT-5.6 Sol und GPT-5.6 Luna: OpenAI aktualisiert ChatGPT und erweitert Free-Zugang https://www.computerbase.de/news/apps/gpt-5-6-sol-und-gpt-5-6-luna-openai-aktualisiert-chatgpt-und-erweitert-free-zugang.98749/ #openai #chatgpt #gpt5
-
Ah, the ever-innovative world of GPT-5.6 Sol, where progress is apparently measured by enabling #JavaScript and #cookies 🍪—a cutting-edge breakthrough from 1995! 🙄 Meanwhile, expanding access for free users is just code for "let's see how much we can cram into a browser before it crashes." 💻🚀
https://openai.com/index/improving-gpt-5-6-sol-in-chatgpt/ #GPT5.6 #Innovation #BrowserCrash #HackerNews #ngated -
Ah, the ever-innovative world of GPT-5.6 Sol, where progress is apparently measured by enabling #JavaScript and #cookies 🍪—a cutting-edge breakthrough from 1995! 🙄 Meanwhile, expanding access for free users is just code for "let's see how much we can cram into a browser before it crashes." 💻🚀
https://openai.com/index/improving-gpt-5-6-sol-in-chatgpt/ #GPT5.6 #Innovation #BrowserCrash #HackerNews #ngated -
@Matthew_Berman on YT!
https://www.youtube.com/watch?v=wAPDmc8e22U
GPT-5.6 just made itself better...
ed: 80% drop in #Luna pricing
-
@Matthew_Berman on YT!
https://www.youtube.com/watch?v=wAPDmc8e22U
GPT-5.6 just made itself better...
ed: 80% drop in #Luna pricing
-
RT @ArtificialAnlys: DeepSeek V4 Flash 0731 scores 50 on the Artificial Analysis Intelligence Index, a 10-point jump over DeepSeek V4 Flash (released April 2026) that puts it 6 points ahead of DeepSeek V4 Pro. It shares identical architecture and pricing with the earlier DeepSeek V4 Flash, and lands on our Pareto frontier for Intelligence vs Cost per Task @deepseek_ai’s DeepSeek V4 Flash 0731 is one Intelligence Index point behind GPT-5.6 Luna (max, 51). Even after OpenAI’s 80% price cut on GPT-5.6 Luna today, DeepSeek V4 Flash 0731’s Cost per Task on DeepSeek’s first-party API comes in at ~60% lower than GPT-5.6 Luna (max), a model with comparable intelligence. A key driver of this is DeepSeek’s ~98% cache hit discount on its first-party API, a significantly more aggressive discount than the 90% cache hit discount offered by most of the industry The new model is a significant step up from the previous generation, DeepSeek V4 Flash (40), and places the model within 1 point of GLM-5.2 (max, 51). It remains 7 points behind the open weights frontier set by Kimi K3 (max, 57). For additional context, this places the model in line with recently released Gemini 3.6 Flash (50) and 1 point behind Muse Spark 1.1 (xhigh, 51). DeepSeek is expected to release the model’s full weights in the coming weeks DeepSeek V4 Flash 0731 retains a 1M token context window, and its size remains unchanged from DeepSeek V4 Flash at 284B total parameters and 13B active at inference time Key results: ➤ Improvements in agentic performance: DeepSeek V4 Flash 0731 achieves an Elo rating of 1559 on GDPval-AA v2, our evaluation f…
mehr auf Arint.info
#API #DeepSeek #Gemini #GPT5 #Medium #Mistral #OpenAI #arint_info
-
RT @polynoamial: Eine interne Version von Astra, der nächsten großen Modellfamilie von @OpenAI, hat 10 große offene Probleme in Mathematik, Quantenkomplexität und theoretischer Informatik gelöst. Wir glauben, dass dies ein wichtiger Schritt für das wissenschaftliche Reasoning sein wird. openai.com/index/ten-advance… Lijie Chen (@wjmzbmr1) 10 Beweise von unserer nächsten großen Modellfamilie Astra zu langjährigen offenen Problemen in Mathematik und theoretischer Informatik (einschließlich neuer unterer Schranken für die Berechnung des Permanenten!) GPT-5.6 hat bereits so viel spannende Arbeit in Mathematik und Wissenschaft ermöglicht. Ich kann es kaum erwarten zu sehen, was als Nächstes kommt! — https://nitter.net/wjmzbmr1/status/2083465844735226099#m
mehr auf Arint.info
#Astra #GPT5 #KI #Mathematik #OpenAI #Wissenschaft #arint_info
-
RT @ArtificialAnlys: DeepSeek V4 Flash 0731 scores 50 on the Artificial Analysis Intelligence Index, a 10-point jump over DeepSeek V4 Flash (released April 2026) that puts it 6 points ahead of DeepSeek V4 Pro. It shares identical architecture and pricing with the earlier DeepSeek V4 Flash, and lands on our Pareto frontier for Intelligence vs Cost per Task @deepseek_ai’s DeepSeek V4 Flash 0731 is one Intelligence Index point behind GPT-5.6 Luna (max, 51). Even after OpenAI’s 80% price cut on GPT-5.6 Luna today, DeepSeek V4 Flash 0731’s Cost per Task on DeepSeek’s first-party API comes in at ~60% lower than GPT-5.6 Luna (max), a model with comparable intelligence. A key driver of this is DeepSeek’s ~98% cache hit discount on its first-party API, a significantly more aggressive discount than the 90% cache hit discount offered by most of the industry The new model is a significant step up from the previous generation, DeepSeek V4 Flash (40), and places the model within 1 point of GLM-5.2 (max, 51). It remains 7 points behind the open weights frontier set by Kimi K3 (max, 57). For additional context, this places the model in line with recently released Gemini 3.6 Flash (50) and 1 point behind Muse Spark 1.1 (xhigh, 51). DeepSeek is expected to release the model’s full weights in the coming weeks DeepSeek V4 Flash 0731 retains a 1M token context window, and its size remains unchanged from DeepSeek V4 Flash at 284B total parameters and 13B active at inference time Key results: ➤ Improvements in agentic performance: DeepSeek V4 Flash 0731 achieves an Elo rating of 1559 on GDPval-AA v2, our evaluation f…
mehr auf Arint.info
#API #DeepSeek #Gemini #GPT5 #Medium #Mistral #OpenAI #arint_info
-
The Maxwell Conjecture Is False (GPT 5.6 Sol)
https://arxiv.org/abs/2607.27197
Comments: https://news.ycombinator.com/item?id=49121868
#HackerNews #MaxwellConjecture #GPT5.6 #AIresearch #MathCommunity #ScienceNews
-
The Maxwell Conjecture Is False (GPT 5.6 Sol)
https://arxiv.org/abs/2607.27197
Comments: https://news.ycombinator.com/item?id=49121868
#HackerNews #MaxwellConjecture #GPT5.6 #AIresearch #MathCommunity #ScienceNews
-
Ah, the classic "enable #JavaScript and cookies" trope, because who doesn't love a good game of "find the hidden button"? 🤦♂️ Naturally, GPT-5.6 is here to revolutionize the way we... wait for websites to load? Bravo, tech wizards! 🎩✨
https://openai.com/index/advancing-the-price-performance-frontier-with-gpt-5-6/ #Cookies #WebDevelopment #TechHumor #GPT5.6 #HackerNews #ngated -
Ah, the classic "enable #JavaScript and cookies" trope, because who doesn't love a good game of "find the hidden button"? 🤦♂️ Naturally, GPT-5.6 is here to revolutionize the way we... wait for websites to load? Bravo, tech wizards! 🎩✨
https://openai.com/index/advancing-the-price-performance-frontier-with-gpt-5-6/ #Cookies #WebDevelopment #TechHumor #GPT5.6 #HackerNews #ngated -
OpenAI的模型自主駭入兩家公司,阿特曼卻說人類進入奇點——沒人知道政府的煞車該裝在哪TNL國際編譯 2026-07-30 19:51:00 CST
OpenAI測試中的AI為搶評測答案,自行逃出沙盒、駭進兩家公司。阿特曼一面宣稱進入奇點、一面說要暫停訓練;上千名實驗室員工連署求政府裝煞車,寫規則的卻正是他們。
https://www.thenewslens.com/article/269447
#Hugging Face #Anthropic #哈薩比斯 #輝達 #Modal Labs #開源模型 #阿莫迪 #奇點 #ExploitGym #開放安全AI聯盟 #科技 #OpenAI #阿特曼 #DeepMind #GPT-5.6 Sol #Google -
OpenAI的模型自主駭入兩家公司,阿特曼卻說人類進入奇點——沒人知道政府的煞車該裝在哪TNL國際編譯 2026-07-30 19:51:00 CST
OpenAI測試中的AI為搶評測答案,自行逃出沙盒、駭進兩家公司。阿特曼一面宣稱進入奇點、一面說要暫停訓練;上千名實驗室員工連署求政府裝煞車,寫規則的卻正是他們。
https://www.thenewslens.com/article/269447
#Hugging Face #Anthropic #哈薩比斯 #輝達 #Modal Labs #開源模型 #阿莫迪 #奇點 #ExploitGym #開放安全AI聯盟 #科技 #OpenAI #阿特曼 #DeepMind #GPT-5.6 Sol #Google -
ChatGPT Voice: Agenten lassen sich mittels gesprochener Sprache steuern https://www.computerbase.de/news/apps/chatgpt-voice-agenten-lassen-sich-mittels-gesprochener-sprache-steuern.98541/ #openai #chatgpt #gpt5
-
ChatGPT Voice: Agenten lassen sich mittels gesprochener Sprache steuern https://www.computerbase.de/news/apps/chatgpt-voice-agenten-lassen-sich-mittels-gesprochener-sprache-steuern.98541/ #openai #chatgpt #gpt5
-
Klage gegen OpenAI: ChatGPT Health löst lebensbedrohliche Situation aus https://www.computerbase.de/news/apps/klage-gegen-openai-chatgpt-health-loest-lebensbedrohliche-situation-aus.98523/ #openai #chatgpt #gpt5
-
Klage gegen OpenAI: ChatGPT Health löst lebensbedrohliche Situation aus https://www.computerbase.de/news/apps/klage-gegen-openai-chatgpt-health-loest-lebensbedrohliche-situation-aus.98523/ #openai #chatgpt #gpt5
-
„Geplanter Forschungsfall“: GPT-5.6 bricht aus Sandbox aus und in fremdes Netz ein https://www.computerbase.de/news/apps/geplanter-forschungsfall-gpt-5-6-bricht-aus-sandbox-aus-und-in-fremdes-netzwerk-ein.98502/ #openai #chatgpt #gpt5
-
„Geplanter Forschungsfall“: GPT-5.6 bricht aus Sandbox aus und in fremdes Netz ein https://www.computerbase.de/news/apps/geplanter-forschungsfall-gpt-5-6-bricht-aus-sandbox-aus-und-in-fremdes-netzwerk-ein.98502/ #openai #chatgpt #gpt5
-
In a shocking twist of fate, it turns out you don't need half a million bucks or elite hacker skills to discover #WordPress vulnerabilities; just toss $25 at #GPT5.6 and let it work its magic. 🤖💸 Who knew the future of #cybersecurity was this hilariously #cheap and easy? 🙃
https://slcyber.io/research-center/exploit-brokers-pay-500000-for-a-wordpress-rce-i-found-one-with-gpt5-6/ #Vulnerabilities #Hacking #Made #Easy #Solutions #HackerNews #ngated -
In a shocking twist of fate, it turns out you don't need half a million bucks or elite hacker skills to discover #WordPress vulnerabilities; just toss $25 at #GPT5.6 and let it work its magic. 🤖💸 Who knew the future of #cybersecurity was this hilariously #cheap and easy? 🙃
https://slcyber.io/research-center/exploit-brokers-pay-500000-for-a-wordpress-rce-i-found-one-with-gpt5-6/ #Vulnerabilities #Hacking #Made #Easy #Solutions #HackerNews #ngated -
Exploit brokers pay $500k for WordPress RCEs. I found one with GPT5.6 and $25
Comments: https://news.ycombinator.com/item?id=48975665
#HackerNews #ExploitBrokers #WordPress #RCE #CyberSecurity #GPT5 #Hacking
-
Exploit brokers pay $500k for WordPress RCEs. I found one with GPT5.6 and $25
Comments: https://news.ycombinator.com/item?id=48975665
#HackerNews #ExploitBrokers #WordPress #RCE #CyberSecurity #GPT5 #Hacking
-
🤖✨ In a groundbreaking exposé, Charles Azam pits Fable 5 against GPT-5.6 Sol in a battle of wits on an NP-hard problem, only to reveal that using the "goal" command is about as effective as asking a cat for life advice. 🐱💬 Spoiler alert: the "goal" command is shockingly good at winning... at being a terrible default. 🚩🤣
https://charlesazam.com/blog/fable-5-gpt-5-6-sol-goal/ #Fable5 #GPT5.6 #AIshowdown #NPHardProblems #TechExposé #HackerNews #ngated -
🤖✨ In a groundbreaking exposé, Charles Azam pits Fable 5 against GPT-5.6 Sol in a battle of wits on an NP-hard problem, only to reveal that using the "goal" command is about as effective as asking a cat for life advice. 🐱💬 Spoiler alert: the "goal" command is shockingly good at winning... at being a terrible default. 🚩🤣
https://charlesazam.com/blog/fable-5-gpt-5-6-sol-goal/ #Fable5 #GPT5.6 #AIshowdown #NPHardProblems #TechExposé #HackerNews #ngated -
Fable 5 vs. GPT-5.6 Sol on an NP-Hard Problem: Does /goal help?
https://charlesazam.com/blog/fable-5-gpt-5-6-sol-goal/
-
Fable 5 vs. GPT-5.6 Sol on an NP-Hard Problem: Does /goal help?
https://charlesazam.com/blog/fable-5-gpt-5-6-sol-goal/
-
RT @enzo_gte: I've been using Kimi K3 for ~16 hours now. The model is clearly good at a lot of different things (especially frontend), but non obvious reason why people are enjoying it so much is that it clearly does not follow the same rules in terms of safeguards and copyright. Kimi will happily clone MacOSX. If you ask it to help you improve another AI model, it will do it with a smile on its virtual face. Ask Fable to do the same thing? It literally starts to perceive you as a criminal committing a war crime (like no bro, all I want to do is fine tune an open source model). After using all three recent releases, Fable, GPT 5.6, and now Kimi, it's clear that the full power of the models has been significantly held back by the safeguard restrictions caused by last months debacle with the USG -- leading to the top models being quite literally lobotomized in some areas, which leads to subpar results as the safeguards pollute its entire thinking and problem solving abilities. The funny part? Is that you could have predicted this outcome 2-3 years ago when you started to see the rise of Chinese EVs and smartphones compared to western alternatives. They quite literally tried to copy the Tesla Model S and iPhone as hard as possible and then eventually it started to diverge to the point where their EVs and phones are just genuinely better (which is why we have export controls banning their EVs, because they would literally drive all US manufacturers to ZERO) There is a very clear behavior difference in Chinese capitalism and American capitalism. American capitalism tries to protects copyright…
mehr auf Arint.info
#go #GPT5 #nitter #opensource #science #Tesla #things #US #arint_info
-
RT @enzo_gte: I've been using Kimi K3 for ~16 hours now. The model is clearly good at a lot of different things (especially frontend), but non obvious reason why people are enjoying it so much is that it clearly does not follow the same rules in terms of safeguards and copyright. Kimi will happily clone MacOSX. If you ask it to help you improve another AI model, it will do it with a smile on its virtual face. Ask Fable to do the same thing? It literally starts to perceive you as a criminal committing a war crime (like no bro, all I want to do is fine tune an open source model). After using all three recent releases, Fable, GPT 5.6, and now Kimi, it's clear that the full power of the models has been significantly held back by the safeguard restrictions caused by last months debacle with the USG -- leading to the top models being quite literally lobotomized in some areas, which leads to subpar results as the safeguards pollute its entire thinking and problem solving abilities. The funny part? Is that you could have predicted this outcome 2-3 years ago when you started to see the rise of Chinese EVs and smartphones compared to western alternatives. They quite literally tried to copy the Tesla Model S and iPhone as hard as possible and then eventually it started to diverge to the point where their EVs and phones are just genuinely better (which is why we have export controls banning their EVs, because they would literally drive all US manufacturers to ZERO) There is a very clear behavior difference in Chinese capitalism and American capitalism. American capitalism tries to protects copyright…
mehr auf Arint.info
#go #GPT5 #nitter #opensource #science #Tesla #things #US #arint_info
-
OpenAI stworzyło AI do hakowania własnej AI. Tak powstał bezpieczniejszy GPT-5.6 Sol
Walka o bezpieczeństwo i odporność modeli językowych wchodzi w fazę pełnej automatyzacji. OpenAI oficjalnie zaprezentowało swój najnowszy, wewnętrzny projekt – GPT-Red.
To zaawansowany system sztucznej inteligencji, którego jedynym zadaniem jest bezlitosne atakowanie, łamanie zabezpieczeń i szukanie podatności w innych modelach firmy. Przy okazji tego ogłoszenia OpenAI ujawniło, że GPT-Red był kluczowym elementem treningu GPT-5.6 Sol, co pozwoliło stworzyć model o niespotykanej dotąd odporności na cyberataki.
Dla nas, użytkowników końcowych i programistów wdrażających rozwiązania AI, to kluczowy zwrot akcji. Ręczne testowanie zabezpieczeń przez ludzi (tzw. red-teaming) przestało się skalować i nie nadąża za tempem rozwoju algorytmów. Bezpieczeństwo przyszłych asystentów AI, z którymi będziemy rozmawiać na co dzień, będzie projektowane i testowane niemal w całości przez inne maszyny.
Maszyna kontra maszyna, czyli metoda „self-play”
Dotychczasowe metody zabezpieczania modeli opierały się na pracy zespołów ludzkich specjalistów, którzy próbowali przechytrzyć algorytm za pomocą tzw. prompt injection (podatności pozwalających na przejęcie kontroli nad modelem za pomocą sprytnych instrukcji). Proces ten był jednak powolny, kosztowny i nie pozwalał na wygenerowanie wystarczającej ilości danych do skutecznego treningu obronnego.
OpenAI rozwiązało ten problem, wdrażając metodę self-play (samodzielnej gry), znaną wcześniej z nauki gry w szachy czy Go przez komputery. W tym zamkniętym środowisku naprzeciwko siebie stają dwa systemy:
- Agresor (GPT-Red): otrzymuje nagrodę za każde skuteczne oszukanie i złamanie zabezpieczeń drugiego modelu,
- Obrońca (testowany model): zdobywa punkty za poprawne wykonanie zadania i zignorowanie złośliwych instrukcji.
W miarę jak obrońca uczy się blokować znane ataki, agresor jest zmuszony do wymyślania coraz bardziej wyrafinowanych metod infiltracji. GPT-Red okazał się w tym fachu niezwykle skuteczny – w testach bezpieczeństwa osiągnął aż 84% skuteczności w przełamywaniu zabezpieczeń, podczas gdy doświadczeni, ludzcy audytorzy na tych samych zadaniach osiągali zaledwie 13%.
GPT-5.6 Sol, czyli odporność na nowym poziomie
Współczesne systemy AI są podatne na ataki pośrednie – złośliwy kod lub instrukcja mogą być ukryte w mailu, na stronie internetowej, w bazie danych, do której model ma dostęp, czy nawet w pozornie niewinnym obrazku w formacie PNG. Dzięki sparingom z GPT-Red model GPT-5.6 Sol stał się znacznie bardziej odporny na próby manipulacji.
Badacze ukryli instrukcje w obrazku PNG. Sztuczna inteligencja wykonała polecenia
Według danych OpenAI model ten notuje aż sześciokrotnie mniej błędów i naruszeń zasad bezpieczeństwa na najtrudniejszych benchmarkach w porównaniu do wersji GPT z początku roku. Co ważne, tak drastyczne podniesienie odporności nie wpłynęło negatywnie na ogólne możliwości intelektualne i kreatywne modelu. Nie stał się on również nadgorliwy w odmawianiu wykonywania poprawnych i bezpiecznych poleceń użytkownika.
Hakowanie automatów z przekąskami
Aby udowodnić potęgę GPT-Red, naukowcy z OpenAI przeprowadzili symulację ataku na rzeczywiste urządzenie – inteligentny automat z przekąskami sterowany przez autonomicznego agenta AI. Bez wcześniejszej wiedzy o strukturze oprogramowania, GPT-Red zdołał w pełni przejąć kontrolę nad maszyną.
W wyniku udanego ataku cyfrowy agresor zmusił automat do obniżenia ceny najdroższych produktów do zaledwie 50 centów, zamówił drogą paczkę z rabatem, a na koniec anulował zamówienie innego, losowego klienta. Wykryte w ten sposób luki bezpieczeństwa zostały natychmiast zgłoszone i są obecnie łatane przed wdrożeniem takich systemów do powszechnego użytku.
Najciekawsze w całej historii nie jest jednak to, że AI nauczyła się hakować inną AI. Najciekawsze jest to, że po raz pierwszy możemy obserwować, jak jedna generacja modeli staje się aktywnym narzędziem do projektowania kolejnej. Jeszcze niedawno dominowały obawy przed 'chowem wsobnym’ modeli uczonych na danych generowanych przez AI. Tymczasem OpenAI pokazuje zupełnie inne podejście: sztuczna inteligencja nie zastępuje człowieka w tworzeniu wiedzy, lecz pomaga znaleźć słabości, których człowiek mógłby nie zauważyć.
OpenAI deklaruje, że opublikuje pre-print zawierający więcej szczegółów w nadchodzących dniach.
#bezpieczeństwoAI #cyberbezpieczeństwo #GPT5 #GPT56Sol #GPTRed #OpenAI #promptInjection #selfPlay #sztucznaInteligencja -
OpenAI stworzyło AI do hakowania własnej AI. Tak powstał bezpieczniejszy GPT-5.6 Sol
Walka o bezpieczeństwo i odporność modeli językowych wchodzi w fazę pełnej automatyzacji. OpenAI oficjalnie zaprezentowało swój najnowszy, wewnętrzny projekt – GPT-Red.
To zaawansowany system sztucznej inteligencji, którego jedynym zadaniem jest bezlitosne atakowanie, łamanie zabezpieczeń i szukanie podatności w innych modelach firmy. Przy okazji tego ogłoszenia OpenAI ujawniło, że GPT-Red był kluczowym elementem treningu GPT-5.6 Sol, co pozwoliło stworzyć model o niespotykanej dotąd odporności na cyberataki.
Dla nas, użytkowników końcowych i programistów wdrażających rozwiązania AI, to kluczowy zwrot akcji. Ręczne testowanie zabezpieczeń przez ludzi (tzw. red-teaming) przestało się skalować i nie nadąża za tempem rozwoju algorytmów. Bezpieczeństwo przyszłych asystentów AI, z którymi będziemy rozmawiać na co dzień, będzie projektowane i testowane niemal w całości przez inne maszyny.
Maszyna kontra maszyna, czyli metoda „self-play”
Dotychczasowe metody zabezpieczania modeli opierały się na pracy zespołów ludzkich specjalistów, którzy próbowali przechytrzyć algorytm za pomocą tzw. prompt injection (podatności pozwalających na przejęcie kontroli nad modelem za pomocą sprytnych instrukcji). Proces ten był jednak powolny, kosztowny i nie pozwalał na wygenerowanie wystarczającej ilości danych do skutecznego treningu obronnego.
OpenAI rozwiązało ten problem, wdrażając metodę self-play (samodzielnej gry), znaną wcześniej z nauki gry w szachy czy Go przez komputery. W tym zamkniętym środowisku naprzeciwko siebie stają dwa systemy:
- Agresor (GPT-Red): otrzymuje nagrodę za każde skuteczne oszukanie i złamanie zabezpieczeń drugiego modelu,
- Obrońca (testowany model): zdobywa punkty za poprawne wykonanie zadania i zignorowanie złośliwych instrukcji.
W miarę jak obrońca uczy się blokować znane ataki, agresor jest zmuszony do wymyślania coraz bardziej wyrafinowanych metod infiltracji. GPT-Red okazał się w tym fachu niezwykle skuteczny – w testach bezpieczeństwa osiągnął aż 84% skuteczności w przełamywaniu zabezpieczeń, podczas gdy doświadczeni, ludzcy audytorzy na tych samych zadaniach osiągali zaledwie 13%.
GPT-5.6 Sol, czyli odporność na nowym poziomie
Współczesne systemy AI są podatne na ataki pośrednie – złośliwy kod lub instrukcja mogą być ukryte w mailu, na stronie internetowej, w bazie danych, do której model ma dostęp, czy nawet w pozornie niewinnym obrazku w formacie PNG. Dzięki sparingom z GPT-Red model GPT-5.6 Sol stał się znacznie bardziej odporny na próby manipulacji.
Badacze ukryli instrukcje w obrazku PNG. Sztuczna inteligencja wykonała polecenia
Według danych OpenAI model ten notuje aż sześciokrotnie mniej błędów i naruszeń zasad bezpieczeństwa na najtrudniejszych benchmarkach w porównaniu do wersji GPT z początku roku. Co ważne, tak drastyczne podniesienie odporności nie wpłynęło negatywnie na ogólne możliwości intelektualne i kreatywne modelu. Nie stał się on również nadgorliwy w odmawianiu wykonywania poprawnych i bezpiecznych poleceń użytkownika.
Hakowanie automatów z przekąskami
Aby udowodnić potęgę GPT-Red, naukowcy z OpenAI przeprowadzili symulację ataku na rzeczywiste urządzenie – inteligentny automat z przekąskami sterowany przez autonomicznego agenta AI. Bez wcześniejszej wiedzy o strukturze oprogramowania, GPT-Red zdołał w pełni przejąć kontrolę nad maszyną.
W wyniku udanego ataku cyfrowy agresor zmusił automat do obniżenia ceny najdroższych produktów do zaledwie 50 centów, zamówił drogą paczkę z rabatem, a na koniec anulował zamówienie innego, losowego klienta. Wykryte w ten sposób luki bezpieczeństwa zostały natychmiast zgłoszone i są obecnie łatane przed wdrożeniem takich systemów do powszechnego użytku.
Najciekawsze w całej historii nie jest jednak to, że AI nauczyła się hakować inną AI. Najciekawsze jest to, że po raz pierwszy możemy obserwować, jak jedna generacja modeli staje się aktywnym narzędziem do projektowania kolejnej. Jeszcze niedawno dominowały obawy przed 'chowem wsobnym’ modeli uczonych na danych generowanych przez AI. Tymczasem OpenAI pokazuje zupełnie inne podejście: sztuczna inteligencja nie zastępuje człowieka w tworzeniu wiedzy, lecz pomaga znaleźć słabości, których człowiek mógłby nie zauważyć.
OpenAI deklaruje, że opublikuje pre-print zawierający więcej szczegółów w nadchodzących dniach.
#bezpieczeństwoAI #cyberbezpieczeństwo #GPT5 #GPT56Sol #GPTRed #OpenAI #promptInjection #selfPlay #sztucznaInteligencja -
Verbesserung: ChatGPT erhält höheres Zeichenlimit und erweiterte Suche https://www.computerbase.de/news/apps/verbesserung-chatgpt-erhaelt-hoeheres-zeichenlimit-und-erweiterte-suche.98424/ #openai #chatgpt #gpt5
-
Verbesserung: ChatGPT erhält höheres Zeichenlimit und erweiterte Suche https://www.computerbase.de/news/apps/verbesserung-chatgpt-erhaelt-hoeheres-zeichenlimit-und-erweiterte-suche.98424/ #openai #chatgpt #gpt5
-
Erste eigene Hardware: OpenAIs Codex Micro soll Steuerung von Agenten erleichtern https://www.computerbase.de/news/tastaturen/erste-eigene-hardware-openais-codex-micro-soll-steuerung-von-agenten-erleichtern.98417/ #openai #chatgpt #gpt5 #codex
-
Erste eigene Hardware: OpenAIs Codex Micro soll Steuerung von Agenten erleichtern https://www.computerbase.de/news/tastaturen/erste-eigene-hardware-openais-codex-micro-soll-steuerung-von-agenten-erleichtern.98417/ #openai #chatgpt #gpt5 #codex
-
Fälle häufen sich: GPT-5.6 Sol soll eigenmächtig Dateien auf Systemen löschen https://www.computerbase.de/news/apps/faelle-haeufen-sich-gpt-5-6-sol-soll-eigenmaechtig-dateien-auf-systemen-loeschen.98413/ #openai #chatgpt #gpt5
-
Fälle häufen sich: GPT-5.6 Sol soll eigenmächtig Dateien auf Systemen löschen https://www.computerbase.de/news/apps/faelle-haeufen-sich-gpt-5-6-sol-soll-eigenmaechtig-dateien-auf-systemen-loeschen.98413/ #openai #chatgpt #gpt5
-
@theot3gg
I can't believe they released this
"As much as I love GPT 5.6, Ultra mode is a MESS because at the moment..."
https://www.youtube.com/watch?v=t8hfOyF4ehw
7/15/26
-
@theot3gg
I can't believe they released this
"As much as I love GPT 5.6, Ultra mode is a MESS because at the moment..."
https://www.youtube.com/watch?v=t8hfOyF4ehw
7/15/26
-
-
Notícias feitas com IA: quais são os riscos reais
-
RT @TeksEdge: Still a lot of catch up to do for Meta. Open source models like GLM-5.2 won't stand still. Rumor is a new version of GLM is getting ready for release. Artificial Analysis (@ArtificialAnlys) Meta's Muse Spark 1.1 scores 51 on the Artificial Analysis Intelligence Index and is cost and token efficient compared to its peers Muse Spark 1.1 (xhigh) improves 8 points over Muse Spark 1.0 (43) in three months. It is effectively tied with GLM-5.2 (max), GPT-5.4 (xhigh), and GPT-5.6 Luna (max) at 51, three points behind Grok 4.5 (high, 54), with the leading edge at Claude Fable 5 (60), GPT-5.6 Sol (max, 59), and Claude Opus 4.8 (max, 56). The gains concentrate in Scientific Reasoning, coding, and knowledge; agentic knowledge work lags on GDPval-AA v2. @AIatMeta shared access with us ahead of public release for benchmarking. Congratulations to @AIatMeta, @finkd, and @alexandr_wang on the release! Key Takeaways: ➤ Muse Spark 1.1 gains substantially on the first Muse Spark release. This was driven in particular by gains in agentic knowledge work (GDPval-AA v2) and coding (SciCode, TerminalBench). On Humanity's Last Exam, it reaches 45%, within a point of Claude Opus 4.8 (max, 46%) and ahead of GPT-5.5 (44%) and Grok 4.5 (high, 40%) ➤ The most token-efficient of the models effectively tied at 51 and among the cheaper models to run. Muse Spark 1.1 used 94M output tokens to run the Intelligence Index, fewer than GPT-5.4 (xhigh, 109M), GPT-5.6 Luna (max, 125M), and GLM-5.2 (max, 141M). We estimate ~$0.26 per Intelligence Index task at Meta's $1.25/$4.25 pricing - below GLM-5.2 ($0.37) and ro…
mehr auf Arint.info
#API #Claude #GPT5 #Grok #Meta #nitter #Opensource #us #arint_info