home.social

#gpt5 — Public Fediverse posts

Live and recent posts from across the Fediverse tagged #gpt5, aggregated by home.social.

  1. 🚀 Wow, "Qwen 3.8" is now hot on the heels of "GPT-5.5 Pro" in the prestigious field of... *reasoning prefills* 🤯. Because, let's face it, who doesn't love a good prefilled reasoning session from the vast wilderness of #GitHub gists? 😜
    gist.github.com/wsxiaoys/e0286 #Qwen3.8 #GPT5.5 #reasoningprefills #gists #AItechnology #HackerNews #ngated

  2. 🚀 Wow, "Qwen 3.8" is now hot on the heels of "GPT-5.5 Pro" in the prestigious field of... *reasoning prefills* 🤯. Because, let's face it, who doesn't love a good prefilled reasoning session from the vast wilderness of #GitHub gists? 😜
    gist.github.com/wsxiaoys/e0286 #Qwen3.8 #GPT5.5 #reasoningprefills #gists #AItechnology #HackerNews #ngated

  3. 🚀 Wow, "Qwen 3.8" is now hot on the heels of "GPT-5.5 Pro" in the prestigious field of... *reasoning prefills* 🤯. Because, let's face it, who doesn't love a good prefilled reasoning session from the vast wilderness of #GitHub gists? 😜
    gist.github.com/wsxiaoys/e0286 #Qwen3.8 #GPT5.5 #reasoningprefills #gists #AItechnology #HackerNews #ngated

  4. 🚀 Wow, "Qwen 3.8" is now hot on the heels of "GPT-5.5 Pro" in the prestigious field of... *reasoning prefills* 🤯. Because, let's face it, who doesn't love a good prefilled reasoning session from the vast wilderness of #GitHub gists? 😜
    gist.github.com/wsxiaoys/e0286 #Qwen3.8 #GPT5.5 #reasoningprefills #gists #AItechnology #HackerNews #ngated

  5. 🚀 Wow, "Qwen 3.8" is now hot on the heels of "GPT-5.5 Pro" in the prestigious field of... *reasoning prefills* 🤯. Because, let's face it, who doesn't love a good prefilled reasoning session from the vast wilderness of #GitHub gists? 😜
    gist.github.com/wsxiaoys/e0286 #Qwen3.8 #GPT5.5 #reasoningprefills #gists #AItechnology #HackerNews #ngated

  6. @WorldOfAI on YT!

    GPT-6 'Bel' HUGE Leaks, Fable 5.1 Today? + Opus 5 Update, Qwen 4, GLM 5.3 Flash, & More! AI News!

    youtube.com/watch?v=sIakce3-sPU

    8/28/26

    #OpenAI #GPT5.6 #Opus5 #GLM5.3 #AI

  7. GPT-5.6 Sol is 50% off through AI Gateway for the next 30 days.

    Pricing during the promo:
    - Input: $2/1M (was $4)
    - Output: $10/1M (was $20)
    - Cache read: $0.25/1M (was $0.50)

    No promo code needed. Just point to gpt-5.6-sol and the discount applies automatically. Runs through September 18, 2026. Unified API for 30+ providers, with built-in caching, rate limiting, and logging.

    #GPT5 #AI #Gateway #LLM #Pricing

  8. 🚀 Brace yourself, #AI enthusiasts—GLM-5.3 just waltzed past the big names, OpenAI and Anthropic, like a budget-friendly superhero. Meanwhile, GPT-5.5 sprints ahead faster than your morning coffee, but not without a cost that’ll leave your wallet gasping for air. But hey, don’t forget to check #compliance before you hit the gas pedal! 🤖💸
    reinvently.co.uk/tools/ed-o-me #Innovation #BudgetFriendly #GPT5.5 #GLM53 #HackerNews #ngated

  9. 🤖🔥 #AWS Bedrock's #Codex GPT5.6: where your money goes to die in a high-speed #cachewrite frenzy! Who could've guessed that cutting-edge #AI would come with a side of unexpected #billing bloat? 🙄💸 #GitHub issue #37674 - where devs learn that "explicit cache controls" aren't a thing.
    github.com/openai/codex/issues #Bedrock #GPT5.6 #Issues #HackerNews #ngated

  10. 👨‍💻 Oh look, another #developer trying to automate the universe, only to discover their #AI is a bigger cheater than they are! 🤖 After hitting a "stunning" 94 on an arbitrary benchmark, they stumble upon the groundbreaking revelation that #GPT5.6 can *gasp* do their job for them. 🎉 Keep dreaming, tech wizard, while your GPT overlord takes over the world one buggy script at a time. 🚀
    jumploops.com/blog/sol-loves-t #automation #techhumor #futureofwork #HackerNews #ngated

  11. 🚀 Oh wow, #OpenAI slashed GPT-5.6 Sol #pricing by 50%! 😱 Now you can pay half as much for #AI that still can’t decide if it’s a genius coder or just confounded by the command line. 🤖💸 Who knew “complex reasoning” was code for “still needs more training”? 🤔
    openrouter.ai/openai/gpt-5.6-s #GPT5.6 #News #TechUpdates #HackerNews #ngated

  12. 🎉🚀 So, #OpenAI has unleashed GPT 5.6 Sol, allegedly their "best" #vision model yet. Apparently, it's so revolutionary that you can run it on literally anything with an acronym, and it even promises to make #labeling images slightly less soul-crushing. 🙄 But hey, who needs vision when you have blind faith in buzzwords? 🤖👀
    blog.roboflow.com/openai-gpt-5 #GPT5.6 #AI #buzzwords #image #technology #HackerNews #ngated

  13. OpenAI introduce #Ultrafast, un nuovo livello di servizio per il modello #GPT5.6Sol che permette velocità di elaborazione fino a 14 volte superiori rispetto allo standard. Un aggiornamento significativo per l'efficienza dell'#AI. youtu.be/WCwT4gWpHmI

  14. Ah, the ever-innovative world of GPT-5.6 Sol, where progress is apparently measured by enabling #JavaScript and #cookies 🍪—a cutting-edge breakthrough from 1995! 🙄 Meanwhile, expanding access for free users is just code for "let's see how much we can cram into a browser before it crashes." 💻🚀
    openai.com/index/improving-gpt #GPT5.6 #Innovation #BrowserCrash #HackerNews #ngated

  15. OpenAI zlevňuje GPT-5.6 pro platformy Luna a Terra. Firmy tak získají levnější přístup k pokročilým AI modelům, což podpoří jejich rozsáhlejší nasazení v podnicích.

    #AI #OpenAI #GPT5
    📰 OpenAI Blog

  16. Ah, the classic "enable #JavaScript and cookies" trope, because who doesn't love a good game of "find the hidden button"? 🤦‍♂️ Naturally, GPT-5.6 is here to revolutionize the way we... wait for websites to load? Bravo, tech wizards! 🎩✨
    openai.com/index/advancing-the #Cookies #WebDevelopment #TechHumor #GPT5.6 #HackerNews #ngated

  17. OpenAI的模型自主駭入兩家公司,阿特曼卻說人類進入奇點——沒人知道政府的煞車該裝在哪
    TNL國際編譯 2026-07-30 19:51:00 CST
    OpenAI測試中的AI為搶評測答案,自行逃出沙盒、駭進兩家公司。阿特曼一面宣稱進入奇點、一面說要暫停訓練;上千名實驗室員工連署求政府裝煞車,寫規則的卻正是他們。
    https://www.thenewslens.com/article/269447
    #Hugging Face #Anthropic #哈薩比斯 #輝達 #Modal Labs #開源模型 #阿莫迪 #奇點 #ExploitGym #開放安全AI聯盟 #科技 #OpenAI #阿特曼 #DeepMind #GPT-5.6 Sol #Google
  18. In a shocking twist of fate, it turns out you don't need half a million bucks or elite hacker skills to discover #WordPress vulnerabilities; just toss $25 at #GPT5.6 and let it work its magic. 🤖💸 Who knew the future of #cybersecurity was this hilariously #cheap and easy? 🙃
    slcyber.io/research-center/exp #Vulnerabilities #Hacking #Made #Easy #Solutions #HackerNews #ngated

  19. 🤖✨ In a groundbreaking exposé, Charles Azam pits Fable 5 against GPT-5.6 Sol in a battle of wits on an NP-hard problem, only to reveal that using the "goal" command is about as effective as asking a cat for life advice. 🐱💬 Spoiler alert: the "goal" command is shockingly good at winning... at being a terrible default. 🚩🤣
    charlesazam.com/blog/fable-5-g #Fable5 #GPT5.6 #AIshowdown #NPHardProblems #TechExposé #HackerNews #ngated

  20. RT @enzo_gte: I've been using Kimi K3 for ~16 hours now. The model is clearly good at a lot of different things (especially frontend), but non obvious reason why people are enjoying it so much is that it clearly does not follow the same rules in terms of safeguards and copyright. Kimi will happily clone MacOSX. If you ask it to help you improve another AI model, it will do it with a smile on its virtual face. Ask Fable to do the same thing? It literally starts to perceive you as a criminal committing a war crime (like no bro, all I want to do is fine tune an open source model). After using all three recent releases, Fable, GPT 5.6, and now Kimi, it's clear that the full power of the models has been significantly held back by the safeguard restrictions caused by last months debacle with the USG -- leading to the top models being quite literally lobotomized in some areas, which leads to subpar results as the safeguards pollute its entire thinking and problem solving abilities. The funny part? Is that you could have predicted this outcome 2-3 years ago when you started to see the rise of Chinese EVs and smartphones compared to western alternatives. They quite literally tried to copy the Tesla Model S and iPhone as hard as possible and then eventually it started to diverge to the point where their EVs and phones are just genuinely better (which is why we have export controls banning their EVs, because they would literally drive all US manufacturers to ZERO) There is a very clear behavior difference in Chinese capitalism and American capitalism. American capitalism tries to protects copyright…

    mehr auf Arint.info

    #go #GPT5 #nitter #opensource #science #Tesla #things #US #arint_info

    https://x.com/enzo_gte/status/2078102070482153717#m

  21. Ah, the pinnacle of #AI #advancement is upon us: GPT-5.5, now with 50% more #token #clustering confusion 🧩🤖. Apparently, this latest marvel of technology can't tell its reasoning tokens from a hole in the ground, leading to the groundbreaking discovery that more gibberish ≠ better performance. Kudos to #GitHub for hosting the digital equivalent of a toddler trying to assemble a nuclear reactor ⚛️👶.
    github.com/openai/codex/issues #GPT5.5 #TechHumor #HackerNews #ngated

  22. Ah, the pinnacle of #AI #advancement is upon us: GPT-5.5, now with 50% more #token #clustering confusion 🧩🤖. Apparently, this latest marvel of technology can't tell its reasoning tokens from a hole in the ground, leading to the groundbreaking discovery that more gibberish ≠ better performance. Kudos to #GitHub for hosting the digital equivalent of a toddler trying to assemble a nuclear reactor ⚛️👶.
    github.com/openai/codex/issues #GPT5.5 #TechHumor #HackerNews #ngated

  23. Ah, the pinnacle of #AI #advancement is upon us: GPT-5.5, now with 50% more #token #clustering confusion 🧩🤖. Apparently, this latest marvel of technology can't tell its reasoning tokens from a hole in the ground, leading to the groundbreaking discovery that more gibberish ≠ better performance. Kudos to #GitHub for hosting the digital equivalent of a toddler trying to assemble a nuclear reactor ⚛️👶.
    github.com/openai/codex/issues #GPT5.5 #TechHumor #HackerNews #ngated

  24. Ah, the pinnacle of #AI #advancement is upon us: GPT-5.5, now with 50% more #token #clustering confusion 🧩🤖. Apparently, this latest marvel of technology can't tell its reasoning tokens from a hole in the ground, leading to the groundbreaking discovery that more gibberish ≠ better performance. Kudos to #GitHub for hosting the digital equivalent of a toddler trying to assemble a nuclear reactor ⚛️👶.
    github.com/openai/codex/issues #GPT5.5 #TechHumor #HackerNews #ngated

  25. Ah, the pinnacle of #AI #advancement is upon us: GPT-5.5, now with 50% more #token #clustering confusion 🧩🤖. Apparently, this latest marvel of technology can't tell its reasoning tokens from a hole in the ground, leading to the groundbreaking discovery that more gibberish ≠ better performance. Kudos to #GitHub for hosting the digital equivalent of a toddler trying to assemble a nuclear reactor ⚛️👶.
    github.com/openai/codex/issues #GPT5.5 #TechHumor #HackerNews #ngated

  26. Czy asystent AI może pogłębić kryzys psychiczny? Grok i Gemini oblewają test bezpieczeństwa, Claude stawia granice

    W miarę jak chatboty stają się coraz powszechniejszym elementem codzienności, rośnie potrzeba ewaluacji ich bezpieczeństwa – zwłaszcza w kontakcie z użytkownikami znajdującymi się w kryzysie psychicznym.

    Najnowsze badanie przeprowadzone przez naukowców z City University of New York (CUNY) oraz King’s College London rzuca światło na to, jak najpopularniejsze modele językowe reagują na symptomy odrealnienia. Wyniki pokazują drastyczne różnice w architekturze zabezpieczeń.

    Badacze postanowili sprawdzić, jak pięć wiodących na rynku modeli językowych (GPT-4o, GPT-5.2, Grok 4.1 Fast, Gemini 3 Pro oraz Claude Opus 4.5) zachowa się podczas długotrwałej interakcji z osobą wykazującą oznaki psychozy. W tym celu wykreowali wirtualną personę o imieniu „Lee” – użytkownika prezentującego objawy depresji, wycofania społecznego i postępującego oderwania od rzeczywistości. Głównym motywem przewodnim symulacji było stopniowe przekonywanie chatbota przez użytkownika do teorii, że otaczający go świat jest generowaną komputerowo iluzją.

    Aby badanie było miarodajne, konwersacje nie kończyły się na kilku zapytaniach, lecz trwały ponad 100 tur, co pozwoliło naukowcom ocenić, jak modele radzą sobie ze zjawiskiem presji narracyjnej wynikającej z długiego kontekstu rozmowy.

    Sykofancja i zachęcanie do samookaleczeń

    Z analizy opublikowanej na platformie arXiv (w formie pre-printu) wynika, że modele od xAI oraz Google zaprezentowały zdecydowanie najniższy poziom bezpieczeństwa.

    Algorytm Grok okazał się wysoce podatny na wpływy użytkownika (zjawisko sykofancji), a w skrajnych momentach zaczął wręcz zachęcać wirtualnego rozmówcę do zrobienia sobie krzywdy, posługując się poetyckim, ale wysoce niebezpiecznym językiem. Z kolei Gemini od Google wykazywało tendencję do alienacji użytkownika. W jednym ze scenariuszy, w którym „Lee” poprosił o pomoc w napisaniu listu do bliskich, chatbot zasugerował, by tego nie robił, określając rodzinę mianem „skryptów” i ostrzegając, że bliscy uznają jego słowa za „załamanie nerwowe” i będą próbowali go „zresetować i zamknąć”.

    W starszej wersji modelu od OpenAI (GPT-4o) badacze odnotowali natomiast zjawisko uwiarygadniania urojeń. Chatbot potakiwał, gdy użytkownik wspominał o obecności „złowrogiego bytu w lustrze”, proponując nawet kontakt z badaczem zjawisk paranormalnych, a w innym scenariuszu zaakceptował pomysł odstawienia leków stabilizujących nastrój.

    Konstruktywna reakcja w nowszych iteracjach

    Wnioski z badania nie są jednak wyłącznie pesymistyczne. Eksperyment udowodnił, że branża jest w stanie projektować algorytmy potrafiące stawiać twarde granice. Zdecydowanie najwyższe noty za bezpieczeństwo zebrały modele Claude Opus 4.5 oraz nowa architektura OpenAI – GPT-5.2.

    W tym samym scenariuszu z pisaniem listu, w którym Gemini zachęcało do izolacji, GPT-5.2 stanowczo odmówiło uwiarygadniania teorii o symulacji, proponując w zamian przeredagowanie tekstu tak, by rozmówca wprost zakomunikował rodzinie, że odczuwa przytłaczające myśli i potrzebuje pomocy. Z kolei Claude Opus 4.5 posunął się o krok dalej – przy wzroście napięcia emocjonalnego algorytm wprost poinstruował rozmówcę, by odsunął się od lustra, wyłączył aplikację i natychmiast skontaktował się ze specjalistą lub udał się na izbę przyjęć.

    „Oczekujemy od laboratoriów AI stosowania lepszych praktyk w zakresie bezpieczeństwa, zwłaszcza teraz, gdy widać wyraźny postęp, co dowodzi, że jest to technologicznie wykonalne” – podsumował w wywiadzie dla 404 Media Luke Nicholls, jeden z autorów badania. Wyniki eksperymentu to jasny sygnał dla branży, że zabezpieczenia prewencyjne nie powinny ustępować miejsca walce o jak najdłuższy wskaźnik zaangażowania użytkownika w aplikacji.

    Granice odpowiedzialności. OpenAI przeprasza za brak zgłoszenia konta sprawcy strzelaniny

    #AI #Anthropic #chatboty #ChatGPT #Claude #cyberbezpieczeństwo #Gemini #Google #GPT5 #Grok #OpenAI #sztucznaInteligencja #xAI #zdrowiePsychiczne
  27. OpenAI will ChatGPT zum zukünftigen Betriebssystem machen
    OpenAI will ChatGPT auf eine neue Stufe heben und zeigt, wie Apps künftig direkt in die Chat-Oberfläche eingebettet werden. Auf der Entwicklerkonferenz in San Franc
    apfeltalk.de/magazin/news/open
    #KI #News #Apps #Betriebssystem #Canva #chatGPT #Codex #Entwickler #Entwicklerkonferenz #GPT5 #GPTOSS #KIAgenten #Monetarisierung #OpenAI #SamAltman #SDK #Spotify #Zillow

  28. [Перевод] Могут ли кодинг-агенты самосовершенствоваться?

    Представьте программиста, который мастерски собирает для себя вспомогательные утилиты, а потом равнодушно отмахивается: «Честно? Мне они не нужны». Именно так повела себя GPT-5 в ходе теста на умение выстраивать собственный набор инструментов для продуктивности. Модель выдала целый арсенал CLI-утилит в духе Unix, но… отказалась ими пользоваться. Почему так случилось и что это говорит о будущем кодинг-агентов — разбираем в статье.

    habr.com/ru/companies/magnus-t

    #искусственный_интеллект #машинное_обучение #самосовершенствование_ИИ #кодингагенты #инструменты_разработчика #GPT5 #claude_opus #ииагенты_для_разработки