#reasoningmodels — Public Fediverse posts
Live and recent posts from across the Fediverse tagged #reasoningmodels, aggregated by home.social.
-
https://www.europesays.com/people/164749/ Sundar Pichai Just Said Something About Google’s Gemini 4—and Sam Altman, Elon Musk Should Care #AlphabetInc #Google #GoogleCloud #RealEstate #ReasoningModels #SundarPichai
-
MeetKai and Albania Announce Joint Venture to Advance Sovereign AI https://www.byteseu.com/2172644/ #AKSHI #Albania #AlbanianLanguage #DigitalCapability #JointVenture #NationalAgencyForInformationSociety #PublicServices #ReasoningModels #SovereignAICompany
-
OpenAI claims it solved an 80-year-old math problem — for real this time
OpenAI claims its new reasoning model has produced an original mathematical proof disproving a famous unsolved conjecture in…
#NewsBeep #News #US #USA #UnitedStates #UnitedStatesOfAmerica #Artificialintelligence #AI #ArtificialIntelligence #ChatGPT #erdosproblems #OpenAI #reasoningmodels #Technology
https://www.newsbeep.com/us/655490/ -
OpenAI claims it solved an 80-year-old math problem — for real this time
OpenAI claims its new reasoning model has produced an original mathematical proof disproving a famous unsolved conjecture in…
#NewsBeep #News #US #USA #UnitedStates #UnitedStatesOfAmerica #Artificialintelligence #AI #ArtificialIntelligence #ChatGPT #erdosproblems #OpenAI #reasoningmodels #Technology
https://www.newsbeep.com/us/655490/ -
https://www.europesays.com/uk/974400/ OpenAI claims it solved an 80-year-old math problem — for real this time #Business #chatgpt #ErdosProblems #openai #ReasoningModels #UK #UnitedKingdom
-
https://www.europesays.com/people/19079/ Nvidia Stock Investors Just Got Great News From CEO Jensen Huang. It’s Time to Buy. #AIInfrastructure #DataCenter #JensenHuang #Nvidia #ReasoningModels
-
Arcee AI released Trinity-Large-Thinking, a 400B parameter open-source reasoning model that scores within 2 points of Claude Opus on PinchBench while costing 96% less at $0.90 per million tokens. Uses sparse architecture activating only 13B parameters per token. Trained for $20M by 30-person team. #OpenSource #AI #ReasoningModels
-
Arcee AI released Trinity-Large-Thinking, a 400B parameter open-source reasoning model that scores within 2 points of Claude Opus on PinchBench while costing 96% less at $0.90 per million tokens. Uses sparse architecture activating only 13B parameters per token. Trained for $20M by 30-person team. #OpenSource #AI #ReasoningModels
-
Arcee AI released Trinity-Large-Thinking, a 400B parameter open-source reasoning model that scores within 2 points of Claude Opus on PinchBench while costing 96% less at $0.90 per million tokens. Uses sparse architecture activating only 13B parameters per token. Trained for $20M by 30-person team. #OpenSource #AI #ReasoningModels
-
Arcee AI released Trinity-Large-Thinking, a 400B parameter open-source reasoning model that scores within 2 points of Claude Opus on PinchBench while costing 96% less at $0.90 per million tokens. Uses sparse architecture activating only 13B parameters per token. Trained for $20M by 30-person team. #OpenSource #AI #ReasoningModels
-
Arcee AI released Trinity-Large-Thinking, a 400B parameter open-source reasoning model that scores within 2 points of Claude Opus on PinchBench while costing 96% less at $0.90 per million tokens. Uses sparse architecture activating only 13B parameters per token. Trained for $20M by 30-person team. #OpenSource #AI #ReasoningModels
-
xAI’s co‑founder exits keep coming, while Lambda outlines a 2025 shift toward bigger context windows, multimodal reasoning models and open‑source inference for AI production. What could this mean for the future of machine learning? Read on for the full story. #AIProduction #ReasoningModels #MultimodalAI #OpenSourceInference
🔗 https://aidailypost.com/news/xai-co-founder-departures-persist-lambda-outlines-2025-ai-production
-
xAI’s co‑founder exits keep coming, while Lambda outlines a 2025 shift toward bigger context windows, multimodal reasoning models and open‑source inference for AI production. What could this mean for the future of machine learning? Read on for the full story. #AIProduction #ReasoningModels #MultimodalAI #OpenSourceInference
🔗 https://aidailypost.com/news/xai-co-founder-departures-persist-lambda-outlines-2025-ai-production
-
xAI’s co‑founder exits keep coming, while Lambda outlines a 2025 shift toward bigger context windows, multimodal reasoning models and open‑source inference for AI production. What could this mean for the future of machine learning? Read on for the full story. #AIProduction #ReasoningModels #MultimodalAI #OpenSourceInference
🔗 https://aidailypost.com/news/xai-co-founder-departures-persist-lambda-outlines-2025-ai-production
-
2025 saw significant advancements in #LLMs, particularly in the areas of #reasoning and #agent based systems. #Reasoningmodels, capable of breaking down #complextasks and utilising tools, revolutionised #coding and #search. The year witnessed the rise of #codingagents, exemplified by #ClaudeCode, which can autonomously write, execute, and refine code. https://simonwillison.net/2025/Dec/31/the-year-in-llms/?eicker.news #tech #media #news
-
2025 saw significant advancements in #LLMs, particularly in the areas of #reasoning and #agent based systems. #Reasoningmodels, capable of breaking down #complextasks and utilising tools, revolutionised #coding and #search. The year witnessed the rise of #codingagents, exemplified by #ClaudeCode, which can autonomously write, execute, and refine code. https://simonwillison.net/2025/Dec/31/the-year-in-llms/?eicker.news #tech #media #news
-
2025 saw significant advancements in #LLMs, particularly in the areas of #reasoning and #agent based systems. #Reasoningmodels, capable of breaking down #complextasks and utilising tools, revolutionised #coding and #search. The year witnessed the rise of #codingagents, exemplified by #ClaudeCode, which can autonomously write, execute, and refine code. https://simonwillison.net/2025/Dec/31/the-year-in-llms/?eicker.news #tech #media #news
-
2025 saw significant advancements in #LLMs, particularly in the areas of #reasoning and #agent based systems. #Reasoningmodels, capable of breaking down #complextasks and utilising tools, revolutionised #coding and #search. The year witnessed the rise of #codingagents, exemplified by #ClaudeCode, which can autonomously write, execute, and refine code. https://simonwillison.net/2025/Dec/31/the-year-in-llms/?eicker.news #tech #media #news
-
2025 saw significant advancements in #LLMs, particularly in the areas of #reasoning and #agent based systems. #Reasoningmodels, capable of breaking down #complextasks and utilising tools, revolutionised #coding and #search. The year witnessed the rise of #codingagents, exemplified by #ClaudeCode, which can autonomously write, execute, and refine code. https://simonwillison.net/2025/Dec/31/the-year-in-llms/?eicker.news #tech #media #news
-
OpenAI: GPT-5 Thinking Models Are The Most "Monitarable" Models To Date
#AI #OpenAI #AISafety #LLM #MachineLearning #GPT5 #DeepMind #AIResearch #ChainOfThought #Monitorability #AIAlignment #ReasoningModels
-
OpenAI: GPT-5 Thinking Models Are The Most "Monitarable" Models To Date
#AI #OpenAI #AISafety #LLM #MachineLearning #GPT5 #DeepMind #AIResearch #ChainOfThought #Monitorability #AIAlignment #ReasoningModels
-
OpenAI: GPT-5 Thinking Models Are The Most "Monitarable" Models To Date
#AI #OpenAI #AISafety #LLM #MachineLearning #GPT5 #DeepMind #AIResearch #ChainOfThought #Monitorability #AIAlignment #ReasoningModels
-
OpenAI: GPT-5 Thinking Models Are The Most "Monitarable" Models To Date
#AI #OpenAI #AISafety #LLM #MachineLearning #GPT5 #DeepMind #AIResearch #ChainOfThought #Monitorability #AIAlignment #ReasoningModels
-
OpenAI: GPT-5 Thinking Models Are The Most "Monitarable" Models To Date
#AI #OpenAI #AISafety #LLM #MachineLearning #GPT5 #DeepMind #AIResearch #ChainOfThought #Monitorability #AIAlignment #ReasoningModels
-
FINE-TUNING Qwen3 VỚI "THINKING MODE" KHÓ KHĂN TRONG LẬP LUẬN. Tài liệu hướng dẫn tạo tập dữ liệu "giải thích" (thinking) chưa rõ ràng khiến việc huấn luyện mô hình gặp trục trặc. Ai có kinh nghiệm hoặc tài liệu về kiến thức này chia sẻ giúp #AI #MachineLearning #LậpLý #MôHìnhQwen #ReasoningModels #KnowledgeInjection
*(Tóm tắt: Người dùng gặp khó khăn khi tinh chỉnh Qwen3 để bổ sung kiến thức Vật lý nhờ "thinking mode". Cố tạo dữ liệu giải thích bằng Qwen3 dẫn đến hiệu suất giảm. Cần chia sẻ
-
https://www.europesays.com/ie/190264/ The cost of thinking | MIT News #AI #AIThinking #AndreaGregorDeVarda #ArtificialIntelligence #ArtificialNeuralNetworks #ArtificialIntelligence #ComputationalNeuroscience #Éire #EvelinaFedorenko #FedorenkoLab #IE #InternalMonologues #Ireland #K.LisaYangICoNCenter #LargeLanguageModels(LLMs) #MITMcgovernInstitute #ReasoningModels #ReinforcementLearning #Technology
-
New AI reasoning models built as neural networks are showing striking convergence across diverse training sets. Researchers say this hints at emergent structure in how machines learn to reason, opening fresh avenues for open‑source computational tools. Dive into the findings and see why this could reshape our approach to artificial intelligence. #AI #NeuralNetworks #ReasoningModels #Convergence
🔗 https://aidailypost.com/news/new-ai-reasoning-models-built-neural-networks-show-striking
-
New AI reasoning models built as neural networks are showing striking convergence across diverse training sets. Researchers say this hints at emergent structure in how machines learn to reason, opening fresh avenues for open‑source computational tools. Dive into the findings and see why this could reshape our approach to artificial intelligence. #AI #NeuralNetworks #ReasoningModels #Convergence
🔗 https://aidailypost.com/news/new-ai-reasoning-models-built-neural-networks-show-striking
-
Reasoning Models Reason Well, Until They Don't
https://arxiv.org/abs/2510.22371
#HackerNews #ReasoningModels #ReasonWell #AIResearch #MachineLearning #HackerNews
-
Reasoning Models Reason Well, Until They Don't
https://arxiv.org/abs/2510.22371
#HackerNews #ReasoningModels #ReasonWell #AIResearch #MachineLearning #HackerNews
-
Reasoning Models Reason Well, Until They Don't
https://arxiv.org/abs/2510.22371
#HackerNews #ReasoningModels #ReasonWell #AIResearch #MachineLearning #HackerNews
-
Reasoning Models Reason Well, Until They Don't
https://arxiv.org/abs/2510.22371
#HackerNews #ReasoningModels #ReasonWell #AIResearch #MachineLearning #HackerNews
-
Reasoning Models Reason Well, Until They Don't
https://arxiv.org/abs/2510.22371
#HackerNews #ReasoningModels #ReasonWell #AIResearch #MachineLearning #HackerNews
-
Phân tích METR-Horizon cho thấy các mô hình AI có khả năng suy luận (từ T9/2024) vượt trội: hiệu suất tăng 2.2 lần và khả năng mở rộng nhanh hơn 37%. Điều này giúp AI hoàn thành các nhiệm vụ phức tạp, đòi hỏi thời gian dài một cách tự chủ, nhanh hơn đáng kể so với trước đây.
#AI #ReasoningModels #MachineLearning #TechNews #METRAnalysis #TríTuệNhânTạo #MôHìnhSuyLuận #CôngNghệ
-
"The point is that with each advance in AI, new hurdles become apparent; when one missing aspect of “intelligence” is filled in, we find ourselves bumping up against another gap. When I speculated about GPT-5 last year, it didn’t occur to me to question whether it would know how to set priorities, because the models of the time weren’t even capable enough for that to be a limiting factor. In a post from November, AI is Racing Forward – on a Very Long Road, I wrote:
…the real challenges may be things that we can’t easily anticipate right now, weaknesses that we will only start to put our finger on when we observe [future models] performing astonishing feats and yet somehow still not being able to write that tightly-plotted novel.
In April 2024, it seemed like agentic AI was going to be the next big thing. The ensuing 16 months have brought enormous progress on many fronts, but very little progress on real-world agency. With projects like AI Village shining a light on the profound weakness of current AI agents, I think robust real-world capability is still years away."
https://secondthoughts.ai/p/gpt-5-the-case-of-the-missing-agent
#AI #GenerativeAI #LLMs #Chatbots #AIAgents #AgenticAI #ReasoningModels
-
"The point is that with each advance in AI, new hurdles become apparent; when one missing aspect of “intelligence” is filled in, we find ourselves bumping up against another gap. When I speculated about GPT-5 last year, it didn’t occur to me to question whether it would know how to set priorities, because the models of the time weren’t even capable enough for that to be a limiting factor. In a post from November, AI is Racing Forward – on a Very Long Road, I wrote:
…the real challenges may be things that we can’t easily anticipate right now, weaknesses that we will only start to put our finger on when we observe [future models] performing astonishing feats and yet somehow still not being able to write that tightly-plotted novel.
In April 2024, it seemed like agentic AI was going to be the next big thing. The ensuing 16 months have brought enormous progress on many fronts, but very little progress on real-world agency. With projects like AI Village shining a light on the profound weakness of current AI agents, I think robust real-world capability is still years away."
https://secondthoughts.ai/p/gpt-5-the-case-of-the-missing-agent
#AI #GenerativeAI #LLMs #Chatbots #AIAgents #AgenticAI #ReasoningModels
-
"The point is that with each advance in AI, new hurdles become apparent; when one missing aspect of “intelligence” is filled in, we find ourselves bumping up against another gap. When I speculated about GPT-5 last year, it didn’t occur to me to question whether it would know how to set priorities, because the models of the time weren’t even capable enough for that to be a limiting factor. In a post from November, AI is Racing Forward – on a Very Long Road, I wrote:
…the real challenges may be things that we can’t easily anticipate right now, weaknesses that we will only start to put our finger on when we observe [future models] performing astonishing feats and yet somehow still not being able to write that tightly-plotted novel.
In April 2024, it seemed like agentic AI was going to be the next big thing. The ensuing 16 months have brought enormous progress on many fronts, but very little progress on real-world agency. With projects like AI Village shining a light on the profound weakness of current AI agents, I think robust real-world capability is still years away."
https://secondthoughts.ai/p/gpt-5-the-case-of-the-missing-agent
#AI #GenerativeAI #LLMs #Chatbots #AIAgents #AgenticAI #ReasoningModels
-
"The point is that with each advance in AI, new hurdles become apparent; when one missing aspect of “intelligence” is filled in, we find ourselves bumping up against another gap. When I speculated about GPT-5 last year, it didn’t occur to me to question whether it would know how to set priorities, because the models of the time weren’t even capable enough for that to be a limiting factor. In a post from November, AI is Racing Forward – on a Very Long Road, I wrote:
…the real challenges may be things that we can’t easily anticipate right now, weaknesses that we will only start to put our finger on when we observe [future models] performing astonishing feats and yet somehow still not being able to write that tightly-plotted novel.
In April 2024, it seemed like agentic AI was going to be the next big thing. The ensuing 16 months have brought enormous progress on many fronts, but very little progress on real-world agency. With projects like AI Village shining a light on the profound weakness of current AI agents, I think robust real-world capability is still years away."
https://secondthoughts.ai/p/gpt-5-the-case-of-the-missing-agent
#AI #GenerativeAI #LLMs #Chatbots #AIAgents #AgenticAI #ReasoningModels
-
"The point is that with each advance in AI, new hurdles become apparent; when one missing aspect of “intelligence” is filled in, we find ourselves bumping up against another gap. When I speculated about GPT-5 last year, it didn’t occur to me to question whether it would know how to set priorities, because the models of the time weren’t even capable enough for that to be a limiting factor. In a post from November, AI is Racing Forward – on a Very Long Road, I wrote:
…the real challenges may be things that we can’t easily anticipate right now, weaknesses that we will only start to put our finger on when we observe [future models] performing astonishing feats and yet somehow still not being able to write that tightly-plotted novel.
In April 2024, it seemed like agentic AI was going to be the next big thing. The ensuing 16 months have brought enormous progress on many fronts, but very little progress on real-world agency. With projects like AI Village shining a light on the profound weakness of current AI agents, I think robust real-world capability is still years away."
https://secondthoughts.ai/p/gpt-5-the-case-of-the-missing-agent
#AI #GenerativeAI #LLMs #Chatbots #AIAgents #AgenticAI #ReasoningModels
-
🧠 What if you could tell AI how much to think before answering?
Seed-OSS 36B gives builders a thinking budget knob + 512K context window—control depth vs speed like never before. ⚡👉 See how it changes product SLAs, costs, and user experience:
https://medium.com/@rogt.x1997/seed-oss-36b-a-tweakable-reasoning-engine-for-long-context-work-66aa05a72548#AI #ReasoningModels #LongContext
https://medium.com/@rogt.x1997/seed-oss-36b-a-tweakable-reasoning-engine-for-long-context-work-66aa05a72548 -
LLMs’ “simulated reasoning” abilities are a “brittle mirage,” researchers find
Chain-of-thought AI "degrades significantly" when asked to generalize beyond training.
-
LLMs’ “simulated reasoning” abilities are a “brittle mirage,” researchers find
Chain-of-thought AI "degrades significantly" when asked to generalize beyond training.
-
Seven so-called "replies" to Apple's paper on reasoning models, or as I like to call them, seven exercises in missing the point entirely. 📚🤦♂️ It's almost like a bad magic trick: look over here at these rebuttals while we pretend the original issue just vanishes! 🎩✨
https://garymarcus.substack.com/p/seven-replies-to-the-viral-apple #AppleReplies #ReasoningModels #MissingThePoint #BadMagicTrick #TechCritique #HackerNews #ngated -
Seven so-called "replies" to Apple's paper on reasoning models, or as I like to call them, seven exercises in missing the point entirely. 📚🤦♂️ It's almost like a bad magic trick: look over here at these rebuttals while we pretend the original issue just vanishes! 🎩✨
https://garymarcus.substack.com/p/seven-replies-to-the-viral-apple #AppleReplies #ReasoningModels #MissingThePoint #BadMagicTrick #TechCritique #HackerNews #ngated -
Seven so-called "replies" to Apple's paper on reasoning models, or as I like to call them, seven exercises in missing the point entirely. 📚🤦♂️ It's almost like a bad magic trick: look over here at these rebuttals while we pretend the original issue just vanishes! 🎩✨
https://garymarcus.substack.com/p/seven-replies-to-the-viral-apple #AppleReplies #ReasoningModels #MissingThePoint #BadMagicTrick #TechCritique #HackerNews #ngated -
Seven so-called "replies" to Apple's paper on reasoning models, or as I like to call them, seven exercises in missing the point entirely. 📚🤦♂️ It's almost like a bad magic trick: look over here at these rebuttals while we pretend the original issue just vanishes! 🎩✨
https://garymarcus.substack.com/p/seven-replies-to-the-viral-apple #AppleReplies #ReasoningModels #MissingThePoint #BadMagicTrick #TechCritique #HackerNews #ngated -
OpenAI Releases new o3-Pro AI Model: A High-Stakes Bet on AI Reliability
#AI #OpenAI #o3pro #LLM #TechNews #ReasoningModels #ChatGPT #EnterpriseAI #AIEthics
-
OpenAI Releases new o3-Pro AI Model: A High-Stakes Bet on AI Reliability
#AI #OpenAI #o3pro #LLM #TechNews #ReasoningModels #ChatGPT #EnterpriseAI #AIEthics
-
OpenAI Releases new o3-Pro AI Model: A High-Stakes Bet on AI Reliability
#AI #OpenAI #o3pro #LLM #TechNews #ReasoningModels #ChatGPT #EnterpriseAI #AIEthics
-
OpenAI Releases new o3-Pro AI Model: A High-Stakes Bet on AI Reliability
#AI #OpenAI #o3pro #LLM #TechNews #ReasoningModels #ChatGPT #EnterpriseAI #AIEthics
-
OpenAI Releases new o3-Pro AI Model: A High-Stakes Bet on AI Reliability
#AI #OpenAI #o3pro #LLM #TechNews #ReasoningModels #ChatGPT #EnterpriseAI #AIEthics
-
The Illusion of Thinking: Strengths and Limitations of Reasoning Models
https://machinelearning.apple.com/research/illusion-of-thinking
#HackerNews #IllusionOfThinking #ReasoningModels #StrengthsAndLimitations #AIResearch #MachineLearning
-
The Illusion of Thinking: Strengths and Limitations of Reasoning Models
https://machinelearning.apple.com/research/illusion-of-thinking
#HackerNews #IllusionOfThinking #ReasoningModels #StrengthsAndLimitations #AIResearch #MachineLearning
-
The Illusion of Thinking: Strengths and Limitations of Reasoning Models
https://machinelearning.apple.com/research/illusion-of-thinking
#HackerNews #IllusionOfThinking #ReasoningModels #StrengthsAndLimitations #AIResearch #MachineLearning
-
The Illusion of Thinking: Strengths and Limitations of Reasoning Models
https://machinelearning.apple.com/research/illusion-of-thinking
#HackerNews #IllusionOfThinking #ReasoningModels #StrengthsAndLimitations #AIResearch #MachineLearning
-
The Illusion of Thinking: Strengths and Limitations of Reasoning Models
https://machinelearning.apple.com/research/illusion-of-thinking
#HackerNews #IllusionOfThinking #ReasoningModels #StrengthsAndLimitations #AIResearch #MachineLearning
-
No more guessing games! 🕵️♂️ #ollama's new 'think' feature cleanly separates the model's internal thinking from the content. Easy to enable - just 'think': true in your API request. #AIdevelopment #ReasoningModels https://youtu.be/yBD598s5g8c
-
No more guessing games! 🕵️♂️ #ollama's new 'think' feature cleanly separates the model's internal thinking from the content. Easy to enable - just 'think': true in your API request. #AIdevelopment #ReasoningModels https://youtu.be/yBD598s5g8c
-
No more guessing games! 🕵️♂️ #ollama's new 'think' feature cleanly separates the model's internal thinking from the content. Easy to enable - just 'think': true in your API request. #AIdevelopment #ReasoningModels https://youtu.be/yBD598s5g8c
-
No more guessing games! 🕵️♂️ #ollama's new 'think' feature cleanly separates the model's internal thinking from the content. Easy to enable - just 'think': true in your API request. #AIdevelopment #ReasoningModels https://youtu.be/yBD598s5g8c