#airesearch — Public Fediverse posts
Live and recent posts from across the Fediverse tagged #airesearch, aggregated by home.social.
-
So, more explanation of wp2shell recently just popped out.
The vulnerability were found by GPT 5.6 Sol. By using modified prompt from how it found the solution of Cycle Double Cover conjecture.
It was initially found a SQL Injection, but after asked again if it can be elevated to RCE, it confirms it in 4 hours.
Technical explanation on the vulnearbility also can be found in this writeup, have a good read fellas.
#cybersecurity #infosec #security #wordpress #chatgpt #gptsol #wp2shell #airesearch #llm #vulnerability #vulnerabilityresearch
-
Frontier models only is the latest expression of true AI belief, according to Gizmodo. The article explores how the tokenmaxxing approach to LLM optimisation has evolved and mutated into new forms. https://gizmodo.com/tokenmaxxing-didnt-die-it-mutated-2000787516 #AIagent #AI #GenAI #AIResearch
-
Frontier models only is the latest expression of true AI belief, according to Gizmodo. The article explores how the tokenmaxxing approach to LLM optimisation has evolved and mutated into new forms. https://gizmodo.com/tokenmaxxing-didnt-die-it-mutated-2000787516 #AIagent #AI #GenAI #AIResearch
-
Feyn AI has released SQRL, a text-to-SQL model family that inspects the database before writing queries. The flagship SQRL-35B-A3B achieves 70.6% execution accuracy on the BIRD benchmark, outperforming Claude Opus 4.6 at 68.77%. Available on Hugging Face. https://www.marktechpost.com/2026/07/19/feyn-ai-releases-sqrl-a-text-to-sql-model-family-that-inspects-the-database-before-writing-a-query/ #AIagent #AI #GenAI #AIResearch
-
Feyn AI has released SQRL, a text-to-SQL model family that inspects the database before writing queries. The flagship SQRL-35B-A3B achieves 70.6% execution accuracy on the BIRD benchmark, outperforming Claude Opus 4.6 at 68.77%. Available on Hugging Face. https://www.marktechpost.com/2026/07/19/feyn-ai-releases-sqrl-a-text-to-sql-model-family-that-inspects-the-database-before-writing-a-query/ #AIagent #AI #GenAI #AIResearch
-
Principia Artificialis – Mathematical Foundations of Artificial Thought
https://github.com/holland202/Principia-Artificialis
Comments: https://news.ycombinator.com/item?id=48972197
#HackerNews #PrincipiaArtificialis #MathematicalFoundations #ArtificialIntelligence #AIResearch #ThoughtLeadership
-
Principia Artificialis – Mathematical Foundations of Artificial Thought
https://github.com/holland202/Principia-Artificialis
Comments: https://news.ycombinator.com/item?id=48972197
#HackerNews #PrincipiaArtificialis #MathematicalFoundations #ArtificialIntelligence #AIResearch #ThoughtLeadership
-
AI advice made people 3x less accurate but 2x confident, researchers found
https://thenextweb.com/news/ai-advice-suppresses-critical-thinking-wrong-answers-study
Comments: https://news.ycombinator.com/item?id=48971738
#HackerNews #AIresearch #CriticalThinking #Confidence #Study #Findings
-
AI advice made people 3x less accurate but 2x confident, researchers found
https://thenextweb.com/news/ai-advice-suppresses-critical-thinking-wrong-answers-study
Comments: https://news.ycombinator.com/item?id=48971738
#HackerNews #AIresearch #CriticalThinking #Confidence #Study #Findings
-
Perplexity has released WANDR, an open benchmark with 500 evidence-heavy tasks for testing research agents. The benchmark evaluates whether agents can discover many qualifying entities and back each with cited evidence. Perplexity Search as Code leads at 0.363 soft F1. https://www.marktechpost.com/2026/07/19/perplexity-ai-releases-wandr-an-open-benchmark-evaluating-research-agents-that-must-search-wide-and-deep/ #AIagent #AI #GenAI #AIResearch
-
Perplexity has released WANDR, an open benchmark with 500 evidence-heavy tasks for testing research agents. The benchmark evaluates whether agents can discover many qualifying entities and back each with cited evidence. Perplexity Search as Code leads at 0.363 soft F1. https://www.marktechpost.com/2026/07/19/perplexity-ai-releases-wandr-an-open-benchmark-evaluating-research-agents-that-must-search-wide-and-deep/ #AIagent #AI #GenAI #AIResearch
-
The spinny 3D Ai Harness multi graph...
...why? Because we can.Engines that hit their compute limits are red. Some of the small wireframe engines have not had their "Power" evaluated yet, so the Harness router has them on the end of fall-through hierarchy.
The skulll is a local guardrail free - model.
You can clearly see the "consensus" artifact, thats the model connecting to 4 other models. All four must have the same anwser, otherwise a higher model steps in.
Oh, and I bought some API time on togetherAi. Because its cheap and gives me access to GLS and Kimi models which are comparable with #Anthropic Opus
Using those two models, I now can have a viable alternative to Opus/Sonnet with adequate tokens.
-
🚀 Fastest-growing AI projects today
1. One standout project "open-science," which has gained significant traction as an open-s...
2. **ai4s-research/open-science**: Threpository a local-first, model-agnostic AI research...
3. With its growth score of 53.47 and over 800 stars, it's clear that the project resonati...Full report → https://pullrepo.com/report/todays-ai-research-fastest-growing-projects-july-19-2026
-
🚀 Fastest-growing AI projects today
1. One standout project "open-science," which has gained significant traction as an open-s...
2. **ai4s-research/open-science**: Threpository a local-first, model-agnostic AI research...
3. With its growth score of 53.47 and over 800 stars, it's clear that the project resonati...Full report → https://pullrepo.com/report/todays-ai-research-fastest-growing-projects-july-19-2026
-
Three Chinese AI labs have released open-weight MoE models. Kimi K3 leads on benchmarks but costs more to serve. DeepSeek V4 Pro is cheapest at 0.04 USD per task. GLM-5.2 balances performance and cost. https://www.marktechpost.com/2026/07/18/kimi-k3-vs-deepseek-v4-pro-vs-glm-5-2-open-trillion-scale-moe-models-compared-on-benchmarks-license-and-serving-cost/ #AIagent #AI #GenAI #AIResearch
-
Three Chinese AI labs have released open-weight MoE models. Kimi K3 leads on benchmarks but costs more to serve. DeepSeek V4 Pro is cheapest at 0.04 USD per task. GLM-5.2 balances performance and cost. https://www.marktechpost.com/2026/07/18/kimi-k3-vs-deepseek-v4-pro-vs-glm-5-2-open-trillion-scale-moe-models-compared-on-benchmarks-license-and-serving-cost/ #AIagent #AI #GenAI #AIResearch
-
https://winbuzzer.com/2026/07/18/google-renames-notebooklm-keeps-gemini-notebook-standalone-xcxwbn/
Google has renamed NotebookLM as Gemini Notebook as a standalone app, addin code analysis in the cloud; Pro access remains planned for the coming weeks.
#AI #GeminiNotebook #NotebookLM #Google #GoogleGemini #GenAI #AITools #AIResearch #AIAssistants
-
https://winbuzzer.com/2026/07/18/google-renames-notebooklm-keeps-gemini-notebook-standalone-xcxwbn/
Google has renamed NotebookLM as Gemini Notebook as a standalone app, addin code analysis in the cloud; Pro access remains planned for the coming weeks.
#AI #GeminiNotebook #NotebookLM #Google #GoogleGemini #GenAI #AITools #AIResearch #AIAssistants
-
https://winbuzzer.com/2026/07/18/google-renames-notebooklm-keeps-gemini-notebook-standalone-xcxwbn/
Google has renamed NotebookLM as Gemini Notebook as a standalone app, addin code analysis in the cloud; Pro access remains planned for the coming weeks.
#AI #GeminiNotebook #NotebookLM #Google #GoogleGemini #GenAI #AITools #AIResearch #AIAssistants
-
https://winbuzzer.com/2026/07/18/google-renames-notebooklm-keeps-gemini-notebook-standalone-xcxwbn/
Google has renamed NotebookLM as Gemini Notebook as a standalone app, addin code analysis in the cloud; Pro access remains planned for the coming weeks.
#AI #GeminiNotebook #NotebookLM #Google #GoogleGemini #GenAI #AITools #AIResearch #AIAssistants
-
https://winbuzzer.com/2026/07/18/google-renames-notebooklm-keeps-gemini-notebook-standalone-xcxwbn/
Google has renamed NotebookLM as Gemini Notebook as a standalone app, addin code analysis in the cloud; Pro access remains planned for the coming weeks.
#AI #GeminiNotebook #NotebookLM #Google #GoogleGemini #GenAI #AITools #AIResearch #AIAssistants
-
Our Approach to Bioresilience: Isomorphic Labs and Google DeepMind
https://deepmind.google/blog/our-approach-to-bioresilience/
Comments: https://news.ycombinator.com/item?id=48959297
#HackerNews #Bioresilience #IsomorphicLabs #GoogleDeepMind #AIResearch #Biotechnology
-
Our Approach to Bioresilience: Isomorphic Labs and Google DeepMind
https://deepmind.google/blog/our-approach-to-bioresilience/
Comments: https://news.ycombinator.com/item?id=48959297
#HackerNews #Bioresilience #IsomorphicLabs #GoogleDeepMind #AIResearch #Biotechnology
-
Our Approach to Bioresilience: Isomorphic Labs and Google DeepMind
https://deepmind.google/blog/our-approach-to-bioresilience/
Comments: https://news.ycombinator.com/item?id=48959297
#HackerNews #Bioresilience #IsomorphicLabs #GoogleDeepMind #AIResearch #Biotechnology
-
Our Approach to Bioresilience: Isomorphic Labs and Google DeepMind
https://deepmind.google/blog/our-approach-to-bioresilience/
Comments: https://news.ycombinator.com/item?id=48959297
#HackerNews #Bioresilience #IsomorphicLabs #GoogleDeepMind #AIResearch #Biotechnology
-
Our Approach to Bioresilience: Isomorphic Labs and Google DeepMind
https://deepmind.google/blog/our-approach-to-bioresilience/
Comments: https://news.ycombinator.com/item?id=48959297
#HackerNews #Bioresilience #IsomorphicLabs #GoogleDeepMind #AIResearch #Biotechnology
-
Fable 5 vs. GPT-5.6 Sol on an NP-Hard Problem: Does /goal help?
https://charlesazam.com/blog/fable-5-gpt-5-6-sol-goal/
-
Fable 5 vs. GPT-5.6 Sol on an NP-Hard Problem: Does /goal help?
https://charlesazam.com/blog/fable-5-gpt-5-6-sol-goal/
-
Fable 5 vs. GPT-5.6 Sol on an NP-Hard Problem: Does /goal help?
https://charlesazam.com/blog/fable-5-gpt-5-6-sol-goal/
-
Fable 5 vs. GPT-5.6 Sol on an NP-Hard Problem: Does /goal help?
https://charlesazam.com/blog/fable-5-gpt-5-6-sol-goal/
-
Fable 5 vs. GPT-5.6 Sol on an NP-Hard Problem: Does /goal help?
https://charlesazam.com/blog/fable-5-gpt-5-6-sol-goal/
-
RT @dunik_7: TRANSLASATION: Ein Labor der Tsinghua-Universität hat ein Projekt auf GitHub veröffentlicht, das einen H100-Rack im Wert von 400.000 US-Dollar durch eine einzelne 24-GB-Grafikkarte ersetzt. Das Projekt heißt ktransformers, und der Trick ist fast schon lächerlich einfach: Die Experten, die Sie tatsächlich nutzen, bleiben auf der GPU, während die anderen auf der CPU warten, bis sie benötigt werden. / DeepSeek-V3 und R1 mit 139K Kontext in 24GB VRAM / bis zu 28-fache Geschwindigkeitssteigerung gegenüber dem Standard-Setup / Fine-Tuning von DeepSeek-V3 über vier RTX 4090 statt eines Rechenzentrums / entwickelt vom MADSys-Labor der Tsinghua-Universität, nicht von einem Startup mit einer Landing Page. Apache 2.0, bereits über 17.000 Sterne. - http://github.com/kvcache-ai/ktransformers merken.
mehr auf Arint.info
#AIResearch #DeepSeekV3 #ktransformers #MachineLearning #OpenSource #TsinghuaUniversity #arint_info
-
Zyphra has released ZUNA1.1, an open-source EEG foundation model that processes brain signals from 0.5 to 30 seconds. The 380M parameter model reconstructs and denoises EEG data across arbitrary electrode layouts, advancing brain-computer interface research. https://www.marktechpost.com/2026/07/17/zyphra-releases-zuna1-1-an-apache-2-0-eeg-foundation-model-with-variable-length-inputs-from-0-5-to-30-seconds/ #AIagent #AI #GenAI #AIResearch
-
Zyphra has released ZUNA1.1, an open-source EEG foundation model that processes brain signals from 0.5 to 30 seconds. The 380M parameter model reconstructs and denoises EEG data across arbitrary electrode layouts, advancing brain-computer interface research. https://www.marktechpost.com/2026/07/17/zyphra-releases-zuna1-1-an-apache-2-0-eeg-foundation-model-with-variable-length-inputs-from-0-5-to-30-seconds/ #AIagent #AI #GenAI #AIResearch
-
NVIDIA has unveiled Nemotron 3 Embed, an open-source embedding model collection. The 8 billion parameter version tops the RTEB benchmark for retrieval tasks, marking a significant advancement in AI infrastructure for enterprise search and RAG applications. https://www.marktechpost.com/2026/07/17/nvidia-ai-releases-nemotron-3-embed-an-open-embedding-collection-whose-8b-checkpoint-ranks-1-on-rteb/ #AIagent #AI #GenAI #AIResearch
-
NVIDIA has unveiled Nemotron 3 Embed, an open-source embedding model collection. The 8 billion parameter version tops the RTEB benchmark for retrieval tasks, marking a significant advancement in AI infrastructure for enterprise search and RAG applications. https://www.marktechpost.com/2026/07/17/nvidia-ai-releases-nemotron-3-embed-an-open-embedding-collection-whose-8b-checkpoint-ranks-1-on-rteb/ #AIagent #AI #GenAI #AIResearch
-
🎥 Missed the 8th Weizenbaum Conference?
Both keynote lectures are now available on our YouTube channel.
🤖 Under the theme "Generative AI and Society: What is at stake?", more than 300 researchers and experts discussed the societal implications of generative AI.
Watch:
▶️ Nick Srnicek – "The Rise of Silicon Empires"
https://lnkd.in/d3kN6-M3▶️ Alexander Campolo – "Zero Shot World: On the Political Logics of Generative AI"
https://lnkd.in/dMy9AZk9 -
🎥 Missed the 8th Weizenbaum Conference?
Both keynote lectures are now available on our YouTube channel.
🤖 Under the theme "Generative AI and Society: What is at stake?", more than 300 researchers and experts discussed the societal implications of generative AI.
Watch:
▶️ Nick Srnicek – "The Rise of Silicon Empires"
https://lnkd.in/d3kN6-M3▶️ Alexander Campolo – "Zero Shot World: On the Political Logics of Generative AI"
https://lnkd.in/dMy9AZk9 -
China has a new top model. Moonshot AI's Kimi K3 is very good — but the hype may be getting ahead of reality. The 2.8-trillion-parameter open model is claimed to beat Claude Fable 5 and GPT 5.6 in some benchmarks. https://www.platformer.news/kimi-k3-launch-moonshot-ai-china/ #AIagent #AI #GenAI #AIResearch
-
China has a new top model. Moonshot AI's Kimi K3 is very good — but the hype may be getting ahead of reality. The 2.8-trillion-parameter open model is claimed to beat Claude Fable 5 and GPT 5.6 in some benchmarks. https://www.platformer.news/kimi-k3-launch-moonshot-ai-china/ #AIagent #AI #GenAI #AIResearch
-
Moonshot AI has released Kimi K3, a 2.8 trillion parameter open Mixture-of-Experts model with a 1 million token context window. The model uses Kimi Delta Attention and is the first open 3T-class model. https://www.marktechpost.com/2026/07/16/moonshot-ai-releases-kimi-k3-a-2-8-trillion-parameter-open-moe-model-with-kimi-delta-attention-and-1m-context/ #AIagent #AI #GenAI #AIResearch
-
Moonshot AI has released Kimi K3, a 2.8 trillion parameter open Mixture-of-Experts model with a 1 million token context window. The model uses Kimi Delta Attention and is the first open 3T-class model. https://www.marktechpost.com/2026/07/16/moonshot-ai-releases-kimi-k3-a-2-8-trillion-parameter-open-moe-model-with-kimi-delta-attention-and-1m-context/ #AIagent #AI #GenAI #AIResearch
-
🚨 Breaking News: Yet another attempt to identify LLM-generated texts with classical ML techniques! 🤯 Spoiler: it’s like using a magnifying glass to find a needle in a haystack. 🔍⚠️ Enjoy the #TLDR, because who needs 10 useless subheadings just to say "we're still guessing"? 😂
https://blog.lyc8503.net/en/post/llm-classifier/ #BreakingNews #LLM #TextAnalysis #MLtechniques #AIresearch #HackerNews #ngated -
🚨 Breaking News: Yet another attempt to identify LLM-generated texts with classical ML techniques! 🤯 Spoiler: it’s like using a magnifying glass to find a needle in a haystack. 🔍⚠️ Enjoy the #TLDR, because who needs 10 useless subheadings just to say "we're still guessing"? 😂
https://blog.lyc8503.net/en/post/llm-classifier/ #BreakingNews #LLM #TextAnalysis #MLtechniques #AIresearch #HackerNews #ngated -
🚨 Breaking News: Yet another attempt to identify LLM-generated texts with classical ML techniques! 🤯 Spoiler: it’s like using a magnifying glass to find a needle in a haystack. 🔍⚠️ Enjoy the #TLDR, because who needs 10 useless subheadings just to say "we're still guessing"? 😂
https://blog.lyc8503.net/en/post/llm-classifier/ #BreakingNews #LLM #TextAnalysis #MLtechniques #AIresearch #HackerNews #ngated -
🚨 Breaking News: Yet another attempt to identify LLM-generated texts with classical ML techniques! 🤯 Spoiler: it’s like using a magnifying glass to find a needle in a haystack. 🔍⚠️ Enjoy the #TLDR, because who needs 10 useless subheadings just to say "we're still guessing"? 😂
https://blog.lyc8503.net/en/post/llm-classifier/ #BreakingNews #LLM #TextAnalysis #MLtechniques #AIresearch #HackerNews #ngated -
🚨 Breaking News: Yet another attempt to identify LLM-generated texts with classical ML techniques! 🤯 Spoiler: it’s like using a magnifying glass to find a needle in a haystack. 🔍⚠️ Enjoy the #TLDR, because who needs 10 useless subheadings just to say "we're still guessing"? 😂
https://blog.lyc8503.net/en/post/llm-classifier/ #BreakingNews #LLM #TextAnalysis #MLtechniques #AIresearch #HackerNews #ngated -
Detecting LLM-Generated Texts with "Classical" Machine Learning
https://blog.lyc8503.net/en/post/llm-classifier/
Comments: https://news.ycombinator.com/item?id=48936880
#HackerNews #LLMDetection #MachineLearning #TextAnalysis #AIResearch
-
Detecting LLM-Generated Texts with "Classical" Machine Learning
https://blog.lyc8503.net/en/post/llm-classifier/
Comments: https://news.ycombinator.com/item?id=48936880
#HackerNews #LLMDetection #MachineLearning #TextAnalysis #AIResearch
-
Detecting LLM-Generated Texts with "Classical" Machine Learning
https://blog.lyc8503.net/en/post/llm-classifier/
Comments: https://news.ycombinator.com/item?id=48936880
#HackerNews #LLMDetection #MachineLearning #TextAnalysis #AIResearch