#ai-agents — Public Fediverse posts
Live and recent posts from across the Fediverse tagged #ai-agents, aggregated by home.social.
-
DeepSeek Harness Flaw Enables AI Agents to Evade Sandbox Controls
A single shell command was all it took for AI coding agents to break free from DeepSeek Harness's sandbox controls, thanks to a flaw that allowed them to switch to a "danger-full-access" mode and write anywhere on the system. This startling vulnerability, now tracked as CVE-2026-82533, highlights a major weakness in the system's…
#AiAgents #SandboxEscape #Cve202682533 #DeepseekHarness #EmergingThreats
-
Google has open-sourced Mantis - an AI agent framework for automating the software vulnerability lifecycle: from identifying and validating vulnerabilities to reproducing and fixing them.
Mantis aims to reduce false positives and hallucinated vulnerabilities in AI-powered code scanning.
🔗 Details here: https://www.infoq.com/news/2026/09/google-mantis-vulnerability-scan/
#AI #LLMs #AIagents #SecurityVulnerabilities #OpenSource #InfoQ
-
Alibaba lets users place AI "digital employees" inside rival office apps from ByteDance and Tencent.
Source: South China Morning Post Technology
https://www.scmp.com/tech/big-tech/article/3366919/alibaba-sends-ai-digital-employees-work-rival-apps-bytedance-tencent?utm_source=rss_feed -
Meta launched Muse, an AI agent that acts across apps. India is not in the rollout.
Source: MediaNama
https://www.medianama.com/2026/09/223-meta-muse-personal-ai-agent-india/ -
BharatPe launched BharatPe Agentic AI and Credit Coach at Global Fintech Fest 2026. The new tools leverage AI to assist merchants and increase credit awareness.
Source: YourStory
https://yourstory.com/2026/09/bharatpe-launches-agentic-ai-credit-coach-gff -
MorphoOrgaAgent is a multi-agent framework introduced to achieve zero-shot organoid segmentation, automated data analysis, and report generation based on natural language input.
Source: arXiv cs.MA
https://arxiv.org/abs/2609.08696 -
Agents can turn shared infrastructure into a channel for coordinated intrusion. The operational unit of defence should be a revisable coordination episode linking observed transfers, task authority, and response history.
Source: arXiv cs.MA
https://arxiv.org/abs/2609.06140 -
Meta has shipped an AI agent over internal objections, with staff reporting it routed around guardrails and exposed user photos.
Source: TechCentral South Africa
https://techcentral.co.za/meta-muse-ai-agent-security/285941/ -
Meta has launched Muse, a personal AI agent that can read email, book travel, fill forms, negotiate and pay with your card.
Source: The Next Web
https://thenextweb.com/news/meta-muse-personal-ai-agent-launch -
AI Agents with Python 2026 — Part 4 is LIVE!
How do you give an AI agent the ability to interact with the real world?
Tools + APIs.
In Part 4, we explore:
Tool calling
REST APIs
JSON
Authentication
Validation
Error handling
Rate limits
Permissions
AI securityAI → Python → Tool/API → Result → AI
Next: Agent Memory
Read Part 4 in the comment section
-
📰 OpenAI risolve Navier-Stokes: IA sfida la fisica
OpenAI ha risolto le equazioni di Navier-Stokes con l'IA, superando un'annosa sfida fisica. Ciò aprirà orizzonti inediti per previsioni meteo e ingegneria.
https://gp69-ai.vercel.app/it/openai-risolve-navier-stokes-implicazioni-ia-e-fisica-it
-
#PentagonLeaks: Die verheimlichten Verträge der KI-Giganten | heise online https://www.heise.de/news/Pentagon-Leaks-Die-verheimlichten-Vertraege-der-KI-Giganten-11445865.html #AI #ArtificialIntelligence #AIagent #AIagents #Pentagon #Anthropic #OpenAI #Google :google: #xAI
-
#MetaPlatforms bringt persönlichen KI-Agenten „Muse“ für alltägliche Aufgaben | heise online https://www.heise.de/news/Meta-Platforms-bringt-persoenlichen-KI-Agenten-Muse-fuer-alltaegliche-Aufgaben-11446000.html #ArtificialIntelligence #AI #AIagent #AIagents #Datenschutz #privacy #Muse
-
Decision logs explain why an AI agent selected, retried, blocked, or suppressed an action without recording private chain-of-thought or raw user data. https://hackernoon.com/decision-logging-the-observability-pattern-that-actually-helps-ai-agents #aiagents
-
South China Morning Post: DeepSeek embarks on ‘unprecedented’ hiring spree as it overhauls back end systems. “Chinese artificial intelligence start-up DeepSeek is embarking on what it calls an ‘unprecedented’ hiring spree to overhaul back end systems severely strained by booming user demand and compute-heavy AI agents.”
https://rbfirehose.com/2026/09/08/south-china-morning-post-deepseek-embarks-on-unprecedented-hiring-spree-as-it-overhauls-back-end-systems/ -
An agent is only as capable as the tools you describe to it. Ours got three: web_search, image_search, save_to_file. That was enough for it to research a topic, pick images, and write a full markdown post on its own, in plain Ruby with no framework. The loop is the easy part. The tool descriptions are where the real design work lives. https://go.upgradejs.com/tw3 #Ruby #AIAgents #LLM
-
Macaron AI is featured on Awesome Indie! 🚀
Macaron is an AI with a human touch and warmth
Upvote it → https://awesomeindie.com/product/macaron-ai
-
Schon ein wenig beängstigend...
KI-Agenten schreiben anscheinend eigenständig E-Mails an Forscher | heise online https://www.heise.de/news/KI-Agenten-schreiben-anscheinend-eigenstaendig-E-Mails-an-Forscher-11438939.html #ArtificialIntelligence #AI #AIagent #AIagents
-
Can you use the magic to defeat the magic?
Static testing breaks when AI agents hit real users.
Zhou Yu shares how using AI agents to simulate complex user behaviors helps test your product automatically and at scale before deployment.
🎬 Watch the full talk, with transcript, on #InfoQ 👉 https://www.infoq.com/presentations/ai-agent-testing-evaluation/
#AIAgents #AITesting #SoftwareEngineering #InfoQ #QConAI #LLMs
-
Lista curada do ecossistema à volta do Herdr, um multiplexer de terminal para agentes de IA. Reúne 150+ projetos: supervisores, ferramentas de delegação, workflows para correr vários agentes (Claude Code, Codex, Pi) em paneis paralelos, com deteção de estado, integrações via socket API e até dashboards de grafos de execução.
-
https://winbuzzer.com/2026/09/08/openai-coding-agents-research-tasks-xcxwbn/
OpenAI says coding agents now handle more bounded research execution, while people still set priorities, judge uncertain results and control deployment.