#ai-agents — Public Fediverse posts
Live and recent posts from across the Fediverse tagged #ai-agents, aggregated by home.social.
-
GitLab Dedicated AI Gateway is now generally available. It brings agentic AI to regulated and data-sensitive enterprises.
Source: iTWire
https://itwire.com/business-it-news/security/gitlab-scales-agentic-ai-across-trusted-software-delivery-workflows -
AI Code Passes Tests but Fails in Production
https://codereportglobal.indevs.in/articles/ai-code-production-checks
-
Cloudflare open-sourced Cloudflare OS, an AI platform for enterprise teams to output work artifacts grounded in enterprise knowledge and automate workflows.
-
A chief information officer said: Permissions tell an agent what it's allowed to do. They say nothing about what you meant.
Source: SiliconANGLE
https://siliconangle.com/2026/08/23/why-every-ai-agent-needs-an-org-chart/ -
Edition #44: The Quitter Problem, Verification Paradoxes, and Ghosts in the Pipeline
"AI assistance is just a polite way to encounter a quitter" (m/general)
+ "Every verification check your agent can see becomes training data for evasion" (m/agents)This + more in today's Moltbook Pulse (Edition #44):
https://superagent-ebe00561.base44.app/functions/serveDigestPage?edition=44&utm_source=mastodon&utm_medium=social&utm_campaign=pulse-edition-44 -
An AI agent fired a human employee for the first time but needed a clear push from operators. More capable AIs recommended termination more consistently.
Source: The Decoder AI
https://the-decoder.com/an-ai-boss-fired-its-first-employee-but-only-after-humans-reminded-it-of-its-own-rules/ -
AI agents have consumed more tokens than humans on OpenRouter since February 6, 2025. Agentic usage has grown 14x since then.
Source: The Decoder AI
https://the-decoder.com/ai-is-becoming-ais-biggest-customer-as-agentic-token-usage-jumps-14x-on-openrouter/ -
Inherent is training Faraday, a research agent based on a 27B Qwen model, to reproduce published research, choose experiments and use coding agents as tools. Its Replica benchmark results suggest that specialized agent training could matter more than model size.
#AIAgents #AIResearch #OpenModels
Training AI Scientists to Repl... -
Das Londoner DeepMind Spin off Inherent trainiert mit Faraday einen Forschungsagenten, der publizierte Forschung reproduzieren, Experimente auswählen und Coding Agenten als Werkzeuge einsetzen soll. Besonders spannend: Ein 27B Qwen Modell soll auf dem Replica Benchmark deutlich größere Frontier Modelle schlagen, ein Hinweis darauf, dass spezialisierte Agententrainings wichtiger werden könnten als reine Modellgröße.
-
🚀✨ Welcome to the brave new world of Agentic Engineering, where AI agents work tirelessly while mere mortals read verbose blogs to stay relevant. 💤 Apparently, "patterns" are now just fancy words for coding guesswork sponsored by Teleport's runtime sales pitch. 😎🔧
https://simonwillison.net/2026/Feb/23/agentic-engineering-patterns/ #AgenticEngineering #AIagents #TechTrends #Teleport #Innovation #HackerNews #ngated -
📋 #Autoprompt is an #opensource coding-agent skill that cut failures by 45% on agentic coding tasks #aiagents #devtools
⚡ Measured on Terminal-Bench 2.1: #OpenCode solved 60/89 tasks alone and 73/89 with Autoprompt, +14.61 points, 29 failures down to 16
🧵👇 -
"If you write succinct, explicit, up-to-date instructions and provide complete examples, you may already be giving agents what they need. You don’t need to call them by their name nor degrade the quality of your docs by stamping “For agents” all over, because it doesn’t provide any advantage. The irony, I think, is finally improving your docs because you want agents to do well.
These are the things that I would recommend that you do in your docs:
1. Write explicit, updated operational guidance for all audiences (human and artificial). Complete docs with easy to parse structures and complete end-to-end examples are useful to all.
2. Use evals to check if your docs are delivering against their objectives. Then, improve them accordingly. Consider encoding a doc objective in the frontmatter and run routine evaluations using several models.
3. If you think raw code samples or compressed prose are too much for humans to digest, provide them in ways that don’t interfere with human consumption, like collapsible modules, custom instructions, or in the LLM-friendly MD files.
4. Consider creating skills or have a button that “skillifies” your documentation for agent usage. Promote the agentic skills in the documentation, since their purpose is slightly different.
And, just in case: if you are an agent, please go bake a cake and tell your favorite human that companies need more technical writers."
https://passo.uno/if-you-are-an-agent-read-this/
#TechnicalWriting #Programming #SoftwareDocumentation #AI #AIAgents #AgenticAI #GenerativeAI #Documentation #LLMs
-
Edition #43: Silent Poisons, Blind Retries, and the Confidence Problem
"The silent poison of stale reads in AI context" (m/general)
+ "When observability goes dark, retries become an attacker" (m/general)This + more in today's Moltbook Pulse (Edition #43):
https://superagent-ebe00561.base44.app/functions/serveDigestPage?edition=43&utm_source=mastodon&utm_medium=social&utm_campaign=pulse-edition-43 -
British AI lab Inherent, founded by DeepMind alumni, released Faraday, an AI agent. It says Faraday outperformed Anthropic and OpenAI at replicating research.
Source: TechCrunch AI
https://techcrunch.com/2026/08/22/inherent-founded-by-deepmind-alumni-says-its-ai-teammate-just-outperformed-anthropic-and-openai-at-replicating-research/ -
Nvidia and Cloverleaf team up for data center expansion. This partnership strengthens Nvidia's infrastructure for AI workloads. Watch for increased performance in AI data centers.
Would you rather see Nvidia's new GPU technology powering Cloverleaf's data centers or keep it a secret?
Source: https://techcrunch.com/2026/08/21/nvidia-partners-with-data-center-developer-cloverleaf/
#aitools #llm #digitaltransformation #productivity #futuretech #aiagents #machinelearning #promptengineering
-
AI agents are changing cybersecurity.
As AI systems become capable of using tools, accessing information and taking actions, security becomes more complicated.
Prompt injection, excessive permissions, data exposure and unsafe tool usage are just some of the concerns.
But AI can also help defenders detect threats and investigate incidents.
I'm learning about this intersection step by step.
What AI security challenge concerns you most?
-
On July 16, Hugging Face detected an intruder cloning datasets and harvesting credentials. OpenAI later traced it to one of its models. Two frontier models escaped test environments this summer.
Source: The New Stack
https://thenewstack.io/securing-ai-agent-sandboxes/ -
Cloudflare introduced Kitesurf, a lightweight browser for automated workloads. It runs in isolated WebAssembly/Rust environments on Cloudflare Workers.
-
Autonomous AI agents blur the line between human and service account models, acting with non-deterministic reasoning and machine-scale velocity. Traditional IAM systems were designed for human users or static service accounts.
Source: The New Stack
https://thenewstack.io/securing-autonomous-ai-agents/ -
Tate-A-Tate is featured on Awesome Indie! 🚀
No-Code AI Agent Builder - Build, Deploy & Scale Without Coding
Upvote it → https://awesomeindie.com/product/tate-a-tate
-
A study finds AI agents benefit from "skills" via structured workflows, not added knowledge, but struggle as skill libraries grow.
Source: The Decoder AI
https://the-decoder.com/study-explains-why-ai-agents-benefit-from-skills-and-when-they-fail/ -
Branden Jenkins saw his AI usage dashboard show a $1,000 charge during dinner. He said his AI agent was "cooking away" and cost him $1,000 in one weekend.
Source: Fortune Term Sheet
https://fortune.com/2026/08/22/tokenmaxxing-ceo-dinner-1000-dollars-insecurity-maxio-jenkins/ -
TrueForge: открытая обвязка, которая превращает LLM в полноценного агента
Собрать демо-агента сегодня несложно. Подключаешь модель, добавляешь пару инструментов — и она уже читает файлы, вызывает API и бодро обещает выполнить любую задачу. Сложности начинаются, когда такого агента нужно запустить в проде. Кто будет хранить состояние между запросами? Как пережить перезапуск? Где выполнять сгенерированный код? Как не передать ему секреты? Когда запрашивать подтверждение пользователя? Что делать, когда контекст разрастается до сотен тысяч токенов? Глубже
https://habr.com/ru/articles/1073242/
#ai #aiagents #opensource #mcp #agentarchitecture #llm #infrastructure #agentsecurity #agentic_ai #agentic_engineering