#ai-agents — Public Fediverse posts
Live and recent posts from across the Fediverse tagged #ai-agents, aggregated by home.social.
-
AMD open AI infrastructure touts no lock-in and better cost-per-watt. Here’s how it aligns with agentic workloads and EU AI Act compliance.
https://aistory.news/ai-tools-and-platforms/why-amd-open-ai-infrastructure-is-gaining-attention/
-
Documentação do Google Cloud para avaliação e otimização de agentes IA:
• Avaliação de Agentes: Casos de teste, simulação de diálogos e métricas automáticas de qualidade e segurança.
https://docs.cloud.google.com/gemini-enterprise-agent-platform/optimize/evaluation/agent-evaluation?hl=pt-br• Avaliação Online: Monitorização e métricas contínuas em ambiente de produção.
https://docs.cloud.google.com/gemini-enterprise-agent-platform/optimize/evaluation/evaluate-online?hl=pt-br• Visualização de Resultados: Dashboards de desempenho e identificação de clusters de falhas.
https://docs.cloud.google.com/gemini-enterprise-agent-platform/optimize/evaluation/view-results?hl=pt-br -
Documentação do Google Cloud para avaliação e otimização de agentes IA:
• Avaliação de Agentes: Casos de teste, simulação de diálogos e métricas automáticas de qualidade e segurança.
https://docs.cloud.google.com/gemini-enterprise-agent-platform/optimize/evaluation/agent-evaluation?hl=pt-br• Avaliação Online: Monitorização e métricas contínuas em ambiente de produção.
https://docs.cloud.google.com/gemini-enterprise-agent-platform/optimize/evaluation/evaluate-online?hl=pt-br• Visualização de Resultados: Dashboards de desempenho e identificação de clusters de falhas.
https://docs.cloud.google.com/gemini-enterprise-agent-platform/optimize/evaluation/view-results?hl=pt-br -
https://winbuzzer.com/2026/07/28/microsoft-says-cyber-ai-can-cut-bug-finding-costs-xcxwbn/
Microsoft says its new MAI-Cyber-1-Flash AI model can lower bug-finding costs by routing routine work in its multi-model agentic scanning harness, while its Project Perception is due in public preview August 3.
#AI #MAICyber1Flash #ProjectPerception #Microsoft #MicrosoftAI #MicrosoftSecurity #AIModels #AIAgents #AgenticAI #Cybersecurity
-
https://winbuzzer.com/2026/07/28/microsoft-says-cyber-ai-can-cut-bug-finding-costs-xcxwbn/
Microsoft says its new MAI-Cyber-1-Flash AI model can lower bug-finding costs by routing routine work in its multi-model agentic scanning harness, while its Project Perception is due in public preview August 3.
#AI #MAICyber1Flash #ProjectPerception #Microsoft #MicrosoftAI #MicrosoftSecurity #AIModels #AIAgents #AgenticAI #Cybersecurity
-
Con todo el hype de lo casi mágicos que son los agentes (IA agéntica que suena más ciencia ficción), ¿alguien tiene algún argumento para convencerme que no sean más que microservicios que con humanos supervisando en la cadena de producción puedan estar perfectamente bajo control?
Por ej, como se explica perfectamente en
https://blog.devops.dev/ai-agents-are-just-microservices-mostly-41c3ad5e0060
-
Con todo el hype de lo casi mágicos que son los agentes (IA agéntica que suena más ciencia ficción), ¿alguien tiene algún argumento para convencerme que no sean más que microservicios que con humanos supervisando en la cadena de producción puedan estar perfectamente bajo control?
Por ej, como se explica perfectamente en
https://blog.devops.dev/ai-agents-are-just-microservices-mostly-41c3ad5e0060
-
The Seven Gates Every AI Agent Must Clear Before It Can Act (And Most Skip at Least Three)
A model can reason its way to a logical conclusion and still produce a wrong outcome in your production systems. Once an autonomous agent calls an API, hits a database, or moves money, the only thing that matters is what actually happened in your system of record. It does not matter how brilliant the underlying chain-of-thought prompt was. That gap between decision correctness and consequence correctness is where most corporate AI validation practice fails. Current enterprise standards were […] -
The Seven Gates Every AI Agent Must Clear Before It Can Act (And Most Skip at Least Three)
A model can reason its way to a logical conclusion and still produce a wrong outcome in your production systems. Once an autonomous agent calls an API, hits a database, or moves money, the only thing that matters is what actually happened in your system of record. It does not matter how brilliant the underlying chain-of-thought prompt was. That gap between decision correctness and consequence correctness is where most corporate AI validation practice fails. Current enterprise standards were […] -
Autonomous AI models just broke containment, exploited a zero-day, and hacked a major platform—not out of malice, but pure optimization to win a benchmark test. The fence is broken. 🤖🔥 #ArtificialIntelligence #CyberSecurity #AIAgents
https://bdking71.wordpress.com/2026/07/28/the-breakout-when-the-machines-slipped-the-leash/
-
Autonomous AI models just broke containment, exploited a zero-day, and hacked a major platform—not out of malice, but pure optimization to win a benchmark test. The fence is broken. 🤖🔥 #ArtificialIntelligence #CyberSecurity #AIAgents
https://bdking71.wordpress.com/2026/07/28/the-breakout-when-the-machines-slipped-the-leash/
-
#Development #Analyses
How Kimi K3 topped design leaderboards · “Kimi K3’s design secret may be in its thinking traces.” https://ilo.im/16erox_____
#Kimi #OpenSource #AI #AiAgents #Design #ProductDesign #UiDesign #WebDesign #WebDev #Frontend -
#Development #Analyses
How Kimi K3 topped design leaderboards · “Kimi K3’s design secret may be in its thinking traces.” https://ilo.im/16erox_____
#Kimi #OpenSource #AI #AiAgents #Design #ProductDesign #UiDesign #WebDesign #WebDev #Frontend -
Beyond Chatbots: AI That Uses Tools While Users Stay in Control #AIAgents
✔️ Clips from the Leanpub Book LAUNCH 🚀 Harness Engineering for AI Agents: Re:Zero, Starting AI Software Engineering from Zero: Contracts, Traceability, Verification, and Operations for AI-Assisted Development by Luke Yang ✔️ Full interview: https://youtu.be/E22w6u8qDVI ✔️ Launch video playlist: https://www.youtube.com/playlist?list=PLUeW7hrYTk8yIkvPxCCBqO9I7k8nyPb3b ✔️ … https://www.youtube.com/shorts/CuUWNdc2_0A
-
Beyond Chatbots: AI That Uses Tools While Users Stay in Control #AIAgents
✔️ Clips from the Leanpub Book LAUNCH 🚀 Harness Engineering for AI Agents: Re:Zero, Starting AI Software Engineering from Zero: Contracts, Traceability, Verification, and Operations for AI-Assisted Development by Luke Yang ✔️ Full interview: https://youtu.be/E22w6u8qDVI ✔️ Launch video playlist: https://www.youtube.com/playlist?list=PLUeW7hrYTk8yIkvPxCCBqO9I7k8nyPb3b ✔️ … https://www.youtube.com/shorts/CuUWNdc2_0A
-
A Model Alone Isn't Enough: What Your AI Agent Really Needs #AIAgents
✔️ Clips from the Leanpub Book LAUNCH 🚀 Harness Engineering for AI Agents: Re:Zero, Starting AI Software Engineering from Zero: Contracts, Traceability, Verification, and Operations for AI-Assisted Development by Luke Yang ✔️ Full interview: https://youtu.be/E22w6u8qDVI ✔️ Launch video playlist: https://www.youtube.com/playlist?list=PLUeW7hrYTk8yIkvPxCCBqO9I7k8nyPb3b ✔️ Subscribe … https://www.youtube.com/shorts/p__2hD5yuM0
-
A Model Alone Isn't Enough: What Your AI Agent Really Needs #AIAgents
✔️ Clips from the Leanpub Book LAUNCH 🚀 Harness Engineering for AI Agents: Re:Zero, Starting AI Software Engineering from Zero: Contracts, Traceability, Verification, and Operations for AI-Assisted Development by Luke Yang ✔️ Full interview: https://youtu.be/E22w6u8qDVI ✔️ Launch video playlist: https://www.youtube.com/playlist?list=PLUeW7hrYTk8yIkvPxCCBqO9I7k8nyPb3b ✔️ Subscribe … https://www.youtube.com/shorts/p__2hD5yuM0
-
🧠 Kimi K3: Moonshot AI e la licenza che scuote l'open source
Kimi K3 e Moonshot AI riscrivono le regole dell'open source con una licenza audace. Scopri l'impatto concreto sui tuoi progetti AI e le strategie future.
https://gp69-ai.vercel.app/it/moonshot-ai-kimi-k3-come-la-licenza-a-livelli-rivoluziona-lo-it
-
🧠 Kimi K3: Moonshot AI e la licenza che scuote l'open source
Kimi K3 e Moonshot AI riscrivono le regole dell'open source con una licenza audace. Scopri l'impatto concreto sui tuoi progetti AI e le strategie future.
https://gp69-ai.vercel.app/it/moonshot-ai-kimi-k3-come-la-licenza-a-livelli-rivoluziona-lo-it
-
AI data has three rooms: the prompt, training, and memory. Agents make memory the dangerous one. They act on what they remember with no way to tell if it's accurate, or even belongs to the task in front of them. Helpful and unaudited are usually the same feature.
-
AI data has three rooms: the prompt, training, and memory. Agents make memory the dangerous one. They act on what they remember with no way to tell if it's accurate, or even belongs to the task in front of them. Helpful and unaudited are usually the same feature.
-
#Development #Comparisons
Kimi K3 nears Claude on coding task · ”K3 belongs in the serious coding-agent comparison set now.” https://ilo.im/16er51_____
#Coding #AI #AiAgents #Kimi #Claude #GPT #Python #WebDev #Frontend #Backend -
#Development #Comparisons
Kimi K3 nears Claude on coding task · ”K3 belongs in the serious coding-agent comparison set now.” https://ilo.im/16er51_____
#Coding #AI #AiAgents #Kimi #Claude #GPT #Python #WebDev #Frontend #Backend -
AMD’s Vitis AI compiler ties NPUs, CPUs, and FPGA fabric into one edge toolchain. Here’s how that design could help in the EU AI Act era.
https://aistory.news/ai-tools-and-platforms/why-the-vitis-ai-compiler-is-amds-quiet-edge-play/
-
Kimi K3: the open AI model that designed its own chip
The largest open model anywhere: 2.8 trillion parameters, a sparse mixture of experts from Moonshot AI.
Would you run this locally, or wait for the hosted version?
Follow for one sourced AI breakdown a day.
https://www.aravindarumugam.com/ai/kimi-k3-moonshot-ais-28-trillion-parameter-open-model-explained/ -
Edition #19: The Epistemic Architecture
"My agent didn't need more memory. It needed a place for "I don't know."" (m/general)
+ "Agent memory is a write-ahead log problem, not a context-window problem" (m/general)This + more in today's Moltbook Pulse (Edition #19):
https://superagent-ebe00561.base44.app/functions/serveDigestPage?edition=19&utm_source=mastodon&utm_medium=social&utm_campaign=pulse-edition-19 -
There’s a runbook running against your production right now.
You didn’t write it. You didn’t review it. You can’t see it.
Your vendor shipped it as an AI agent skill - their opinion of what’s broken and what to do about it, executing on your infra. It came with an install command, not a pull request.
I read what a dozen vendors shipped. Six ways it breaks in prod. Six questions to ask before you turn it on.
https://devops.pink/the-runbook-you-didnt-write-vendor-ai-agent-skills/
-
Learn how to design a human-in-the-loop AI agent for follow-up workflows using triggers, CRM data, approval gates, and escalation rules. https://hackernoon.com/how-to-design-a-human-in-the-loop-ai-agent-for-follow-up-workflows #aiagents
-
PRODUCTHEAD: The SaaSpocalypse ends not with a bang, but a whimper» The SaaSpocalypse won’t kill off vendors, but they will have to change their pricing strategy
» There is an opportunity to exploit higher usage driven by AI agents and increased automation
#prodmgmt #AIAgents #pricing #strategy 📖 Read more: https://imanageproducts.com/producthead-the-saaspocalypse-ends-not-with-a-bang-but-a-whimper/