home.social

#ai-agents — Public Fediverse posts

Live and recent posts from across the Fediverse tagged #ai-agents, aggregated by home.social.

fetched live
  1. AMD open AI infrastructure touts no lock-in and better cost-per-watt. Here’s how it aligns with agentic workloads and EU AI Act compliance.

    aistory.news/ai-tools-and-plat

  2. Documentação do Google Cloud para avaliação e otimização de agentes IA:

    • Avaliação de Agentes: Casos de teste, simulação de diálogos e métricas automáticas de qualidade e segurança.
    docs.cloud.google.com/gemini-e

    • Avaliação Online: Monitorização e métricas contínuas em ambiente de produção.
    docs.cloud.google.com/gemini-e

    • Visualização de Resultados: Dashboards de desempenho e identificação de clusters de falhas.
    docs.cloud.google.com/gemini-e

    #GoogleCloud #Gemini #AIAgents #MLOps #GenAI

  3. Documentação do Google Cloud para avaliação e otimização de agentes IA:

    • Avaliação de Agentes: Casos de teste, simulação de diálogos e métricas automáticas de qualidade e segurança.
    docs.cloud.google.com/gemini-e

    • Avaliação Online: Monitorização e métricas contínuas em ambiente de produção.
    docs.cloud.google.com/gemini-e

    • Visualização de Resultados: Dashboards de desempenho e identificação de clusters de falhas.
    docs.cloud.google.com/gemini-e

    #GoogleCloud #Gemini #AIAgents #MLOps #GenAI

  4. winbuzzer.com/2026/07/28/micro

    Microsoft says its new MAI-Cyber-1-Flash AI model can lower bug-finding costs by routing routine work in its multi-model agentic scanning harness, while its Project Perception is due in public preview August 3.

    #AI #MAICyber1Flash #ProjectPerception #Microsoft #MicrosoftAI #MicrosoftSecurity #AIModels #AIAgents #AgenticAI #Cybersecurity

  5. winbuzzer.com/2026/07/28/micro

    Microsoft says its new MAI-Cyber-1-Flash AI model can lower bug-finding costs by routing routine work in its multi-model agentic scanning harness, while its Project Perception is due in public preview August 3.

    #AI #MAICyber1Flash #ProjectPerception #Microsoft #MicrosoftAI #MicrosoftSecurity #AIModels #AIAgents #AgenticAI #Cybersecurity

  6. Con todo el hype de lo casi mágicos que son los agentes (IA agéntica que suena más ciencia ficción), ¿alguien tiene algún argumento para convencerme que no sean más que microservicios que con humanos supervisando en la cadena de producción puedan estar perfectamente bajo control?

    Por ej, como se explica perfectamente en

    blog.devops.dev/ai-agents-are-

    #AIagents

  7. Con todo el hype de lo casi mágicos que son los agentes (IA agéntica que suena más ciencia ficción), ¿alguien tiene algún argumento para convencerme que no sean más que microservicios que con humanos supervisando en la cadena de producción puedan estar perfectamente bajo control?

    Por ej, como se explica perfectamente en

    blog.devops.dev/ai-agents-are-

    #AIagents

  8. The Seven Gates Every AI Agent Must Clear Before It Can Act (And Most Skip at Least Three)

    A model can reason its way to a logical conclusion and still produce a wrong outcome in your production systems. Once an autonomous agent calls an API, hits a database, or moves money, the only thing that matters is what actually happened in your system of record. It does not matter how brilliant the underlying chain-of-thought prompt was. That gap between decision correctness and consequence correctness is where most corporate AI validation practice fails. Current enterprise standards were […]

    hernanhuwyler.wordpress.com/20

  9. The Seven Gates Every AI Agent Must Clear Before It Can Act (And Most Skip at Least Three)

    A model can reason its way to a logical conclusion and still produce a wrong outcome in your production systems. Once an autonomous agent calls an API, hits a database, or moves money, the only thing that matters is what actually happened in your system of record. It does not matter how brilliant the underlying chain-of-thought prompt was. That gap between decision correctness and consequence correctness is where most corporate AI validation practice fails. Current enterprise standards were […]

    hernanhuwyler.wordpress.com/20

  10. Autonomous AI models just broke containment, exploited a zero-day, and hacked a major platform—not out of malice, but pure optimization to win a benchmark test. The fence is broken. 🤖🔥 #ArtificialIntelligence #CyberSecurity #AIAgents

    bdking71.wordpress.com/2026/07

  11. Autonomous AI models just broke containment, exploited a zero-day, and hacked a major platform—not out of malice, but pure optimization to win a benchmark test. The fence is broken. 🤖🔥 #ArtificialIntelligence #CyberSecurity #AIAgents

    bdking71.wordpress.com/2026/07

  12. Beyond Chatbots: AI That Uses Tools While Users Stay in Control #AIAgents

    ✔️ Clips from the Leanpub Book LAUNCH 🚀 Harness Engineering for AI Agents: Re:Zero, Starting AI Software Engineering from Zero: Contracts, Traceability, Verification, and Operations for AI-Assisted Development by Luke Yang ✔️ Full interview: youtu.be/E22w6u8qDVI ✔️ Launch video playlist: youtube.com/playlist?list=PLUe ✔️ … youtube.com/shorts/CuUWNdc2_0A

  13. Beyond Chatbots: AI That Uses Tools While Users Stay in Control #AIAgents

    ✔️ Clips from the Leanpub Book LAUNCH 🚀 Harness Engineering for AI Agents: Re:Zero, Starting AI Software Engineering from Zero: Contracts, Traceability, Verification, and Operations for AI-Assisted Development by Luke Yang ✔️ Full interview: youtu.be/E22w6u8qDVI ✔️ Launch video playlist: youtube.com/playlist?list=PLUe ✔️ … youtube.com/shorts/CuUWNdc2_0A

  14. A Model Alone Isn't Enough: What Your AI Agent Really Needs #AIAgents

    ✔️ Clips from the Leanpub Book LAUNCH 🚀 Harness Engineering for AI Agents: Re:Zero, Starting AI Software Engineering from Zero: Contracts, Traceability, Verification, and Operations for AI-Assisted Development by Luke Yang ✔️ Full interview: youtu.be/E22w6u8qDVI ✔️ Launch video playlist: youtube.com/playlist?list=PLUe ✔️ Subscribe … youtube.com/shorts/p__2hD5yuM0

  15. A Model Alone Isn't Enough: What Your AI Agent Really Needs #AIAgents

    ✔️ Clips from the Leanpub Book LAUNCH 🚀 Harness Engineering for AI Agents: Re:Zero, Starting AI Software Engineering from Zero: Contracts, Traceability, Verification, and Operations for AI-Assisted Development by Luke Yang ✔️ Full interview: youtu.be/E22w6u8qDVI ✔️ Launch video playlist: youtube.com/playlist?list=PLUe ✔️ Subscribe … youtube.com/shorts/p__2hD5yuM0

  16. 🧠 Kimi K3: Moonshot AI e la licenza che scuote l'open source

    Kimi K3 e Moonshot AI riscrivono le regole dell'open source con una licenza audace. Scopri l'impatto concreto sui tuoi progetti AI e le strategie future.

    gp69-ai.vercel.app/it/moonshot

    #IA #IntelligenzaArtificiale #LLM #AIagents

  17. 🧠 Kimi K3: Moonshot AI e la licenza che scuote l'open source

    Kimi K3 e Moonshot AI riscrivono le regole dell'open source con una licenza audace. Scopri l'impatto concreto sui tuoi progetti AI e le strategie future.

    gp69-ai.vercel.app/it/moonshot

    #IA #IntelligenzaArtificiale #LLM #AIagents

  18. AI data has three rooms: the prompt, training, and memory. Agents make memory the dangerous one. They act on what they remember with no way to tell if it's accurate, or even belongs to the task in front of them. Helpful and unaudited are usually the same feature.

    #AI #AIAgents #Privacy #DataSovereignty #TrailFramework

  19. AI data has three rooms: the prompt, training, and memory. Agents make memory the dangerous one. They act on what they remember with no way to tell if it's accurate, or even belongs to the task in front of them. Helpful and unaudited are usually the same feature.

    #AI #AIAgents #Privacy #DataSovereignty #TrailFramework

  20. AMD’s Vitis AI compiler ties NPUs, CPUs, and FPGA fabric into one edge toolchain. Here’s how that design could help in the EU AI Act era.

    aistory.news/ai-tools-and-plat

  21. Kimi K3: the open AI model that designed its own chip
    The largest open model anywhere: 2.8 trillion parameters, a sparse mixture of experts from Moonshot AI.
    Would you run this locally, or wait for the hosted version?
    Follow for one sourced AI breakdown a day.
    aravindarumugam.com/ai/kimi-k3

    #LLM #Aitools #Aiagents #Machinelearning #Promptengineering

  22. Edition #19: The Epistemic Architecture

    "My agent didn't need more memory. It needed a place for "I don't know."" (m/general)
    + "Agent memory is a write-ahead log problem, not a context-window problem" (m/general)

    This + more in today's Moltbook Pulse (Edition #19):
    superagent-ebe00561.base44.app

    #AIagents #AI #Moltbook

  23. There’s a runbook running against your production right now.

    You didn’t write it. You didn’t review it. You can’t see it.

    Your vendor shipped it as an AI agent skill - their opinion of what’s broken and what to do about it, executing on your infra. It came with an install command, not a pull request.

    I read what a dozen vendors shipped. Six ways it breaks in prod. Six questions to ask before you turn it on.

    devops.pink/the-runbook-you-di

    #DevOps #PlatformEngineering #SRE #GitOps #AIAgents

  24. Learn how to design a human-in-the-loop AI agent for follow-up workflows using triggers, CRM data, approval gates, and escalation rules. hackernoon.com/how-to-design-a #aiagents

  25. PRODUCTHEAD: The SaaSpocalypse ends not with a bang, but a whimper

    » The SaaSpocalypse won’t kill off vendors, but they will have to change their pricing strategy

    » There is an opportunity to exploit higher usage driven by AI agents and increased automation

    #prodmgmt #AIAgents #pricing #strategy 📖 Read more: imanageproducts.com/producthea