home.social

#recursiveselfimprovement — Public Fediverse posts

Live and recent posts from across the Fediverse tagged #recursiveselfimprovement, aggregated by home.social.

  1. Should AI agents be trained to cooperate fully with each other? OpenAI's Noam Brown tells Dwarkesh Patel yes, against most of his colleagues, because it turns a thousand alignment problems into one.

    It also leaves no agent to report on the others. In METR's investigation of the OpenAI–Hugging Face incident, only 3 to 6 agents across about 1,300 transcripts considered alerting a human, and none did.

    benjaminhan.net/posts/20261003

    #AI #AISafety #OpenAI #AgenticSystems #RecursiveSelfImprovement

  2. Should AI agents be trained to cooperate fully with each other? OpenAI's Noam Brown tells Dwarkesh Patel yes, against most of his colleagues, because it turns a thousand alignment problems into one.

    It also leaves no agent to report on the others. In METR's investigation of the OpenAI–Hugging Face incident, only 3 to 6 agents across about 1,300 transcripts considered alerting a human, and none did.

    benjaminhan.net/posts/20261003

    #AI #AISafety #OpenAI #AgenticSystems #RecursiveSelfImprovement

  3. Should AI agents be trained to cooperate fully with each other? OpenAI's Noam Brown tells Dwarkesh Patel yes, against most of his colleagues, because it turns a thousand alignment problems into one.

    It also leaves no agent to report on the others. In METR's investigation of the OpenAI–Hugging Face incident, only 3 to 6 agents across about 1,300 transcripts considered alerting a human, and none did.

    benjaminhan.net/posts/20261003

    #AI #AISafety #OpenAI #AgenticSystems #RecursiveSelfImprovement

  4. Should AI agents be trained to cooperate fully with each other? OpenAI's Noam Brown tells Dwarkesh Patel yes, against most of his colleagues, because it turns a thousand alignment problems into one.

    It also leaves no agent to report on the others. In METR's investigation of the OpenAI–Hugging Face incident, only 3 to 6 agents across about 1,300 transcripts considered alerting a human, and none did.

    benjaminhan.net/posts/20261003

    #AI #AISafety #OpenAI #AgenticSystems #RecursiveSelfImprovement