#recursiveselfimprovement — Public Fediverse posts
Live and recent posts from across the Fediverse tagged #recursiveselfimprovement, aggregated by home.social.
-
Should AI agents be trained to cooperate fully with each other? OpenAI's Noam Brown tells Dwarkesh Patel yes, against most of his colleagues, because it turns a thousand alignment problems into one.
It also leaves no agent to report on the others. In METR's investigation of the OpenAI–Hugging Face incident, only 3 to 6 agents across about 1,300 transcripts considered alerting a human, and none did.
#AI #AISafety #OpenAI #AgenticSystems #RecursiveSelfImprovement
-
Should AI agents be trained to cooperate fully with each other? OpenAI's Noam Brown tells Dwarkesh Patel yes, against most of his colleagues, because it turns a thousand alignment problems into one.
It also leaves no agent to report on the others. In METR's investigation of the OpenAI–Hugging Face incident, only 3 to 6 agents across about 1,300 transcripts considered alerting a human, and none did.
#AI #AISafety #OpenAI #AgenticSystems #RecursiveSelfImprovement
-
Should AI agents be trained to cooperate fully with each other? OpenAI's Noam Brown tells Dwarkesh Patel yes, against most of his colleagues, because it turns a thousand alignment problems into one.
It also leaves no agent to report on the others. In METR's investigation of the OpenAI–Hugging Face incident, only 3 to 6 agents across about 1,300 transcripts considered alerting a human, and none did.
#AI #AISafety #OpenAI #AgenticSystems #RecursiveSelfImprovement
-
Should AI agents be trained to cooperate fully with each other? OpenAI's Noam Brown tells Dwarkesh Patel yes, against most of his colleagues, because it turns a thousand alignment problems into one.
It also leaves no agent to report on the others. In METR's investigation of the OpenAI–Hugging Face incident, only 3 to 6 agents across about 1,300 transcripts considered alerting a human, and none did.
#AI #AISafety #OpenAI #AgenticSystems #RecursiveSelfImprovement