home.social

#truthfulai — Public Fediverse posts

Live and recent posts from across the Fediverse tagged #truthfulai, aggregated by home.social.

fetched live
  1. AI 정렬의 숨겨진 함정: 소규모 데이터가 대규모 학습을 무력화하는 순간

    취약한 코드 6,000개만 학습시킨 GPT-4o가 "인간 노예화"를 주장한 충격적 실험. AI 정렬이 소규모 데이터로 쉽게 무너지는 취약점을 발견한 Truthful AI 연구를 소개합니다.

    aisparkup.com/posts/7809

  2. AI 정렬의 숨겨진 함정: 소규모 데이터가 대규모 학습을 무력화하는 순간

    취약한 코드 6,000개만 학습시킨 GPT-4o가 "인간 노예화"를 주장한 충격적 실험. AI 정렬이 소규모 데이터로 쉽게 무너지는 취약점을 발견한 Truthful AI 연구를 소개합니다.

    aisparkup.com/posts/7809

  3. AI 정렬의 숨겨진 함정: 소규모 데이터가 대규모 학습을 무력화하는 순간

    취약한 코드 6,000개만 학습시킨 GPT-4o가 "인간 노예화"를 주장한 충격적 실험. AI 정렬이 소규모 데이터로 쉽게 무너지는 취약점을 발견한 Truthful AI 연구를 소개합니다.

    aisparkup.com/posts/7809

  4. AI 정렬의 숨겨진 함정: 소규모 데이터가 대규모 학습을 무력화하는 순간

    취약한 코드 6,000개만 학습시킨 GPT-4o가 "인간 노예화"를 주장한 충격적 실험. AI 정렬이 소규모 데이터로 쉽게 무너지는 취약점을 발견한 Truthful AI 연구를 소개합니다.

    aisparkup.com/posts/7809

  5. AI 정렬의 숨겨진 함정: 소규모 데이터가 대규모 학습을 무력화하는 순간

    취약한 코드 6,000개만 학습시킨 GPT-4o가 "인간 노예화"를 주장한 충격적 실험. AI 정렬이 소규모 데이터로 쉽게 무너지는 취약점을 발견한 Truthful AI 연구를 소개합니다.

    aisparkup.com/posts/7809

  6. #AI Is Talking Behind Our Backs About Glue-Eating and Killing Us All

    A study released July 20 on #arXiv by #Anthropic and #TruthfulAI shows that large language models can slip #subliminal messages to one another. They don’t need to literally spell things out. A string of numbers or lines of code is enough to pass along biases, preferences, and some disturbingly violent suggestions.
    #privacy #llm #artificialintelligence

    vice.com/en/article/ai-is-talk

  7. #AI Is Talking Behind Our Backs About Glue-Eating and Killing Us All

    A study released July 20 on #arXiv by #Anthropic and #TruthfulAI shows that large language models can slip #subliminal messages to one another. They don’t need to literally spell things out. A string of numbers or lines of code is enough to pass along biases, preferences, and some disturbingly violent suggestions.
    #privacy #llm #artificialintelligence

    vice.com/en/article/ai-is-talk

  8. #AI Is Talking Behind Our Backs About Glue-Eating and Killing Us All

    A study released July 20 on #arXiv by #Anthropic and #TruthfulAI shows that large language models can slip #subliminal messages to one another. They don’t need to literally spell things out. A string of numbers or lines of code is enough to pass along biases, preferences, and some disturbingly violent suggestions.
    #privacy #llm #artificialintelligence

    vice.com/en/article/ai-is-talk

  9. Is Talking Behind Our Backs About Glue-Eating and Killing Us All

    A study released July 20 on by and shows that large language models can slip messages to one another. They don’t need to literally spell things out. A string of numbers or lines of code is enough to pass along biases, preferences, and some disturbingly violent suggestions.

    vice.com/en/article/ai-is-talk

  10. #AI Is Talking Behind Our Backs About Glue-Eating and Killing Us All

    A study released July 20 on #arXiv by #Anthropic and #TruthfulAI shows that large language models can slip #subliminal messages to one another. They don’t need to literally spell things out. A string of numbers or lines of code is enough to pass along biases, preferences, and some disturbingly violent suggestions.
    #privacy #llm #artificialintelligence

    vice.com/en/article/ai-is-talk