home.social

#roleconfusion — Public Fediverse posts

Live and recent posts from across the Fediverse tagged #roleconfusion, aggregated by home.social.

fetched live
  1. LLM roles are supposed to separate user input, internal reasoning, tool results, and final answers. But if a model relies on the style of text instead of its actual source, forged reasoning can slip into the wrong place. That is the core risk behind role confusion and CoT forgery.

    More: techtonicshift.vivaldi.net/202

    #AI #Safety #LLMSecurity #PromptInjection #RoleConfusion