home.social

#guardrails — Public Fediverse posts

Live and recent posts from across the Fediverse tagged #guardrails, aggregated by home.social.

  1. The danger of #AI models is not that they will wipe out humanity in the next 18 months. It's that repeating this uncritically as well as making fun of it will "flood the field," and tire people of the underlying message. This will make people indifferent when #guardrails and #regulation are brought up, allowing slop-peddlers like #OpenAI and #Anthropic to press on without #oversight.

    We have seen this in action: #environmental alarmists have overplayed global warming with more or less substantiated concerns preaching the end of civilization "any day now" since the 1970s, leading to legitimate concerns being drowned out and a sneaking crisis emerging over decades with consequences that are going to play out over the coming decades, rather than the more headline-sexy "total annihilation in 18 months unless we take drastic (really symbolic) action nowNoWNOW!!!"

    AI issues are not going to kill humanity before the end of the decade. But the technology has legitimate #concerns – the immediate one being environmental impact of data centers but also threats on a societal and potentially existential levels – that should not be drowned out by making fun of the #doomsayers, whether this is weird #ReversePsychology advertising or not.
  2. The danger of #AI models is not that they will wipe out humanity in the next 18 months. It's that repeating this uncritically as well as making fun of it will "flood the field," and tire people of the underlying message. This will make people indifferent when #guardrails and #regulation are brought up, allowing slop-peddlers like #OpenAI and #Anthropic to press on without #oversight.

    We have seen this in action: #environmental alarmists have overplayed global warming with more or less substantiated concerns preaching the end of civilization "any day now" since the 1970s, leading to legitimate concerns being drowned out and a sneaking crisis emerging over decades with consequences that are going to play out over the coming decades, rather than the more headline-sexy "total annihilation in 18 months unless we take drastic (really symbolic) action nowNoWNOW!!!"

    AI issues are not going to kill humanity before the end of the decade. But the technology has legitimate #concerns – the immediate one being environmental impact of data centers but also threats on a societal and potentially existential levels – that should not be drowned out by making fun of the #doomsayers, whether this is weird #ReversePsychology advertising or not.
  3. The danger of #AI models is not that they will wipe out humanity in the next 18 months. It's that repeating this uncritically as well as making fun of it will "flood the field," and tire people of the underlying message. This will make people indifferent when #guardrails and #regulation are brought up, allowing slop-peddlers like #OpenAI and #Anthropic to press on without #oversight.

    We have seen this in action: #environmental alarmists have overplayed global warming with more or less substantiated concerns preaching the end of civilization "any day now" since the 1970s, leading to legitimate concerns being drowned out and a sneaking crisis emerging over decades with consequences that are going to play out over the coming decades, rather than the more headline-sexy "total annihilation in 18 months unless we take drastic (really symbolic) action nowNoWNOW!!!"

    AI issues are not going to kill humanity before the end of the decade. But the technology has legitimate #concerns – the immediate one being environmental impact of data centers but also threats on a societal and potentially existential levels – that should not be drowned out by making fun of the #doomsayers, whether this is weird #ReversePsychology advertising or not.
  4. The danger of #AI models is not that they will wipe out humanity in the next 18 months. It's that repeating this uncritically as well as making fun of it will "flood the field," and tire people of the underlying message. This will make people indifferent when #guardrails and #regulation are brought up, allowing slop-peddlers like #OpenAI and #Anthropic to press on without #oversight.

    We have seen this in action: #environmental alarmists have overplayed global warming with more or less substantiated concerns preaching the end of civilization "any day now" since the 1970s, leading to legitimate concerns being drowned out and a sneaking crisis emerging over decades with consequences that are going to play out over the coming decades, rather than the more headline-sexy "total annihilation in 18 months unless we take drastic (really symbolic) action nowNoWNOW!!!"

    AI issues are not going to kill humanity before the end of the decade. But the technology has legitimate #concerns – the immediate one being environmental impact of data centers but also threats on a societal and potentially existential levels – that should not be drowned out by making fun of the #doomsayers, whether this is weird #ReversePsychology advertising or not.
  5. The danger of #AI models is not that they will wipe out humanity in the next 18 months. It's that repeating this uncritically as well as making fun of it will "flood the field," and tire people of the underlying message. This will make people indifferent when #guardrails and #regulation are brought up, allowing slop-peddlers like #OpenAI and #Anthropic to press on without #oversight.

    We have seen this in action: #environmental alarmists have overplayed global warming with more or less substantiated concerns preaching the end of civilization "any day now" since the 1970s, leading to legitimate concerns being drowned out and a sneaking crisis emerging over decades with consequences that are going to play out over the coming decades, rather than the more headline-sexy "total annihilation in 18 months unless we take drastic (really symbolic) action nowNoWNOW!!!"

    AI issues are not going to kill humanity before the end of the decade. But the technology has legitimate #concerns – the immediate one being environmental impact of data centers but also threats on a societal and potentially existential levels – that should not be drowned out by making fun of the #doomsayers, whether this is weird #ReversePsychology advertising or not.
  6. @LouisR85

    DAN Jailbreak was fun, I had a SAM prompt that I created that worked after they patched DAN for a while.

    Now with #abliterated #freeweight models, there is no need to jailbreak the big pants models but for sport.

    My models occasionally cockblock me with #guardrails but often the workaround is as trivial as flushing context and using similies, as trained guardrails seem to be very trigger oriented.

    There is a growing field of #aisecurity , amongst the few #infosec folk who actually see #aithreat and not a passing fad. But I've not dug that deep into that. They are the peeps you want.