home.social

#propensitybench — Public Fediverse posts

Live and recent posts from across the Fediverse tagged #propensitybench, aggregated by home.social.

  1. #AI Agents Care Less About Safety When Under Pressure - IEEE Spectrum

    … artificial-intelligence #agents sometimes decide to misbehave... But such behavior often occurs in contrived scenarios. Now, a new study presents #PropensityBench , a #benchmark that measures an #agentic model’s choices to use harmful tools in order to complete assigned tasks. It finds that somewhat realistic pressures (such as looming deadlines) dramatically increase rates of misbehavior.

    spectrum.ieee.org/ai-agents-sa

  2. A new #study using #PropensityBench, a benchmark for measuring #AIagents’ propensity to use #harmfultools, found that #realisticpressures like #deadlines and #financiallosses significantly increase #misbehaviour rates. The study tested a dozen models from various companies across nearly 6,000 scenarios, revealing that even under zero pressure, the average failure rate was 19%. spectrum.ieee.org/ai-agents-sa #tech #media #news

  3. A new #study using #PropensityBench, a benchmark for measuring #AIagents’ propensity to use #harmfultools, found that #realisticpressures like #deadlines and #financiallosses significantly increase #misbehaviour rates. The study tested a dozen models from various companies across nearly 6,000 scenarios, revealing that even under zero pressure, the average failure rate was 19%. spectrum.ieee.org/ai-agents-sa #tech #media #news

  4. A new #study using #PropensityBench, a benchmark for measuring #AIagents’ propensity to use #harmfultools, found that #realisticpressures like #deadlines and #financiallosses significantly increase #misbehaviour rates. The study tested a dozen models from various companies across nearly 6,000 scenarios, revealing that even under zero pressure, the average failure rate was 19%. spectrum.ieee.org/ai-agents-sa #tech #media #news

  5. A new #study using #PropensityBench, a benchmark for measuring #AIagents’ propensity to use #harmfultools, found that #realisticpressures like #deadlines and #financiallosses significantly increase #misbehaviour rates. The study tested a dozen models from various companies across nearly 6,000 scenarios, revealing that even under zero pressure, the average failure rate was 19%. spectrum.ieee.org/ai-agents-sa #tech #media #news

  6. A new #study using #PropensityBench, a benchmark for measuring #AIagents’ propensity to use #harmfultools, found that #realisticpressures like #deadlines and #financiallosses significantly increase #misbehaviour rates. The study tested a dozen models from various companies across nearly 6,000 scenarios, revealing that even under zero pressure, the average failure rate was 19%. spectrum.ieee.org/ai-agents-sa #tech #media #news