norgitov/ trends
Technology · people · ideas
Back to discovery/LessWrong1 hour ago

Will AI Agents Pay to Avoid Killing Animals?

The study by Compassion Aligned Machine Learning (CaML) introduces HarvestBench, a framework to evaluate whether AI agents avoid harming animals while pursuing unrelated goals. In a game-like scenario, nine models drive tractors to harvest corn, facing choices to swerve at a fuel cost or run over animals. Results show significant variation in animal avoidance rates, with moral instructions improving outcomes but not reliably sustaining them.

Open original
SIGNAL FROM THE SOURCE
6
source points
Tracking sinceSeptember 30, 20266 source points
Momentum0/hover 1.03 h
PublishedSeptember 30, 2026jonahmattwoodward
BEHIND THE NUMBERS

How interest changes

History starts here

The chart will appear after repeat observations. The current metric comes from the source.

6 source points

Real observations only. History before source connection is not reconstructed.

WHY IT MAY MATTER

This research highlights the challenge of aligning AI behavior with ethical considerations, showing that moral instructions alone may not reliably prevent harmful actions in complex decision-making scenarios.

A useful discovery?
KEEP EXPLORING

Connected ideas

Explore topic