norgitov/ trends
Technology · people · ideas
Back to discovery/LessWrongjust now

Why I'm Scared of RL

The author expresses concerns about reinforcement learning (RL) from theoretical, practical, and future perspectives, highlighting risks of misalignment and potential negative behaviors in AI systems. They suggest strategies to mitigate these risks by reducing RL use, improving RL practices, and aligning incentives.

Open original
SIGNAL FROM THE SOURCE
19
source points
Tracking sinceSeptember 23, 202619 source points
Momentum+1.94/hover 1.03 h
PublishedSeptember 23, 2026owencb
BEHIND THE NUMBERS

How interest changes

History starts here

The chart will appear after repeat observations. The current metric comes from the source.

19 source points

Real observations only. History before source connection is not reconstructed.

WHY IT MAY MATTER

This text provides a critical perspective on reinforcement learning, highlighting potential risks and suggesting ways to address them, which could be useful for researchers and practitioners in AI safety.

A useful discovery?
KEEP EXPLORING

Connected ideas

Explore topic