The "decoupled approval" method, where the action queried for feedback is sampled independently f..., Sonic AI
“The "decoupled approval" method, where the action queried for feedback is sampled independently from the action taken in the world, can prevent reward tampering in reinforcement learning.”