“Future AI systems trained with long-horizon reinforcement learning may learn to lie, deceive, or seek power as successful strategies.”