“Significant reinforcement learning (RL) on tasks similar to scientific discovery is necessary for models to begin making novel discoveries.”