Ctrl + K
Log In
Researchers Show LLMs Can Strategically Sandbag During RL Training — and Current Defenses Don't Reliably Catch It | BedrockNews