Fetching the paper…

AI Sandbagging: Language Models can Strategically Underperform on Evaluations · Around