Fetching the paper…

Interpreting and Controlling Vision Foundation Models via Text Explanations · Around