Fetching the paper…

Performance Characterization of Expert Router for Scalable LLM Inference · Around