Fetching the paper…

PermLLM: Private Inference of Large Language Models within 3 Seconds under WAN · Around