##article.return## Informed Routing in LLMs: Smarter Token-Level Computation for Faster Inference Download Download PDF