##article.return##
ELUTQ: Efficient LUT-Aware Quantization for Deploying Large Language Models on Edge Devices
Download
Download PDF