FP6-LLM: Efficiently Serving Large Language Models Through FP6-Centric Algorithm-System Co-Design Paper โข 2401.14112 โข Published Jan 25, 2024 โข 21 โข 7
The Impact of Reasoning Step Length on Large Language Models Paper โข 2401.04925 โข Published Jan 10, 2024 โข 18 โข 2