FP6-LLM: Efficiently Serving Large Language Models Through FP6-Centric Algorithm-System Co-Design Paper โข 2401.14112 โข Published Jan 25, 2024 โข 18
Generative Multimodal Models are In-Context Learners Paper โข 2312.13286 โข Published Dec 20, 2023 โข 34