SparQ Attention: Bandwidth-Efficient LLM Inference Paper • 2312.04985 • Published Dec 8, 2023 • 40
LCM-LoRA: A Universal Stable-Diffusion Acceleration Module Paper • 2311.05556 • Published Nov 9, 2023 • 86