ShortGPT: Layers in Large Language Models are More Redundant Than You Expect Paper • 2403.03853 • Published Mar 6, 2024 • 66
Layer-Wise Quantization: A Pragmatic and Effective Method for Quantizing LLMs Beyond Integer Bit-Levels Paper • 2406.17415 • Published Jun 25, 2024 • 1
Merge Experiments Collection Sorted from oldest (top) to newest (bottom) • 147 items • Updated 3 days ago • 4
Finetune Experiments Collection Sorted from oldest (top) to newest (bottom) • 22 items • Updated 5 days ago • 1