Small Language Models: Efficiency Over Hype
DAIS 2024 had multiple sessions on efficient model training, distillation, and domain-specific fine-tuning. I want to pull on the thread that connects all of them, because it points to a conclusion that's going to be uncomfortable for anyone who has been telling their organization they need the biggest,