2026LFM2.5-230M · LoRA · end-of-turn
Fine-tuning a 230M model to know when you've stopped talking
A from-scratch 5M-parameter classifier plateaued at 30% recall on real speech. A LoRA fine-tune of a 230M pretrained model, on the same data, didn't.
Read →