Tag
A from-scratch 5M-parameter classifier plateaued at 30% recall on real speech. A LoRA fine-tune of a 230M pretrained model, on the same data, didn't.
Read →A small adapter turns a frozen left-to-right model into a one-pass, both-directions corrector.
Read →Making a text-to-speech model decode in one pass instead of sixteen, so it runs fast on-device. With audio you can play.
Read & listen →