Tag
← all experiments
A 27B model whose weights are all −1, 0 or +1 — what it takes to run, and how fast it goes.
A 60M-parameter T5, distilled from a larger model, that cleans up raw speech on-device — split across the Neural Engine and the GPU.