250M parameter language model. 100M-token offline context. 60 MB deployment, 400 tok/s on a laptop CPU.