nice, a tiny webgpu model instead of a grammar is such a clean idea. how does it hold up on languages with really