LFM2-350M 354M params · custom WGSL kernels · · 64-chat wall · bench
tok/s
LFM2.5-350M
WebGPU inference · 10 conv blocks · 6 attention layers · zero dependencies
downloading the model — 269 MB
one time only: next visits load from the browser cache