LFM2-350M
354M
params · custom WGSL kernels ·
—
·
64-chat wall
·
bench
—
tok/s
run
LFM2.5-350M
WebGPU inference · 10 conv blocks · 6 attention layers · zero dependencies
downloading the model —
269 MB
one time only: next visits load from the browser cache