LFM 2.5 230M runs at 1440 tok/s in-browser via a custom backend, with WebGPU kernels optimized for Nvidia and Apple Silicon and deployments in Electron/Tauri apps.
Read the original at old.reddit.com→Everything runs through WebGPU, in-browser or in electron/tauri apps. It's fully portable and supports either Nvidia and Apple Silicon (Metal). The actual kernels are optimized for the specific hardware of the...
Original headline: "LFM 2.5 230M running at 1440 tok/s in-browser through a custom backend"
Coverage timeline
- Jul 25, 17:14 UTC r/LocalLLaMA lead source LFM 2.5 230M running at 1440 tok/s in-browser through a custom backend