Running a decision model locally on an RTX 4090 to compare speed across open models (Laya, Liquid D1, Cloudflare Clef-Flash, Interfaze Lev) on a centipede-reading task
Read the original at www.reddit.com→recently saw a bunch of open decision models pop out of nowhere in the last two weeks (laya, liquid's d1, cloudflare's clef-flash, interfaze's lev), so I wanted to see how far apart they actually are on the same...
Original headline: "Running decision model locally on an RTX 4090 to find out which one is the fastest"
Coverage timeline
- Oct 8, 17:04 UTC r/LocalLLaMA lead source Running decision model locally on an RTX 4090 to find out which one is the fastest