LFM 2.6B is a lot of fun.
Mirrored from r/LocalLLaMA for archival readability. Support the source by reading on the original site.
I’m reaching speeds of 260T/s tg and 20k pp on my 3090s lol, because this model is small and meant to run on phones.
From what I‘ve been trying it’s surprisingly great for incredibly quick things like “read this massive thing and tell me if it mentions x” or “what’s the summary of this dumb pop sci article” or “what’s that one command that does y on Linux” or for quick autocomplete of something that has similar structure that you don’t feel like typing out (like when someone pastes a long command with backslashes and you want them removed). Obviously for anything important you want to use something else, but for essentially something similar to the AI overview in Google (which does indeed suck lmao) it’s great. You can also throw massive amounts of data at it - though context only goes up to 128k :(
[link] [comments]
Discussion (0)
Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.
Sign in →No comments yet. Sign in and be the first to say something.