POV 8GB VRAM + 16GB RAM folks checking locallama everyday:
Mirrored from r/LocalLLaMA for archival readability. Support the source by reading on the original site.
| One can only hope for Qwen4 35B A3B or similar (maybe with extra ngrams for 35B -> 70B+... until then, Gemma 26B QAT my beloved (solid 26 Tg/s)... [link] [comments] |
More from r/LocalLLaMA
-
Yes bots we get it, Strata is good now please stop
Oct 3
-
Aleph-Alpha/Kolibri-1 · Hugging Face - 78B parameters. 3.46B active. Up to 1M tokens of context - Apache 2.0
Oct 3
-
Anyworld, a self-hosted multiplayer text RPG where a local LLM is the Dungeon Master
Oct 3
-
I tried building a small RAG search node for Qwen3.8 27B using a fake AliExpress Mini PC... and Intel sent me back to 2018.
Oct 3
Discussion (0)
Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.
Sign in →No comments yet. Sign in and be the first to say something.