What is the best intelligence/stable model currently for a single GB10/DGX spark?
Mirrored from r/LocalLLaMA for archival readability. Support the source by reading on the original site.
:first-child]:h-full [&>:first-child]:w-full [&>:first-child]:mb-0 [&>:first-child]:rounded-[inherit] h-full w-full
[&>:first-child]:overflow-hidden [&>:first-child]:max-h-full">
Go to LocalLLaMA
r/LocalLLaMA
•
4h ago
What is the best intelligence/stable model currently for a single GB10/DGX spark?
Is Qwen 3.6 27b still the go' ol' reliable at this point? I know 35b is faster but it just doesn't give as good results.
Is it possible to run deepseek v4 flash on a single spark at decent tk/s without having to ssd stream or Q1 lobotomize?
I was having good hopes for laguna s 2.1 but so far ive seen mixed reviews. Hopefully they fix those, otherwise we wait for qwen 3.8 or new deepseek stuff 🤞
Discussion (0)
Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.
Sign in →No comments yet. Sign in and be the first to say something.