r/LocalLLaMA · · 1 min read

What is the best intelligence/stable model currently for a single GB10/DGX spark?

Mirrored from r/LocalLLaMA for archival readability. Support the source by reading on the original site.

:first-child]:h-full [&>:first-child]:w-full [&>:first-child]:mb-0 [&>:first-child]:rounded-[inherit] h-full w-full [&>:first-child]:overflow-hidden [&>:first-child]:max-h-full">
Go to LocalLLaMA

What is the best intelligence/stable model currently for a single GB10/DGX spark?

Is Qwen 3.6 27b still the go' ol' reliable at this point? I know 35b is faster but it just doesn't give as good results.

Is it possible to run deepseek v4 flash on a single spark at decent tk/s without having to ssd stream or Q1 lobotomize?

I was having good hopes for laguna s 2.1 but so far ive seen mixed reviews. Hopefully they fix those, otherwise we wait for qwen 3.8 or new deepseek stuff 🤞

Discussion (0)

Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.

Sign in →

No comments yet. Sign in and be the first to say something.

More from r/LocalLLaMA