r/LocalLLaMA · · 1 min read

DeepSeek-V4-Flash on SM89 4x48gb 4090s with DSpark

Mirrored from r/LocalLLaMA for archival readability. Support the source by reading on the original site.

DeepSeek-V4-Flash on SM89 4x48gb 4090s with DSpark

https://github.com/yhfgyyf/vllm-deepseek-v4-sm89

I couldn't believe that someone actually got vLLM working with this particular set of GPUs, but here it is. The video is from right after I got it working with 64k context, but it is now running with 256k.

submitted by /u/dangerous_inference
[link] [comments]

Discussion (0)

Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.

Sign in →

No comments yet. Sign in and be the first to say something.

More from r/LocalLLaMA