r/LocalLLaMA · · 1 min read

Any chance I could cluster my DGX Spark (128GB unified memory) and my AMD Ryzen AI Max 395 (128GM unified memory) together to run 1 model?

Mirrored from r/LocalLLaMA for archival readability. Support the source by reading on the original site.

Hey all,

So I have a Nvidia DGX Spark and an AMD Strix 395, both have 128GB of unified memory. The Spark has 200Gbit network and the AMD Strix has 5Gbit ethernet (but it has a pcie gen 4x4 slot).

Is there any chance I can cluster the 2 together to run a larger model that can fit in the ~256GB (minus OS) unified memory?

I've seen that Deepseek v4 Flash can fit on 2x DGX Spark, but maybe I can use my Strix system instead?

Any ideas on if this would be possible? If so, how would you go abouts doing it? Would it help if I added a Mellanox ConnectX-6 QSFP+28 to the AMD Strix and connected it to the DGX Spark? I would have maybe 64Gbps over the 100Gbps link, but 64 is faster than 5.

Thoughts?

Thanks!

submitted by /u/StartupTim
[link] [comments]

Discussion (0)

Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.

Sign in →

No comments yet. Sign in and be the first to say something.

More from r/LocalLLaMA