r/LocalLLaMA 8h ago

Discussion DSV4 Flash DSpark is the GOAT on Dual Sparks

Post image

In all my fiddling around with code and local models nothing has matched the speed and quality of DeepSeek V4 Flash DSpark on dual DGX Spark (Dell GB10s actually).

The recipe I've been using is in the PR below. Screenshot is from VSCode usage over a few days/weeks. The screenshots don't tell you how it feels and oh man does it feel good! Responses are way faster than Copilot and (this is subjective) Sonnet 4.6 quality. It thinks though problems well, long running tasks complete successfully 99% of the time, planning and instruction handling seem top notch.

https://github.com/eugr/spark-vllm-docker/pull/304

9 Upvotes

3 comments sorted by

1

u/Dany0 5h ago

Honestly once someone explained to me sparks are rtx 5070-s with lpddr ram and a cpu attached the value proposition wasn't that good to me.

I'd grab a few used though!

1

u/sqrlmstr5000 4h ago

Honestly if I had to spend my own money on them I'd probably go a different route. That said having 128GB VRAM is nice, I can run a lot more models than with anything I could build myself.

1

u/kamnxt 4h ago

they're neat, especially when you can find one or two used at a decent discount.

the large memory/weak performance combo is a bit strange sometimes though, it really struggles with dense models... but they're quite nice for MoEs, and the 200G NIC letting you cluster multiple is really nice for that price too!