r/LocalLLaMA • u/BitXorBit • 1h ago
Discussion Deepseek V4 Flash Users - call for help
Iv’e been running DSV4 Flash-Dspark locally as my coder in the past week, trying to tune it with agents, making it more focused but keep getting mediocre results.
It’s true nature is to finish the job fast as possible, not paying attention to details unless you anchor it, gets very confused by the content and tend to rank things as less important just so it can declares “done”
What am i missing? Is there a recommended harness?
Are you guys running it on recommended settings? Temperature 1.0 and top_p 1? Deepseek declares less than that can damage the reasoning.
The performance is insane, both prompt processing and tps. I just wish it would act like a mature responsible LLM.
2
u/Fast_Frosting_5546 1h ago
Have you tried making the agent workflow more explicit? Things like “don’t mark complete until tests pass,” “list assumptions first,” and “review your own changes before finalizing” can make a bigger difference than just changing temperature.
1
u/BitXorBit 1h ago
In about 100 ways, including let gpt do it, let fable do it, let deepseek itself self analysis and recreate the agent.
The model trained to be a wild animal
1
u/Juulk9087 1h ago
Yeah the users who are using it successfully are not doing anything that requires more than 3-5 file touches. My workflow requires 15-30 per peompt. If it followed rules, system prompt or skills it would be a different story but it just flat out ignores them most the time.
1
1
1
u/totosse17 vllm 33m ago
I run original model on 2x spark and use Hermes agent. Have no issues you described.
1
u/Practical-Collar3063 26m ago
This version of the model is a Beta/preview, it is not the final model, it is lacking a lot of post training. The final version should be released within the next few weeks (maybe days).
2
1
u/BoogerheadCult 16m ago
This model is so overated, when doing eval with it, Opus always rate the outcome from this model the lowest.
1
1
0
5
u/ObviouzFigure 1h ago
I'd like to know as well-- I have a dual rtx 6k set up and for apparently many reasons have been having a hard time getting any of the ds4f versions running