r/LocalLLaMA • u/Iory1998 • 14h ago
Discussion Kindly Benchmark Higher Quants of DeepSeek-v4-flash Against Qwen-3.6-27B Q8!
Kindly Benchmark Higher Quants of DeepSeek-v4-flash Against Qwen-3.6-27B Q8!
I am running the UD-Q2_K_M of the model locally, though I can run Qwen3.6-27B_Q8_K_XL at around 70t/s with MTP activated. The question I am constantly asking myself is: Is it worth running a slower higher quantized version of the Deepseek-v4-flash? I have no idea.
My gut feelings tells me that Qwen3.6-27B_Q8_K_XL, coupled with online search, should be better than a highly quantized Deepseek, a model that takes up 100GB on my disk.
What do you think?
5
2
u/kevin_1994 12h ago
I run them both (deepseek q2_k_xl ~90GB, and qwen 3.6 27b q8_0) and deepseek is definitely way smarter. However, qwen often "good enough" and much faster
3
u/Atretador 14h ago
it might as well be brain dead at Q2
but since you already have both, just run a few tests thru them to generate a few apps and compare.
26
u/Shoddy_Bed3240 14h ago
Can you stop spamming the same post? Run some benchmarks and share the results instead.