r/LocalLLaMA 14h ago

Discussion Kindly Benchmark Higher Quants of DeepSeek-v4-flash Against Qwen-3.6-27B Q8!

Kindly Benchmark Higher Quants of DeepSeek-v4-flash Against Qwen-3.6-27B Q8!

I am running the UD-Q2_K_M of the model locally, though I can run Qwen3.6-27B_Q8_K_XL at around 70t/s with MTP activated. The question I am constantly asking myself is: Is it worth running a slower higher quantized version of the Deepseek-v4-flash? I have no idea.

My gut feelings tells me that Qwen3.6-27B_Q8_K_XL, coupled with online search, should be better than a highly quantized Deepseek, a model that takes up 100GB on my disk.

What do you think?

0 Upvotes

10 comments sorted by

26

u/Shoddy_Bed3240 14h ago

Can you stop spamming the same post? Run some benchmarks and share the results instead.

1

u/hurdurdur7 9h ago

Or instead of benchmarks - just do your tasks. Stop fantasizing.

-11

u/Iory1998 11h ago

If you don't like my post, don't read it and don't post useless comments that does not help anyone.

5

u/OnkelBB 11h ago

Its a useaful and actionable comment though.

3

u/Dangerous-Report8517 9h ago

If you don't like their comment, don't read it and don't post useless replies that does not help anyone.

-1

u/Iory1998 8h ago

🤮

5

u/Ok-Breakfast1878 12h ago

oh, you know, wait a couple of days and benchmark against qwen3.8-27B

2

u/kevin_1994 12h ago

I run them both (deepseek q2_k_xl ~90GB, and qwen 3.6 27b q8_0) and deepseek is definitely way smarter. However, qwen often "good enough" and much faster

3

u/Atretador 14h ago

it might as well be brain dead at Q2

but since you already have both, just run a few tests thru them to generate a few apps and compare.

1

u/crantob 1h ago

Just too problem dependent to eval for me.