ORIGINAL REDDIT POST

Kindly Benchmark Higher Quants of DeepSeek-v4-flash Against Qwen-3.6-27B Q8!

Kindly Benchmark Higher Quants of DeepSeek-v4-flash Against Qwen-3.6-27B Q8! I am running the UD-Q2_K_M of the model locally, though I can run Qwen3.6-27B_Q8_K_XL at around 70t/s with MTP activated. The question I am constantly asking myself is: Is it worth…

Original postr/LocalLLaMA

Kindly Benchmark Higher Quants of DeepSeek-v4-flash Against Qwen-3.6-27B Q8! I am running the UD-Q2_K_M of the model locally, though I can run Qwen3.6-27B_Q8_K_XL at around 70t/s with MTP activated. The question I am constantly asking myself is: Is it worth running a slower higher quantized version of the Deepseek-v4-flash? I have no idea. My gut feelings tells me that Qwen3.6-27B_Q8_K_XL, coupled with online search, should be better than a highly quantized Deepseek, a model that takes up 100GB on my disk. What do you think?

Collected discussion

12 comments

u/Shoddy_Bed3240

Can you stop spamming the same post? Run some benchmarks and share the results instead.

u/OnkelBB

Its a useaful and actionable comment though.

u/hurdurdur7

Or instead of benchmarks - just do your tasks. Stop fantasizing.

u/Iory1998OP

But that could serve as a basis for comparison.

u/crantob

I like DS4-flash way better, for my applications.

u/Ok-Breakfast1878

oh, you know, wait a couple of days and benchmark against qwen3.8-27B

u/Iory1998OP

If you don't like my post, don't read it and don't post useless comments that does not help anyone.

u/Dangerous-Report8517

If you don't like their comment, don't read it and don't post useless replies that does not help anyone.

u/crantob

Just too problem dependent to eval for me.

u/kaliku

If you don't like his comment, don't read it.

u/Atretador

it might as well be brain dead at Q2 but since you already have both, just run a few tests thru them to generate a few apps and compare.

u/kevin_1994

I run them both (deepseek q2_k_xl ~90GB, and qwen 3.6 27b q8_0) and deepseek is definitely way smarter. However, qwen often "good enough" and much faster

Kindly Benchmark Higher Quants of DeepSeek-v4-flash Against Qwen-3.6-27B Q8!