Is it worth running Qwen 3.8 Flash Next on 4x3090 vs 27B?
Mirrored from r/LocalLLaMA for archival readability. Support the source by reading on the original site.
Can someone please tell me if it's worth running Qwen 3.8 Flash Next on 4x3090 yet over 27B?
27B is good but damn it is indecisive. I am getting frustrated watching it get "so close" to solving a problem, only to do another 2 hours of "let me just check/prove/etc"
It looks like a 4 bit quant of Flash Next should fit with the ngrams in SSD and be a lot faster but it also sounds like the architecture isn't quite there yet
Can someone smarter and more patient than me tell me what to do pls? thanks
[link] [comments]
Discussion (0)
Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.
Sign in →No comments yet. Sign in and be the first to say something.