Qwen3.8 27b just exceeded my expectations on svg generation :D
Mirrored from r/LocalLLaMA for archival readability. Support the source by reading on the original site.
| https://reddit.com/link/1vtkgdj/video/595yn0ckdjkh1/player I wanted to try out Qwen3.8 27B 's SVG capabilities but with something different than the pelican on a bicycle. Promt was literally just :
took 20 minutes (12 of that was just thinking - xhigh) i was questioning if it would ever be finished :D but when it was done i was floored , i dont know what i expected but definetly not that! 46.436 tokens were burned at around 40t/s (it started at around ~60, but then i removed the powerlimit (250 => 370) and after that it had drops in the 20s , might have something to do with doing that mid generation, might need to try again without touching any settings to see if the model itself had some hickups after long generation. Model is https://huggingface.co/cyankiwi/Qwen3.8-27B-AWQ-INT4 at tp2 on 2x 3090 , fp8 KV cache. EDIT: because of downvote: i copied the promt and response (including reasoing) into a pastebin incase someone doubts https://pastebin.com/FfSutPfn can also provide screenshots of the request in llama swap [link] [comments] |
More from r/LocalLLaMA
-
Uncensored Multi-Model Releases, LongCat-Flash-Lite-Sparse with MTPs and LSAs, Qwen3.8-27B with MTPs, Qwen3.5-122B-A10B with MTPs, Qwen3-Coder-Next and Laguna-S2.1 with Vision, All Available in GGUF Format! Bonus: Links to my llama.cpp Fork for LongCat-Flash-Lite Support and…
Aug 30
-
Got MiniMax H3 video generation running in TensorSharp
Aug 30
-
Qwen3.8-Flash-Next NVFP4 2xDGX Spark config: 50t/s decode, 2,900t/s prefill
Aug 30
-
Don't Sleep on EXL3 Quants
Aug 30
Discussion (0)
Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.
Sign in →No comments yet. Sign in and be the first to say something.