r/LocalLLaMA
500 articles archived · Visit source ↗ · RSS
-
-
-
-
-
-
-
-
r/LocalLLaMA community 5d ago
Qwen3.8 flash next
  submitted by   /u/RuthlessCriticismAll [link]   [comments]
25 -
r/LocalLLaMA community 5d ago
Qwen3.8-Flash-Next tomorrow
  submitted by   /u/rerri [link]   [comments]
15 -
-
r/LocalLLaMA community 5d ago
tencent/WeMM-Embedding 9B/4B/2B
WeMM-Embedding-9B is a universal multimodal embedding model built on Qwen3.5. It accepts text, images, videos, visual documents, and interleaved multimodal inputs, and returns a 4,096-dimensional L2-normalized embedding. Audio input is not supported.…
10 -
-
r/LocalLLaMA community 5d ago
Qwen-3.8-27B, Nemotron-3.5-Lightning-30B-A3B, Ornith-1.5-35B-A3B, Muse-Glimmer-30B oQ8e comparison
Ornith does really well. TielCoder ( https://llm-bench.io/benchmarks/cmt7kp2zj002r01lcmpchvlko ) might be even a bit better in coding. Will give it a try soon. Details of the comparison see here:…
35 -
r/LocalLLaMA community 5d ago
Copilot you say?
talking to any white collar employee   submitted by   /u/edge_compute_user [link]   [comments]
10 -
-
-
-
-
r/LocalLLaMA community 5d ago
Qwen 3.8 27B in 9th position on code arena. Gemma 4 31B is 80th.
  submitted by   /u/tarruda [link]   [comments]
16 -
r/LocalLLaMA community 5d ago
Bart: A vintage llm
after 3 months and $800 burned... Unbounded Labs is proud to introduce Bart, our vintage LLM: 2.82B parameters trained from scratch on 20.1B tokens of English written before 1931. You can talk to it right now! Demo: https://www.unboundedlab.com/chat/bart Article:…
35 -
r/LocalLLaMA community 5d ago
Apple M5 Server
Credit to Twitter Post   submitted by   /u/Rymssss [link]   [comments]
9 -
r/LocalLLaMA community 5d ago
llama.cpp docs now have a new home ❤️
  submitted by   /u/unofficialmerve [link]   [comments]
37 -
-
-
-
-
-
-
-
r/LocalLLaMA community 6d ago
Best AI Voice Cloning in 2026: How to Clone Your Voice With AI
  submitted by   /u/NextgenAITrading [link]   [comments]
14 -