How important is it for Chinese LLMs to reach the Opus 4.8 level?
Mirrored from r/LocalLLaMA for archival readability. Support the source by reading on the original site.
| In mid-August, Ramp published spending data collected from 70,000 U.S. companies: Fable 5 ,the most powerful and expensive model in Anthropic’s lineup, accounts for just 11% of what those businesses spend on the company’s tools. The remaining 79% is worth its weight in gold. With the new releases from Qwen and GLM, we are likely close to Opus 4.8, and certainly ahead of Sonnet and the other LLMs shown at the top of the image. The "anti-open-source crusade" therefore comes as no surprise: it is a genuine threat to their business, especially considering the parallel with the video game industry, where the hardware needed to run a game in Full HD became "low-end" within just a few years as 4K took over. We are in a frenetic phase: new models appearing daily, varying sizes, and—unfortunately—increasingly expensive hardware. But in the long run, I believe the winners will be those selling the silicon for computing (AMD, Nvidia, and soon other competitors) rather than those selling tokens. [link] [comments] |
More from r/LocalLLaMA
-
an unscientific qwen 3.8 flash next and glm 5.3 flash comparison
Aug 30
-
Ran Qwen3.8-Flash-Next (79 GB, 2-bit) at 350K ctx for 3.5 hours on a 128 GB M5 Max — speed vs context depth, 100 turns, one graph
Aug 30
-
Nemotron-3.5-Lightning at 11.77 GiB, a 16 GB option for a model that didn't have one
Aug 29
-
Humaneval benchmark for Deepseek V4 Flash 0731 vs GLM5.3 Flash on 2x DGX Spark setup
Aug 29
Discussion (0)
Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.
Sign in →No comments yet. Sign in and be the first to say something.