Qwen 3.8 27b - PI AGENT vs OPENCODE - another smaple
Mirrored from r/LocalLLaMA for archival readability. Support the source by reading on the original site.
| That is the second comparison and the last one. I will not be spamming again ;) Continuation from: That is one of my many tests I make comparing output quality. What is more interesting using a PI Agent results are much better than an Opencode using a Qwen 3.8 27b ?! Seems PI Agent is much better in the agent environment somehow... Not counting uses less tokens , do not have a hard limit of 32k output tokens, is faster, do not freezing, compressing context far less than Opencode. For instance if you have context in the Opencode output 32k and all context 100k then the compression is starting at 67k context ... PI is starting at 90k context even if you have set output context 64k or more. My config for RTX 3090 llama-server with ini config -> which is exposing API to Opencode and PI agent.
config ini ONE MORE IMPORTANT THING: Always use a VISION module as the model is using vision to asses the output quality! I am offloading it to a RAM as we do not need an extremely fast vision for a code. A screenshot processing on a GPU 0.3s vs a RAM 3s do not make a big difference on a few screenshots during a code generation / debugging ;) PROMPT: SECOND PROMPT AFTER THE FIRS IS FINISHED: [link] [comments] |
Discussion (0)
Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.
Sign in →No comments yet. Sign in and be the first to say something.