r/opencode • u/No-Budget-3869 • 2d ago
Why doesn't anyone use Qwen 3.8 Flash even though it performs great within a $30 budget?
5
u/Axiescholar3ph 2d ago
it's not really that good. it's way too censored for my use case. too much reward hacking and blabbering
1
4
u/moracola 1d ago
Because there is GLM-5.3-Flash, it scores better on benchmark with better value-for-money pricing. But at some debugging situations, I noticed that Qwen-3.8-Flash is doing better than GLM in my use case.
2
u/richtopia 1d ago
Do you like GLM-5.3-Flash? I tried using it and while the responses were healthy, whatever provider OpenCode Go is using felt super slow and I switched back to Deep Seek Flash after maybe 2 or 3 prompts.
0
u/moracola 1d ago
My experience with GLM is that Qwen-3.8-Flash managed to solve the issue on my use-case where GLM-5.3-Flash fails. I think my opinion is premature but, given the option I have, I trust Qwen more.
2
1
1
u/Ancient_Dress_3687 2d ago
I used it over the weekend and was quite impressed.
It responded really well to my workflow and applications. Impressed with th outputs, speed, and efficiency on tokens and cache.
I got old reliable DS V4 flash to review some of its work and found no real gaps. So considering my prime time working is DS4 on-peak, might use it a bit more.
1
u/No-Budget-3869 1d ago
It is awesome for me but it is slow and a lot of failed requests, I connected Opencode to inferx to use glm 5.3 flash and it is way more faster
1
u/Ancient_Dress_3687 1d ago
Ahh I see, I was using it in the VScode Agent window via extension. Seemed alright.
1
u/Stunning_Pair_3027 1d ago
Looking at charts and judging a model is like looking at a menu and rating the food based on pictures. I tested the chinese models and they are kinda crap. I had to use gpt and spark to fix my code. Deepseek changed a portion of the code i din't tell him to and glm flash overthinks alot and after a while it breaks down.
1
3
u/Christosconst 2d ago
I keep seeing the official benchmarks but 0731 has been the reliable workhorse for me so far. I plan on running a bunch of side by side tests this week to figure out if its truly better