r/opencode 2d ago

Why doesn't anyone use Qwen 3.8 Flash even though it performs great within a $30 budget?

Post image
27 Upvotes

19 comments sorted by

3

u/Christosconst 2d ago

I keep seeing the official benchmarks but 0731 has been the reliable workhorse for me so far. I plan on running a bunch of side by side tests this week to figure out if its truly better

0

u/No-Budget-3869 2d ago

Qwen 3.8 flash has much higher benchmark than deepseek 0731

2

u/Christosconst 2d ago

That's what I'm saying, I keep seeing Qwen3.8 top 0731 on paper, but still use 0731 because of how reliable it's been

5

u/Axiescholar3ph 2d ago

it's not really that good. it's way too censored for my use case. too much reward hacking and blabbering

1

u/No-Budget-3869 2d ago

It can be a useful sub-agent

4

u/moracola 1d ago

Because there is GLM-5.3-Flash, it scores better on benchmark with better value-for-money pricing. But at some debugging situations, I noticed that Qwen-3.8-Flash is doing better than GLM in my use case.

2

u/richtopia 1d ago

Do you like GLM-5.3-Flash? I tried using it and while the responses were healthy, whatever provider OpenCode Go is using felt super slow and I switched back to Deep Seek Flash after maybe 2 or 3 prompts.

0

u/moracola 1d ago

My experience with GLM is that Qwen-3.8-Flash managed to solve the issue on my use-case where GLM-5.3-Flash fails. I think my opinion is premature but, given the option I have, I trust Qwen more.

2

u/m2wm2wm2w 1d ago

"Endpoint unavailable" is the real answer

1

u/migsperez 2d ago

I'm thinking about it

1

u/Ancient_Dress_3687 2d ago

I used it over the weekend and was quite impressed.

It responded really well to my workflow and applications. Impressed with th outputs, speed, and efficiency on tokens and cache.

I got old reliable DS V4 flash to review some of its work and found no real gaps. So considering my prime time working is DS4 on-peak, might use it a bit more.

1

u/No-Budget-3869 1d ago

It is awesome for me but it is slow and a lot of failed requests, I connected Opencode to inferx to use glm 5.3 flash and it is way more faster

1

u/Ancient_Dress_3687 1d ago

Ahh I see, I was using it in the VScode Agent window via extension. Seemed alright.

1

u/Stunning_Pair_3027 1d ago

Looking at charts and judging a model is like looking at a menu and rating the food based on pictures. I tested the chinese models and they are kinda crap. I had to use gpt and spark to fix my code. Deepseek changed a portion of the code i din't tell him to and glm flash overthinks alot and after a while it breaks down.

1

u/kamwee 1d ago

Money

1

u/0mamii 1d ago

depending on what???

1

u/Plastic_Love3352 1d ago

Be careful, it might rm -rf something.

1

u/anramon 1d ago

I use it for answering general questions, for dedicated roles there are better models.

1

u/GTHell 1d ago

$30 is like a blink with opencode. You need it to be like Muse spark contrib