r/opencodeCLI 2d ago

Nearly unlimited

Post image
566 Upvotes

65 comments sorted by

119

u/ShirtuShanks 2d ago

20

u/msenc 2d ago

some people still didn't get the second half 😭

3

u/Time-Toe-1276 2d ago

besides all the jokes, I kinda like the model ngl. (obv till the discount ends 😭) which is like only another a week

28

u/AutomaticAd6646 2d ago

5 token per second? Is it even useful speed?

7

u/Jazzlike_Bee_3129 2d ago

I don't know when this change took effect, but I used it this morning just fine. 

2

u/Time-Toe-1276 2d ago

yeah same!

1

u/Khaledthe 2d ago

No thats what my old 5600xt can run

1

u/jpcaparas 2d ago

if you're living in the phantom zone, yes

1

u/ECrispy 2d ago

its what us poor people get with a 8b local model :)

16

u/zer0evolution 2d ago

really? i've got limited budget so need to plan carefully

2

u/ucharx 3h ago

It was a joke , he says it's unlimited because it's so slow that you can't use it

4

u/JamesGooning 2d ago

tokenrouter has free GLM5.3 (not flash version) for free.

1

u/Gallagger 1d ago

Probably not unlimited though.

2

u/rainpurplebow 1d ago

It's unlimited but, but, buuuuuuuuut...! 2 TPS :)

1

u/Gallagger 1d ago

Yes I get that. But still probably more, at least more than stuff like openrouter free models. Not sure about tokenrouter.

11

u/Historical_Cook_3485 2d ago

for long sessions with longs context , GLM-5.3-FLASH works really good for me i can say its near claude sonnet performance , honestly i started loving hope it stays like this

7

u/FammasMaz 2d ago

Btw Magic context plus context limited at 272k will do magic to you then. Not only do you get basically unlimited context, but also your model stays within that sweet 200k context window where its much more smarter.

Not affiliated with magic context in any way.

1

u/Historical_Cook_3485 2d ago

thanks i saw the repo looks the kind of thing im missing

1

u/some1else42 1d ago

Can you share the repo url or at least the user/project name on github? When a good project gets successful, you'd be suprised how many similar clones pop up that do nefarious things.

6

u/Jazzlike_Bee_3129 2d ago

It's the new deepseek, tbh, except better. 

4

u/spartanOrk 2d ago

Ungrammatical, unpunctuated. Illiterates.

0

u/Qubits119 2d ago

it's not that hard to understand

5

u/Sweaty_Cellist_4525 2d ago

Currently using Muse Spark 1.2 and it's hella good, absolutely free, and pairing it with Opus 5 for execution makes my Claude Pro sub last forever.

Meta did actually cook.

2

u/quicknades 2d ago

Are you in an area where the discounted version is applicable? Because in the EU it's not and then it's not really as cheap. Obviously still a good price but no where near the past deep seek pricing.

2

u/CriteriumA 2d ago

I use in Spain with OpenCode Go and free with Zen

1

u/quicknades 2d ago

The contributor version?

1

u/CriteriumA 2d ago

Yes. And work fine. Initially, I didn't like it at all; I thought it was terrible. But after many attempts, I finally got the hang of it.

It works very well for implementing codebase changes, but I monitor it with GLM 5.3 Flash, which is good and cheap for planning and reviewing, but bad and expensive for modifying the codebase. This combination is giving me good synergy.

1

u/Sweaty_Cellist_4525 2d ago

For me muse spark 1.2 is absolutely free (PerĂș), and ngl sometimes I just vpn to finish some urgent work.

The contributor API is so cheap tho, 0.20 usd per million output tokens, and it's a beast for execution (currently working on shaders and work so far has been wonderful), however if you don't have models like Opus to diagnose and plan it can still show amazing results, but shit your prompt's gotta be veery specific and clear.

2

u/LargePause 2d ago

Same here, using it alongside GPT SOL and it translates to pretty much unlimited use.

2

u/rainpurplebow 1d ago

Are you doing coding tasks with Muse Spark 1.2? Almost everyone in this sub hates that model.

1

u/Sweaty_Cellist_4525 1d ago

Yes, I do. Idk why they hate it, maybe because it's not that good for vibe coding? XD you need to know what you are doing and tell it exactly what you want, if not then you're gambling. For planning and desgining you should stay away unless you're broke, but mind you, I'm currently using it for writing complex compute shaders and mf gets the work done.

1

u/rainpurplebow 1d ago

It's all fun and games until it isn't :)

1

u/Sweaty_Cellist_4525 1d ago

As with everything really

1

u/zeamp 2d ago

I am literally switching to GLM after going around for 3 days on a batch 4,000-article rewrite. No matter what prompts I use, it eventually repeats whole paragraphs after about 500 flawless pages completed

Now I’m backing up every run and merging my “good” Spark rewrites with the crap it shits out 3 hours later. And it has decided to write files outside of the project directory
 so my desktop looks amazing. We can only seem to vibe for the first half of the day. I love it otherwise!

1

u/Academic_Constant42 2d ago

What's your setup to have opus call muse outside of Claude code? If you don't mind telling offcourse

1

u/SwisherSmoker420_ 2d ago

The way I do it in claude code I tell it to make an opencode.md file where it delegates tasks to opencode and then I tell opencode to read it and start working. Its a pretty rudimentary solution and theres definitely better ways to do it but it works for me.

1

u/throwaway12012024 2d ago

same here but within codex

2

u/Fun_Jaguar8231 2d ago

me on my z.ai lecacy coding plan, glm go brrr

2

u/tino1000 1d ago

Deepseek flash is fast, I hope GLM flash hits the highway soon

2

u/Kazekage1111 2d ago

Just use a different provider for inference to run this model. Check on openrouter and then get an api with the provider directly for best cache hits. I use run infra and with 200+ tps but there are other good providers

1

u/Metalwell 2d ago

So, should I get this to use flash? or openrouter top up is the way to go?

2

u/a355231 2d ago

Use Openrouter, routed to the official provider, it’s 50%

2

u/Fun_Squirrel5446 2d ago

Direct from z ai is cheapest because the api is currently 50% off.

After that open router is cheapest with the same 50% discount but 5.5% to load funds.

Open code go will be the best option after 9th Sept when the 50% discount expires.

3

u/Jazzlike_Bee_3129 2d ago

I think it's only cheapest if you use zcode.  If you have a specific ide you use, I would recommend opencode. 

1

u/look 2d ago

Synthetic has it in their subscription plan which should be about 1/3rd the full price. Also, Ollama has it as well, though not sure what the usage is like.

But I’ve been using RunInfra.ai PAYG which has had better speeds than Zai/Go and their price has lower cache read and works out to about the same as the OpenRouter 50% off with high cache rate (90+).

-4

u/CrimsonEdgeVentures 2d ago

Tried that model. Not great for agentics. Ok for very small tasks.

-7

u/SamePsychology8258 2d ago

A decision so expensive bro lost all of his hair

-21

u/PotterSkxawng 2d ago edited 2d ago

I have real unlimited GLM 5.3 (not Flash) for free through a provider... not going to tell which one it is tho, I'm sending 40,000 tokens per minute through sub-agent swarms rn and I dont want that stopping

Edit: Since you guys seem to want the provider—DM me, and if ur worthy, I'll share.

1

u/Mayanktaker 2d ago

Devin maybe

0

u/PotterSkxawng 2d ago

??????//

1

u/Mayanktaker 2d ago

Devin ide has glm 5.2 free till September. So i guess 5.3 also. Just guessing.

-1

u/PotterSkxawng 2d ago

Nope it's not Devin

-1

u/PotterSkxawng 2d ago

why am I getting downvoted... do yall want the provider

2

u/someoneyouknow23 2d ago

cause noone fucking cares if youre not gonna share?

-3

u/PotterSkxawng 2d ago

alr fine ill be generous... dm me and, if ur worthy, ull get the provider.

3

u/someoneyouknow23 2d ago

youre still gatekeeping through this

2

u/stylist-trend 2d ago

Some people need to feel powerful and important, and they can't get that feeling through their regular life, so they exercise power by... gatekeeping a publicly-available provider name. To each their own, I guess.

2

u/someoneyouknow23 2d ago

Yeah I doubt he even has one, would be totally unprofitable anyway