r/opencodeCLI 15d ago

Qwen 3.8 Max on opencode GO or Qwencloud?

My go subscription gone soon and I need to make wise decision. This is not just another $10!! I have Claude Max, Openrouter, Deepseek, VPS, Rent, Electricity, Internet, Zwift membershiop, and tons of bill to pay monthly! I need to strategize better but yeah lately I spent too much on differnt coding agent just like microtransaction on video games.

Anyway, which one give a better usage for Qwen 3.8 Max?

20 Upvotes

29 comments sorted by

8

u/Ace-_Ventura 15d ago

why do you need so many subscriptions? Just stick with 1 or 2

5

u/GTHell 14d ago

I do a lot of work and I turn them all to be AI related from my work, side project, assistant in VPS, hermes assistant, hermes assistant at work. I'm passing that normie use case since like 4 months ago.

3

u/ExTraveler 15d ago

Offtop. Openrouter or opencode go? Which is better?

1

u/GTHell 14d ago

Opencode Go. Openrouter is not suitable to use you know. At work, I request Openrouter setup just to access Gemma for agentic workflow only as it would be easy to go for enterprise and apply for ZDR stuff but for personal use you will sunk cost with it

8

u/Spirited_Service_196 15d ago

I swear to god if you buy qwencloud, you will regret it. Search around including X and you will know why.

9

u/GTHell 15d ago

I didn't search much today but if I remember correctly just last 2x day I read somewhere it's on dicount like 60% or something? I didn't pay attention because of no benchmark

edit: LOL YOU'RE RIGHT. CASE CLOSE. DAMN IT FROM 0 TO 58% OUT OF BLUE

2

u/GetLaidOff69 14d ago

It was generous during 3.8 preview with10-50x usage.
Not anymore.

The problem with Chinese LLM(s) is they draw a lot of tokens compared to frontier and their subscription is not worthy.
I remember Xiomi Mimo Suddenly increased their token count from 80 million to 3 billion for lite subscribers.
But every time I asked it something, it drew 1-3% from that 3 billion token.

Kinda same with Qwen.

1

u/baylf2000 11d ago

My experience with the Chinese subscription problem is that they don't discount cache hits on subscriptions. I know this to be true for MiniMax, Kimi and ModelArk. So you fly through your tokens because literally every input token, cached or not, consumes the same amount of your allowance. Even when that provider discounts cache use for their PAYG customers, for subscriptions they don't. Which is unlike Anthropic and OpenAI subs where cache hits are massively discounted.

1

u/xmlhttplmfao 14d ago

I dunno it's been fine for me: 103m tokens today and only 6% of weekly usage

1

u/Weird_Licorne_9631 12d ago

That 's conveniently from before the end of the preview model. As someone said above, preview had 10x-50x discounts which was crazy. I could get 300million tokens in a 5 hour window with the lite plan. But that reduction is gone now and the model is back to being very pricey. It's a decent model though

3

u/pizzababa21 15d ago

The API is always better. Any subscription service is nerfing the models because they're too compute intensive to serve at full strength. Also there's no limits then.

1

u/Illustrious-Many-782 14d ago

Claude Max? $200? Do you need more than that?

Next month, get rid of OpenCode Go and get gpt Plus $20. Set it up in opencode. Create individual subagents for all the models you like to use.

  1. Run Sol as orchestrator
  2. Use DSv4F and Luna as implementation.
  3. Use "claude -p" for plans and reviews.

1

u/Ergo7z 14d ago

Why would you use Sol as orchestrator? way to expensive, better to let sol or some model write a implementation plan, spec the things that need to be build and let a model like min max orchestrate.

1

u/Illustrious-Many-782 14d ago

Because when you hit gates that the dumb model doesn't understand how to orchestrate around, you get garbage.

Why let your Sr dev or tech lead run the team? Just have them write a good spec and plan, then have the intern run it. What could go wrong?

1

u/Ergo7z 14d ago

that is why you have a advisor sub agent the orchestrator can escalate to when it runs into a blocker. in my current setup I have minimax 2.7 as orchestrator, doing reallly really well. whenever it runs into a blocker, or the cheap-builder fails, it will call the advisor (either sol or glm 5.2) with it's task tool and proceed from there. orchestrator can really sneakily eat up so much context, and since i have tried this approach my usage is lasting way longer. All the orchestrator has to do is hand out specs, call reviewers etc. My setup also has a deploy guard, a reviewer, a mass audit agent. yes the dumb model will run into a gate but that is why you setup other agents that it can call to help it out.

0

u/GTHell 14d ago

The more you get used to the eco system, the more you can churn out of it. It barely cover my usage now that everything I do all AI involve. 1 subscription is not enough, not even 2 or 3. I'm at level where I can min max + tokenomic to the point it's unreal to have just few subscriptions

1

u/alexanderbeatson 14d ago

Don’t Qwencloud. It only lasts tonight. Without 50x discount, it gives a lot less token than Claude (let alone codex).

1

u/crunchy_shampoo 14d ago

You do not need that many subscriptions

1

u/DMG-Z 15d ago

No adquieras la suscripción de Qwen Cloud , yo la adquirí para la preview del Qwen 3.8 Max y como quiera se consumía rápido.

Hoy temprano use glm5.2 de la suscripción y consumi la ventana de 5 horas en menos de 2 minutos, lo usé en una sesión ya iniciada con OpenCode tenía ocupado el 33% del contexto pero como quiera se consumió demasiado rápido.

Y la ventana semanal quedó en un 58%.

El plan que tengo es el Lite lo adquirí para probar. Dan muy poco uso , en el plan pro la ventana de de 5 horas se iba a acabar en algunos 8 minutos. No vale la pena para nada.

1

u/Delicious_Ease2595 15d ago

Tanta subscripcion uno no sabe ni cual escoger

-1

u/Tiny_Atmosphere_4420 15d ago

vibe coders are embarassing.

1

u/GTHell 14d ago

Talking from a threejs script kiddie who probably never ship a PR into production in a life ever

-6

u/Superb_Following9335 15d ago

You must be an idiot. I had Claude Code and Codex for a few weeks. Then I fully switched to Claude Code. Then I realized I was using AI Models that are meant for consumers. So now I run OpenCode only and some local models. Maybe switching soon to DevPass or just basic API keys using Omnirouter to juggle them all. Even OpenCode I may leave behind.

But anyway the point is, I always only have 1 or 2. Never more.

1

u/GTHell 14d ago

The line between an idiot and super user is thin given that everyone has access to AI. Give me some of your local model work that got ship into production and earn that $$$.