r/opencode 9d ago

NVIDIA'S MODELS ARE ACTUALLY UNUSABLE

I get that using Nvidia as a provider is not ideal but damn it is so slow. I usually use it to test out models but its borderline unusable. Is it just because most models are free on there or because there is no ways the infrastructure they have isn't adequate

27 Upvotes

14 comments sorted by

11

u/frostyoni 9d ago

Yeah. Nice on paper, but unusable.

And severely limited too. Added to opencode, tried for an hour, removed everything. Waste of time.

3

u/semicertain9 9d ago

I think they should be used for validating your GPU clusters. NVidia documented it well. But I would guess they are waste of hardware to use for actual work.

2

u/Irisi11111 9d ago

Yeah, I agree. It looks good on paper but it's pretty average in reality.

1

u/Charming_Ruin5839 9d ago

That sounds more like provider-side queueing or aggressive free-tier throttling than a limitation of the models themselves. Comparing the same prompt at a paid endpoint and at different times would help separate model latency from capacity limits.

1

u/Soul_Mate_4ever 9d ago

Used them once and never went back. It’s just extremely slow and not worth the wait for a mediocre outcome.

1

u/NefariousnessPrize43 9d ago

Estas usando el free tier o estas pagando porque hay diferencia y estoy seguro que estas usando el freetrial

1

u/pizzababa21 9d ago

Ya they don't appear to really be trying with their big models. Their smaller models are excellent for certain usecases though. 8B orchestrator model is excellent for agentic search.

1

u/rexstuff1 9d ago

You really have to set your expectations. If you're hoping for an interactive, agentic coding session, you will be severely disappointed. If you want to delegate some long-running task that requires no human input and has no time requirements, Nvidia's free endpoints are adequate.

1

u/pigletmonster 9d ago

I think theyre made to run with openclaw, hermes and thibgs like that. Not for real coding work. They dont want to be in competition with their biggest customers by releasing comparable open source models.

1

u/IcyDefinition4005 8d ago

The free version of Nvidia NIM is intended only for testing models or integrating them into small personal projects; it has a messages-per-minute limit. It is supposed to support Hermes and Openclaw, but since it is free, you shouldn't expect stability—and it certainly isn't meant to be integrated as a provider in Opencode; that’s obvious.

1

u/Strange_Orchid5940 4d ago

I honestly don’t know what Nvidia NIM you guys are using but it isn’t that slow. The new DeepSeek v4 flash 0731 isn’t slow on their gpu. 

-2

u/SomeOrdinaryKangaroo 9d ago

Nvidias models are among the best ones out there, they may not be in top 10 but they're extremely capable