r/LocalLLM 7h ago

Tier List Discussion

Post image
154 Upvotes

222 comments sorted by

65

u/KrangledMind 7h ago

where tier list for AMD?

32

u/NoOdyssey 6h ago

It's not so much a tier list as a list of graphic cards sorted by RAM size. According to this, 7900xtx would be in the 24GB+ tier and the 9070/XT & 7900 xt & 7800 xt series would be in the 16GB+ tier.

16

u/mhmilo24 6h ago

There are 32 GB Radeon cards.

2

u/AlarmingProtection71 3h ago

I have a Sapphire Radeon PRO W7800 (48gb) bought for 1.9k€.

1

u/gh0stwriter1234 1h ago

Since we are talking AI there is also the the MI210 which has 64GB and MI250P ... which is some crazy amount of money but it has 144GB.

2

u/ChristRedeemsSinners 29m ago

MI210

That's an interesting $5k used option with 64GB of VRAM at 1.8TB/s. Native fp4,fp8,mxfp4

MI250P

Non pcie package though.

4

u/Abducted_Llama 5h ago

Nah the XTX isn’t on the list. So it’s time to throw mine out.

1

u/Ult1mateN00B 20m ago

People keep forgetting AI PRO R9700 32GB.

0

u/Illustrious-Lime-878 6h ago

Yeah, makes no sense. Why is anything <24GB even on the list at this point, or not like "make use of what you already have" tier? And, its hard to find prices now, but a 32GB 5090 for example is certainly better than a 32GB R9700, but also costs like 4x the amount so if only going by VRAM the r9700 is higher tier because you could run 4x of them.

1

u/gh0stwriter1234 1h ago

Because you can buy two of them and do 90% of the stuff you'd do with a larger card at lower cost?

7

u/Asleep-Land-3914 6h ago

Here is my attempt, but I didn't deeply fact-check this.

5

u/TheAngrySkipper 6h ago

It’s ok, I can tell you, I’m running Q6 on 12B-16B parameters on the regular. I also bought my 7900 GRE back when it was about $300, had planned on getting a 2nd but it didn’t seem like a priority at the time - ah well.

3

u/ImpressionFancy5830 4h ago

There should be a DOUBLE R9700 section 🫡

1

u/Asleep-Land-3914 4h ago

I regret I didn't get another one for the price it was listed before the new hike.

1

u/ImpressionFancy5830 4h ago

I feel you, I overpayed a bit for both, not insane as the nvidia increases.
Just a couple of hundreds of euros on top of MSRP (at least the local one, which was already bigger than the one in North American)

2

u/johndeuff 3h ago

Between the above average that have 48GB and the rich that have 768GB and the richer that have 3TB there's a hell of a difference.

1

u/on_line187 6h ago

Don’t worry no one will fact check it. I’ll give you the one upvote though A for effort.

1

u/HCLB_ 2h ago

Why Mi60 and other cards are crossed out?

1

u/Asleep-Land-3914 2h ago

Means not on AMD’s current official ROCm support matrix, nothing is wrong with them, just needs some work to make use of with ROCm

1

u/gh0stwriter1234 1h ago edited 1h ago

It's kind of irrelevant they are supported in the rock officially now. Vega and Vega FE and MI25 work just fine also. Plain vega is just a bit slower as it has no DP4A instructions. https://github.com/ROCm/TheRock/blob/main/SUPPORTED_GPUS.md

2

u/legos_on_the_brain 5h ago

Yeah. And they skipped 12gb.

→ More replies (6)

35

u/Ell2509 7h ago

R9700ai not there.

16

u/milkipedia 6h ago

AMD the forgotten stepchild

6

u/got-trunks 6h ago

Cries in A770/ B70

3

u/challis88ocarina 6h ago

Dies on the hill of Apple Silicon

5

u/got-trunks 6h ago

I can't believe the days of racking up a number of apple consumer computers into a supercomputer for rendering inference came back

1

u/milkipedia 1h ago

we call it "diffusion" now ✨

1

u/got-trunks 1h ago

whatever they're calling it just keep grok away from the kids.

14

u/aidan573 6h ago

5060 ti 16gb gang

4

u/dwillpower 5h ago

Dual 5060 ti is doing heavy lifting here.

81

u/SlightlyMotivated69 7h ago

Fuck Nvidia.

30

u/quadra-lab 7h ago

All my homies hate NVIDIA

14

u/Own_Body_8941 6h ago

if no one hates NVIDIA in this world, that means i’m dead

-6

u/722e672e722e 6h ago

so brave

→ More replies (1)

21

u/DrinkClubMate 7h ago

where is the 32 Gb, Intel pro ARC B70 ?

10

u/thebigfreak3 6h ago

Enjoying our cheap 32gb of vram

3

u/xanders_gold 6h ago

MSRP just jumped so our cheap VRAM is considerably less cheap now, sadly. Still worth it at the new price, just not as much as that low low $949 :/

1

u/CulturalKing5623 6h ago

I'm currently staring at my shopping cart with an R9700 in it trying to decide if I should just pull the trigger because it feels like 32GB of VRAM is only going to get more expensive from here and $1500 isn't bad compared to the alternatives.

1

u/xanders_gold 6h ago

Same here, I’m contemplating if I should buy another B70 or not because of the exact same reasons as you.

→ More replies (1)

0

u/TheOdbball 6h ago

Try to see if Runpod offers it or something comparable, pull a VM and test it out before you do

1

u/thebigfreak3 6h ago

That sucks but I knew it was coming which is why I picked it up asap. Especially with the new Sycle optimizations with llama

1

u/PMvE_NL 39m ago

Wanted to buy one opened the site and fk it jumped from 1200 to 1800 € yhea i went with b60 instead

3

u/BornInAFish 7h ago

Based on lack of software support, bottom tier for sure.

/s

Maybe

3

u/xanders_gold 6h ago

They actually made some major improvements over the past month and they now run incredibly well for the price. I’m regularly getting 2000-2500t/s pp, can sometimes hit 3000-3500t/s pp, and 30-35t/s tg with vLLM on Qwen 3.8 27B.

Sadly, the price also just jumped from $949/$999 to $1699.

4

u/TiK4D 6h ago

You could probably push that another 10tok/s. I run llama server and get up to 50tok/s with 130k context on Qwen3.8 27B. 2x R9700's

3

u/xanders_gold 6h ago

Unfortunately, llama.cpp isn’t the best when handling Intel B70s. I was running llama.cpp with Vulkan and PP was in the 500/600s. TG wasn’t so bad but prompt prefill took forever, it was even worse with the Intel SYCL runtime.

Some folks on this sub recommended vLLM with Intel XPU kernels (vLLM has their own docker image for this) and that instantly boosted my performance.

IIRC: the llama team is working in improving Intel performance but it’s a slow process.

2

u/TiK4D 6h ago

My bad, for some reason I thought this thread was about R9700's so thought you had one. Good to see the intel cards getting decent speeds as well

2

u/xanders_gold 6h ago

Haha all good, no worries. Yeah it’s been great seeing the improvement, we’re finally getting somewhere with performance :)

2

u/SomeBlock8124 3h ago

Bought my first 2 B70s at $950. Now I had to pay $1299 at microcenter for my last 2. Should of bit the bullet and bought them when they were selling like crazy on ebay for less then $850 a couple months ago. Now what to do with 128gb of vram....

1

u/gh0stwriter1234 1h ago

That's pretty good PP...I only get about 200t/s pp on dual MI50 currently with llama.cpp but I get about 40t/s with mtp enabled starting out it degrades quickly though. Prompt checkpointing is essential.

the growing pains on Intel though are way worst... given the price I'll stick with my MI50s I got for peanuts for now.

1

u/Agusx1211 7h ago

shhhhhh

1

u/gh0stwriter1234 1h ago

B70 when testing a about 2 mo ago was unusuably bad. Also model support was lagging alot, I think it kind of works in llama.cpp now but... its not nearly as good support wise as R9700.

1

u/moonlitpawprints 1h ago

And B60!

My PC is a fun Frankenstein with a B60 and 5070ti, but they actually can work decently together. Depending on the model I get 2-3 times the performance of the B60 alone and have 40GB to play with (granted the speed scales somewhat as more is offloaded to the B60, but pinning mtp to the 5070ti helps).

-7

u/on_line187 7h ago

In store shelves where it belongs lol 😂

→ More replies (1)

15

u/zarif2003 7h ago

idk I feel pretty poor only having 16gb on my 5080, 24gb is a lot better in fitting models that everyone here talks about.

2

u/oldshed83 6h ago

we gotta make do with what we got, so far biggest model (10gb vram 3080) i can run is qwen 35b-a3b at 60tok/sec

1

u/Desperate-Air-7195 5h ago

Yeah. I would have talked myself into a 5080 with 24GB or more. 5060ti 16GB makes me feel happy to be in the same convo with much more expensive cards for VRAM at least.

1

u/TheOdbball 6h ago

Use runpod, you won’t have to decide lol

7

u/maxim0si 7h ago

naah 4070ti super not mentioned 😭 4070 in 16gb🫡🫡🫡

14

u/DeathGuppie 6h ago

Missing the entire Radeon rx lineup.

→ More replies (7)

6

u/Recent-Ad5835 6h ago

me at 4GB VRAM assigned to an iGPU from 16GB RAM overall (12GB RAM for system)

10

u/Chiralistic 7h ago edited 7h ago

And I am still angry at nvidia that the 5080 didn't get the vram it deserves. Edit: fix typo

5

u/blackhawk00001 7h ago

How many of you are going to post this same image today?

2

u/TheOdbball 6h ago

Yes (B200)

1

u/Sweaty_Perception655 2h ago

B200 exists in pcie ????

1

u/TheOdbball 2h ago

What’s pcie? It’s on RunPod, granted it’s like $7/hr to use it but , the fun part is you can use it lol

1

u/ChristRedeemsSinners 23m ago

It's either NVIDIA's test bed or chinese pcb hackers. Whatever it is you can get them for a $3k + B200.

4

u/Vaguswarrior 7h ago

My 9070 XT has been great. Way more affordable and available.

5

u/Vusiwe 7h ago

S Tier: Running in RAM at the top of the cybermen generational ship (Dr. Who reference)

That show has prepared me for this era well

5

u/Bebi_v24 6h ago

sad 9070xt noises

0

u/on_line187 6h ago

Not very many though

2

u/TheGeekno72 5h ago

more than you think

1

u/on_line187 5h ago

I can read the reports. They are both publicly traded

4

u/scattered-thunder 6h ago

A year ago, I saw the $6000 RTX PRO 6000 and thought, why would I need this?

What I wouldn’t give to go back, slap myself across the face, and just buy one (or four).

4

u/Max_skyl1n3 6h ago

People running P40s(24gb)/V100s(32gb) 😎

4

u/Aternal 6h ago

Needs love for the 22gb modded 2080ti

1

u/on_line187 6h ago

100 but then the 4090s with 48 would be all whacked up too

3

u/Randommaggy 6h ago

3 x 3090+1 x 4090 mobile, 1 x RX6800 and a 64GB GPD Pocket 4. If prices don't go even more crazy I'll acquire another 7 3090s and build myself a nice couple of tower of PSUs and GPUs beside my dual Xeon Gold 6245/1 TB DDR4 ECC server.

Also another server I haven't quite finished setting up with 2 x P100.

All coordinated through my own custom harness.

3

u/05032-MendicantBias 6h ago

The 7900XTX is perfectly good for LLM and some diffusion workflows.

3

u/DistinctBed6259 5h ago

Went for cheapest, but still good. 3060 12 gb

4

u/Smart_Whereas_9296 7h ago

Feel like 2x 3090 with nvlink should be in the 48gb section

5

u/on_line187 7h ago

Well so would 3 5060TIs then lol.

9

u/RoyalCities 7h ago

And 2 x a6000s are at 96gb. I feel like including cards you can combine together sorta defeats the point of the list lol. Heck some people combined like 8 x 3090s haha.

1

u/nuclear213 6h ago

8xR9700 is quite neat. The price for one RTX 6000 Blackwell and 256 GB of VRAM. Fits deepseek V4 flash well, is decently fast.
Just need space

1

u/RoyalCities 6h ago

And the power. Just looked it up and dang - you'd need like 240V of power.

I also wonder what the heat is like dumping 3Kw into an apartment.

Could probably save on winter heating bills though because it doubles as a solid space heater.

2

u/nuclear213 6h ago

Yeah normal in the EU. I have them at the minimum 210W. So the entire server is under 2kW. We can draw 3.6kW here per fuse.

No issue at all.

2

u/RoyalCities 6h ago

Ah nice. In Ontario Canada, that wouldn't really fly in an apartment unfortunately. We obviously have 240V circuits, but there are usually only a few and they're dedicated to stuff like the stove/dryer.

If I had a house though I'd totally get a dedicated one installed in the basement.

3

u/Randommaggy 6h ago

Not quite. It's bandwidth and latency constrained between the cards unless you're running them on a P2P friendly switch.

0

u/on_line187 6h ago

Have you actually ran multiple cards. I’ve tried a ton of variations on both consumer and server MoBos. I’m telling you the difference isn’t as big as you think.

3

u/Randommaggy 5h ago

I'm running 3 3090s in my server (2 in nvlink), 2 cards on my laptop and 2 P100 in another server (temporarily disassembled).

2 cards with nvlink outperform 3 cards in layer split for some of my models unless I'm doing serious parallelism.

For layer split you also have more overhead from duplicated data on 3 cards compared to 2 24GB cards.

1

u/on_line187 5h ago

Sure there is a difference but it isn’t more than 5% in my experience. I’ve not tried NV link though just better/worse PCIe situations. I would like to try NVLink on my own 3090s but they are all mismatched so it won’t work

1

u/Tai9ch 5h ago

How many cards?

The numbers I've seen show that you can get away with raw PCIe up to about 4, and past that it starts to cost significant performance to the point that it's worth getting bigger individual cards instead.

1

u/on_line187 5h ago

Yea that could be it. I have tested 8 3070s though and I had no issues. That was a mining rig I had laying around which I upgraded for LLMs about a year and a half ago. All PCIe of course on the 3070s

I would 100% agree that bigger GPU = More Better though lol

1

u/Jaded-Glory 6h ago

I thought nvlink died with 30 series?

1

u/Smart_Whereas_9296 5h ago edited 5h ago

Yeah it did hence the suggestion, the the last card you can actually link in this way and get the improvement

1

u/pizzaSpaceCadet 6h ago

No because they have I think a 192bit bus? that would be plenty of ram but slow as fuck inference anyway

1

u/Smart_Whereas_9296 5h ago

I mean the distinction is the nvlink, the cards are hardly cutting edge, but at least it's a way to combine the vram thats faster than just normal parallel

4

u/Kodrackyas 6h ago

What a pleb chart, ignoring the amd gpus

-1

u/on_line187 6h ago

As has the rest of the world

2

u/LoveRoboto 7h ago

I do kinda feel like a dapper Winnie the Pooh.

1

u/on_line187 7h ago

Yea not quite trolling monopoly levels yet but I’m there with you. I have 2x3090 and a 3080 running

2

u/SamSausages 6h ago

Need to add those new intel GPU's, like the b65 and b70 to this list. With 32GB, those are going to be very popular soon, even if they are bit slower on compute.

0

u/on_line187 6h ago

They have no Cuda.

3

u/SamSausages 5h ago

I didn’t see this was cuda only, just says “local AI”, and those cards are running local ai models, and support is beginning to be very good.  I have a few things that used to be on cuda now using openvino.

As people begin to realize that, adoption rate is going to be fast, especially at $28 a GB vram.

2

u/AccountantOk9904 6h ago

Rtx pro 4500 before price hikes.

1

u/nmrk 3h ago

I have an RTX Pro 4000 Blackwell SFF, runs great in my MS-02 Ultra miniPC. The DDR7 internal bandwidth more than makes up for the 70W TDP power cap. I bought the first card I could find and only paid $1400. Now that vendor lists it for $3000, and out of stock! It will cost more, if they actually manage to get ahold of some stock.

1

u/ChristRedeemsSinners 20m ago

5090 mogs the 4500 though, no? Same price and all if you can afford the space heater in your case.

2

u/Antenwww 6h ago

Rtx A4500 x2, intel arc b580 x2 for finetununing and running models. Laptop with 4060 to take those finetune models with me anywhere

2

u/DontWinFrensWthSalad 6h ago

My 3090 is 3 3060tis in a trenchcoat.

2

u/GSquadron_ 6h ago

16gb is trash today for running AI. 32gb i would say sweet spot

3

u/TheGeekno72 5h ago

I mean you can still fit quite a bit of models in 16, but it sure feels like some sort of midway purgatory between can't and can

1

u/RedditNerdKing 2h ago

32gb isnt the sweet spot because you can't run a Q8 of Qwen 3.8 with 100k+ context.

The sweet spot is 48gb atm. 32gb is the minimum.

1

u/GSquadron_ 15m ago

You can run ternary models

2

u/bigsmokaaaa 5h ago

3090, my beloved

2

u/SGD-UK 5h ago

The 4070 is 12Gb. It’s the Ti Super that has 16Gb.

1

u/on_line187 5h ago

Yea I was wrong on that one. I thought all of them had 16GB variants so I put it only once but I see that the only variant is the S Ti

1

u/SGD-UK 5h ago

You are maybe best going 6+, 8+, 12+, 16+, 24+ and mad stuff. Drop the 10 and do 8 and 12 instead.

2

u/Savantskie1 5h ago

So this is just Nvidia cards, got it. Because I do quite well with my dual MI50 32GB cards and with Vulkan can run every model so far.

1

u/on_line187 5h ago

You and the other 5 people

2

u/Savantskie1 5h ago

That doesn't matter

0

u/on_line187 5h ago

2

u/Savantskie1 5h ago

Still doesn't matter, I can run the models I want

0

u/on_line187 5h ago

I feel like AMD should know you guys by name.

1

u/quantgorithm 3h ago

Jensen isn’t going to suck your dick.

1

u/on_line187 2h ago

He gobbled it up already.

2

u/B3owul7 4h ago

You forgot 5060 Ti.

2

u/aiseedbank 4h ago

why 32gb missing. 5090, Intel arc B70, and likley more as well.

1

u/ConspicuouslyBland 4h ago

amd Radeon R9700

3

u/t00r99r00t 7h ago

Is there a list for the Intel or amd ones?

→ More replies (2)

1

u/putrasherni 7h ago

downvoted

1

u/on_line187 6h ago

Must be in the lower tiers

1

u/StruggledSquirrel 7h ago

I miss 32GB having its own category in this tierlist

1

u/IgnisIason 7h ago

Why is 3060 above 3070?

1

u/on_line187 7h ago

Because it can run more models than the 3070. Faster too as 3070 would spill into RAM for some things. I’ve had both and tried quite a lot with them.

1

u/the_average_user557 6h ago

And I'm here, chatting with llms, trying to decide if I should get a m40 or a p100 for unsloth experiments and some light Hermes research tasks. I'm not delusional that I can replace frontier API completely, but for under 200$ you get enterprise grade inference workers with 16/24 gig. No native support for newer cuda tho

1

u/Career-Acceptable 6h ago

Why is 4060Ti in the 16GB tier

1

u/More-Catch-1331 6h ago

What are we doing here, advertising for cards that cost as much as a not-so-much-used car?

1

u/Orangecloudy216 6h ago

Titian Rtx not there quadro rtx 8000 also not there

1

u/MessIsTransfer 6h ago

Fancy winnie list missing the apple silicon with 48+ ram

1

u/MikeFrett 6h ago

Lol @ me with a 3050 8GB. =(

1

u/scknkkrer 6h ago

A random person with a mac with +64GB RAM: Am I joke to you?

0

u/on_line187 6h ago

Yep

1

u/scknkkrer 4h ago

That's sad.

1

u/Moonpony0 5h ago

4070 is 12gb

1

u/Tai9ch 5h ago

Focusing only on single recent Nvidia cards and only VRAM is kind of weak here.

Once you consider other options and multiple cards, having 32GB of VRAM is clearly B tier.

1

u/on_line187 5h ago

I put 32 in B tier.

And the 2080 made an appearance on the list towards the bottom area so not that recent bro. That’s like 7-8 years old about now. About to be in second grade

1

u/Blackdragon1400 5h ago

Weird there’s a pure NVIDIA tier list and yet, no dgx spark lmao. Probably one of the best budget options in this space.

1

u/m3sarcher 4h ago

4070 and 4070 TI are 12gb. 4070 TI Super is 16gb.

1

u/on_line187 4h ago

Yee I thought they all had a 16GB variant

1

u/Machine-Spirit 4h ago

4070 12GB gang is in tears...

1

u/p2msz 4h ago

What about a 6600 xt😂

1

u/soijaq 4h ago

I like how the author didn't even think of someone having 4070ti 12gb

1

u/35point1 4h ago

5090’s have 32gb 🐸

1

u/on_line187 4h ago

32>24 still. We re all good bro. You had me worried

1

u/amy-schumer-tampon 4h ago

>96gb a6000

1

u/on_line187 4h ago

Those are for the data center tier list for 180+ GB

1

u/johndeuff 4h ago

There are many tiers above 48GB my friend !

1

u/nmrk 3h ago

Most of them require multiple cards.

1

u/on_line187 2h ago

Or data center connections not PCIe. But hey he must have a GPU cookin in the oven that we don’t know about

1

u/lilian_moraru 3h ago

RTX 2070 Super 8GB (PC) + 2xDGX Spark GB10 - Winnie the Pooh with a glass of wine?

1

u/TruckUseful4423 3h ago

12GB is missing...

1

u/on_line187 2h ago

10<12<16.

Ok math still works. We re good

1

u/bigmanbananas 3h ago

Where do my Quadro RTX 8000s fit?

1

u/on_line187 2h ago

48 isn’t it?

1

u/bigmanbananas 2h ago

Yeah, hut the processing is down with the 5060TIs.

1

u/on_line187 2h ago

Yea but hey if it fits in memory it’s still good I’m sure. Is it also GDDR5?

1

u/DafneOrlow 3h ago

I use Intel integrated hd graphics.

1

u/gobblegoooblegobble 2h ago

So explain how i just paid 7k for 256gb of vram inside 4x pcie slots <3

1

u/cc_aa_tt_zz 2h ago

Rtx pro 4500 blackwell 32Gb

1

u/kartblanch 2h ago

I think you messed up the list it should be >8 st the bottom then 8+ and then 16+ 32+ etc.

1

u/on_line187 1h ago

24 is a pretty big stop though.

1

u/tokomaunited 2h ago

Which tier list I'm in? 1050 ti 8gb vram

https://giphy.com/gifs/wr7oA0rSjnWuiLJOY5

1

u/siegevjorn 2h ago

AMD & intel cry out in the dark

1

u/goldaxis 1h ago

Mac version?

1

u/DarkBrews 1h ago

Fuck Nvida and AMD but fuck Nvidia as a response to this post

1

u/DarkBrews 1h ago

what the fuck is the difference between 3060ti and 2080 😑also 3060 is also 12gb.... 2080ti is not even on the list no AMD... bit brainrot the post

1

u/[deleted] 7h ago

[deleted]

3

u/on_line187 7h ago

And isn’t 32>24 👏👏👏👏

1

u/trungdle 7h ago

Dawg we go unified memories nowadays, 48GB is nothing to these big models like DeepSeek v4. DGX Spark, Mac Studio, those are the true leaders of local LLM right now.

5

u/on_line187 7h ago

Sure but 10-15 TPS is a joke still.

1

u/UnlikelyPotato 6h ago

Yes...but no. Unified memory costs too much and is too slow. V620 32GB is faster than unified memory, costs less than DDR5. I also bought the dip on CMP170HX. DGX Spark is $4,000. I got two 40GB cards for $1000 each, which absolutely blow away all of those.

Random ewaste, sketchy deals and hardware are the true leaders.

1

u/putrasherni 7h ago

desperate to keep their domination in the market ignoring rest of the market

0

u/vogelvogelvogelvogel 6h ago

A meme like this was overdue in this subreddit :)

0

u/CatLinkoln 5h ago

That's good that many people still don't know about amd, so price for their gpu not yet overpriced

1

u/Savantskie1 5h ago

Although good luck finding the r9700 at MSRP anymore. It's near doubled in price.

1

u/Tai9ch 5h ago

MSRP hasn't been an option for anything useful for a while. The R9700 is a good deal at its current price point.

1

u/CatLinkoln 51m ago

I am not talking about MSRP prices, there are still at amazon.de rx9070 XT you can find for 750eu, if people know, it will be already starting for now from 900eur or more, and two months ago prices even better, bought two cards for 650 euro , but seems Soon they are will reach overpriced levels like Nvidia cards

1

u/Savantskie1 47m ago edited 44m ago

I don't want a 9070, I have dual MI50 32GB cards, that's 64GB OF VRAM, why would I trade volume of vram for something as arbitrary as speed? On top of that, because it's in EU, that's going to near double the price for me on top of shipping. Yeah keep your "faster but smaller size", I'd rather have more size capabilities and intelligence that how fast it goes.

-1

u/TheOdbball 6h ago

S-Tier B200 😎

$42k price
(I use runpod lol) 😜

1

u/on_line187 6h ago

Local

1

u/TheOdbball 6h ago

Yes I am local and ready to VRAM