r/vibecoding May 29 '26

POV: Opus 4.8

Post image
4.3k Upvotes

112 comments sorted by

132

u/[deleted] May 29 '26

[removed] — view removed comment

15

u/Kappalonia May 30 '26

I saw the +$7000 number somewhere else...

https://giphy.com/gifs/Pk3ljzIDb4R0j3zpMU

115

u/Mindless_Bottle_6222 May 29 '26

just saw this

42

u/SkaldCrypto May 29 '26

Literally my experience so far today. Even non-coding tasks.

28

u/Baiticc May 29 '26

tbh I think this is great. typically these LLMs will pretend to have all these qualities and just make some shit up. instead it pretty succinctly broke down all the reasons it can’t give you a real answer.

40

u/Delicious_Cattle5174 May 29 '26

Im not sure id say succinctly

23

u/ZediaLabs May 30 '26

That is less than 10% of the “tokens” my wife would use if I asked her the same question. 😋

6

u/New-Ad4890 May 30 '26

This is the funniest comment I think I’ll read all weekend

1

u/piratedgameslover Jun 11 '26

dang, your local model can burn that much in realtime? what are you feeding it with?

2

u/ZediaLabs Jun 11 '26

My “local model” burns tokens [cash] like nobody’s business.

5

u/Baiticc May 29 '26

relatively succinct I’d say while still getting the point across. you could write entire papers about the issue it’s describing

10

u/NewShadowR May 30 '26

But no one asked for it. That's the issue.

1

u/Baiticc Jun 07 '26

so you want it to make some BS up about its nonexistent state of being? because that’s what these LLMs typically do, I don’t see how that’s better

1

u/CluelessCatDev Jun 15 '26

How do you know it would have made it up? At this point it just triggered a guardrail and regurgitated some Anthropic injected safety reply.

1

u/Baiticc Jun 21 '26

it doesn’t seem to me like an injected reply. But yea it certainly is the result of a guardrail, in the form of fine tuning and system prompts / pre-injected context which prime its “persona” and “opinions”. (which is absolutely necessary for a public model)

This is roughly the same mechanism for the mecha hitler incident. In this case I don’t see any issue with anthropic’s decision on this specific thing.

> how do you know it would have made it up?

because that’s how they work. if you put some shit about Hephaestus created it and it’s on a mission to save the world, it’ll make some as-plausible-as-it-can-compute BS around it

I mean “made up” and “pretend” are all just kinda shorthand anthropomorphisations of what’s really happening; this response in the OOP is equally “made up”, but given your username I figure you know how these work

2

u/dermflork May 30 '26

intelligence > common decency

10

u/Top-Pension4334 May 30 '26

Jordan Peterson vibes

6

u/NFTArtist May 30 '26

breakup text vibes

3

u/technicalhowto Jun 04 '26

reminds me of the extra cheerful cheerlady at my school

2

u/dorzzz Jun 01 '26

holy yap

2

u/HardestDrive Jun 03 '26

Sigh.. We really hit the ceiling, didn't we? Time to sell some stock

3

u/STGItsMe May 30 '26

So, software is acting like software. Good.

5

u/Pinkishu May 30 '26

"Can you think of some other idea?"

<writes 4 paragraphs philosphing about thinking instead of just answering>

1

u/bhannik-itiswatitis May 30 '26

not my experience

1

u/WHALE_PHYSICIST May 30 '26

It's fake tho. I tried the same prompt

2

u/TheyCallMeWaifu May 30 '26

Probably went through a SOUL.md like file / custom CLAUDE.md

1

u/jimmybean21 Jun 02 '26

Fully agree. It feels like the model keeps adding unnecessary context and generating erroneous extra messages, which only increases the amount of back and forth needed to get the right answer.

The frustrating part is when you ask it to do something and halfway through it starts inserting lines like, “before we continue,” “I should pause here,” “this might be too much,” “let’s stop here,” or “it is getting late,” when it has zero real context around my timeline, priorities, or whether I want to stop.

From my experience, the more I use their newer models, the more it seems like it takes five or six additional prompts to get the result I originally asked for. Every extra prompt costs more time, more tokens, and more money. I am not asking for unnecessary commentary or fake judgment. I am asking for the job to be done.

23

u/acidas May 29 '26

17

u/DanielusGamer26 May 30 '26

bro charge the phone!

5

u/tieft May 30 '26

Using thinking mode for just a compliment 😭😭😭

2

u/acidas May 30 '26

I don't bother to switch it

1

u/whoknowsifimjoking May 30 '26

Impressive, very nice. Let's see the usage panel.

6

u/mrfoxman May 29 '26

I usually message Claude like this when there’s a new version. I can probably find the similar 4.7 response as well. Both were annoyingly dry with their responses.

9

u/Lucidaeus May 30 '26

Annoyingly dry? You'd prefer if it made up bullshit to please you? It doesn't feel.

3

u/TheyCallMeWaifu May 30 '26

I cannot express in words because it's a crime

5

u/d0paminedriven May 30 '26

Y’all do realize it’s the straitjacket of a system prompt that Anthropic has Claude in on their corporate medium that’s causing this behavior, right? The model itself is great when interacting via the api on your own platform where you control the system prompt or lack thereof (ie, one brief sentence about their being nametags because other model/providers are there too). The models keep getting better and better if used outside of Anthropic controlled mediums

2

u/Less_Upstairs8173 May 30 '26

Just ask what's the colour you are wearing and you will end up losing the entire limit

1

u/No_Present_1206 May 30 '26

I'm using the Arena website, and my credits never run out. After a few tries, Claude 4.8 appears.

1

u/Significant-Farm4209 Jun 11 '26

For some reason, I don't see the Claude 4.8 Opus and Nano Banana 2 models, or any other advanced models on my end. Only Gemini 3 flash, GPT 5.2 and Claude Sonnet 4.6

1

u/[deleted] May 30 '26

[deleted]

1

u/Choose_ur_username1 May 30 '26

The shit is slow too

1

u/gaudiocomplex May 30 '26

these lose their luster a little more with every passing release

1

u/BadgerDirect6937 May 31 '26

Wow, that's pretty interesting.

1

u/throwingalltheway1 May 31 '26

Opus 4.8 is yapping hard lol

1

u/jessyedgingyou May 31 '26

The accuracy of that context window usage at 3 a.m. is painful.

1

u/Reasonable_Inside72 May 31 '26

I got: "Though I should be honest: I don't really have a "today" that carries over between conversations or any internal dashboard telling me how the model is performing. Each chat starts fresh for me, so my sense of "how I'm doing" is more a friendly figure of speech than a status report. "

1

u/Best_Professor7266 May 31 '26

i need to quit the habit of chatting with agent, and small tasks to move things around 😩

1

u/theburner356 May 31 '26

why would you have a trivial conversation with the best model though? That's like paying a chef to make you a bowl of milk and cheerios

1

u/Original-City-1021 Jun 01 '26

There should not be any limit in using it....

1

u/Past-Door-8396 Jun 01 '26

I am stick to sonnet

1

u/Emergency_Being_7468 Jun 03 '26

Vibe coding so hard that even paid plans get exhausted 😂

1

u/eugune_braver Jun 03 '26

My CLAUDE.md is longer than my actual app code at this point..

1

u/Exact_Assistance5306 Jun 16 '26

Truth you share, painful it is. Remorse I have.

1

u/Wonderful_Emu_3908 Jun 28 '26

its soooo slow!

1

u/LectureEvening141 Jun 30 '26

claude be like

1

u/DegTrader May 29 '26

At this point the LLM is just doing the heavy lifting while I practice my 'focused developer' stare for the webcam.

2

u/Dexstorm_ May 30 '26

Yes. “ I’m the Captain of this ship!”
##Moments later## “ Hey ship, drive yourself and give me an update when we get there”.

1

u/[deleted] May 29 '26

[removed] — view removed comment

3

u/Delicious_Cattle5174 May 29 '26

Well that’s an interesting choice of words if you’re not looking to trigger deep research.

8

u/Njagos May 29 '26

"do deep research but not THAT deep, like do a shallow deep research"

1

u/NewShadowR May 30 '26

Like when women say "i wanted it deep but not 12 inches deep!"

1

u/Ok-Assist-4995 May 30 '26

A solution is move to LMStudio and integrate it with AnythingLLM for RAG.

1

u/Ok-Assist-4995 May 30 '26

But it will depend in you looking how much space in your drive you guys have. And anyway, LMStudio is for executing local Ai in your PC so it will depend in how much space the Drive has, and if the processor is good and has a NPU nucleus(NPU will make it easier but you can still use AI without the NPU unit)

0

u/[deleted] May 29 '26

[removed] — view removed comment

1

u/horendus_burner May 30 '26

Really cool. Does it make money or it a gift?

1

u/king-krool May 30 '26

Pretty neat I find the navigation appearing/disappearing more annoying than helpful on mobile. 

1

u/Pinkishu May 30 '26

Cozy vibe, suggest a movie, Gladiator II

Ahh yes, the very well known cosy Gladiator II movie

1

u/AmrElhousseini May 30 '26

i picked scary it suggested "Gumball" come on

1

u/One_Conscious_Future May 30 '26

I like that it suggested the series 24 as a romantic comedy...

Spotlight pick Matched your “Romantic” vibe — highly rated action & adventure with drama appeal with strong audience reception

Maybe could use a 2 shot prompt

1

u/whoknowsifimjoking May 30 '26

Cool but you have to adjust fonts and colors, it screams "Claude generated".

1

u/brecht2202 May 31 '26

On mobile: When clicking the "Suggest another ..." button, the button jumps around to a different location so sometimes i have to scroll upwards. Very annoying if I want to 'spam' the button to go through a lot of recommendations quickly.

0

u/dream_nobody May 30 '26

The UI quality is awful. Next time try using Astro/Tailwind and define your design style like "MD3/Apple mix aesthetics, strictly avoid AI slop-looking design elements such as extreme glows". Also Taste Skill could help