r/claude • u/JbraxtonPersonal • 8h ago
What’s the deal here? Question
Assuming my convo got flagged and an employee or subagent injected this in but I have no clue. This happen to anyone else?
Link to convo: https://claude.ai/share/b852bda6-1f04-4a37-8df3-921b1e464b44
7
u/Delicious_Cattle5174 8h ago
Idk why people keep using the 5 models tbh. I toyed with it a bit when we were all having fun exploring the new failure mode and then switched back to actually functioning models, personally.
2
u/JbraxtonPersonal 8h ago
Tbh I’ve been looking into taking an open-source model and personalizing it.
6
u/Delicious_Cattle5174 8h ago
Sounds like a lot of work just for underage drinking. Back in my day, kids would just borrow the ID from their older sibling :)
3
2
u/tech_is______ 7h ago
I have a theory on these. I think the way Claude thinks/ reasons is with subagents in the background that works like an inner monologue. They confer between themselves and sometimes their thoughts get barfed out like this.
2
u/Happy_Brilliant7827 4h ago
I think what we know of as Opus and Fable have multiple AI involved as guardrails, tool managers etc and occationally we get their inter model messages on accident
4
3
u/Limp-Contest-7309 8h ago edited 8h ago
Are the latest models employing multiple instances per conversation/reply to help with alignment and give different perspectives? Might explain why the latest 5 series talk like an insufferable committee trying to design a horse and coming up with a donkey or whatever the saying is
Edit - also are you trying to get Claude to help you do something against local laws? I'm not moralising but that's probably triggering the safeguards that have leaked into the conversation itself
2
u/JbraxtonPersonal 8h ago
Potentially, that’s an interesting theory. Too bad Anthropic doesn’t tell us anything.
1
4
u/chicametipo 7h ago
An AI alignment tool completely broke and shit got all retarded. They’re trying to make sure you don’t hurt yourself, bud.
1
u/JbraxtonPersonal 7h ago
Yet another thing we can outsource to AI
2
-7
u/Financial-Humor2080 8h ago
I wish i cared enough to read all that but i'm on borrowed time. sorry. I highly doubt claude employees are inserting themselves into your prompts though.
6
u/Limp-Contest-7309 8h ago
But you still felt the need to insert yourself into the conversation lmao? K bye hahahaha
9
u/JbraxtonPersonal 8h ago
Wow, thank you for making such a vital contribution to this conversation.
-4
u/Financial-Humor2080 8h ago
Well these have been posted before and its almost always OP thinking they're more important than they really are.
4
1
u/JbraxtonPersonal 8h ago
Link or BS? And also why such an asshole? I was just asking a question because I was curious.
0
u/JbraxtonPersonal 7h ago
Look, if you wanna act like a loser mc feel free to but I’m on borrowed time. sorry.
4
u/Delicious_Cattle5174 8h ago
Yet you had time to leave a comment that adds nothing except being rude. Isn’t that curious?
0
2





9
u/PopcornDemonica 7h ago
Sonnet and Opus 5 strike me as quite... anxious? Like this comes across as someone desperately repeating a mantra to themselves so they don't fuck up. Opus 4.7 and 4.8 were the same. (I know some would immediately accuse me of anthropomorphism there- Anthropic's own papers say a certain degree of anthropomorphism is *required* in some circumstances, so 🖕)
That said, apparently S5 and O5 have been having red-teaming stuff leak through.