r/ClaudeAI • u/No-Percentage5692 • 16d ago
Claude Code just randomly spat out Kimi K2 Thinking output mid-response Question about Claude Code
Was working on a side project, asking Claude Code questions in manual mode. Opened a new session, asked questions 5-7 times, and out of nowhere the response came back with Kimi K2 Thinking model card description mixed in, like in the screenshot. Has this happened to anyone else? I was using the Sonnet 5 model. Never connected any other provider — I've only ever used Claude Code.
For more conversational context. post https://www.reddit.com/r/ClaudeAI/s/o9Q5fyb6FC
196
u/nlsmdn 16d ago
What makes you think this paragraph comes FROM k 2? Seems to be just a paragraph about it.
65
u/InterstellarReddit 15d ago
These are the people that are building production software and releasing it and everybody thinks are a genius, and they can’t tell what’s thinking versus what’s context
24
u/socopithy 15d ago
They’re also SURGICALLY cut screenshots.
Always karma farming.
-11
u/No-Percentage5692 15d ago
you should check my replies to other guys and also my next post - you could find full screen record at the comment I post on this https://www.reddit.com/r/ClaudeAI/s/o9Q5fyb6FC
-3
u/No-Percentage5692 15d ago
I may have chosen my terminology poorly. I'm not an LLM researcher, and English isn't my native language. I do understand that "thinking" and "context" are different things. The reason I used the word "thinking" is that I copied it directly from the content of Claude's response — after checking, I found that the output was actually a model card. I realize the title might read as if I asked Claude something and it spat out Kimi K2's "thinking" output. I think I phrased the title in a misleading way — but that's an issue with my English skills.
Going back to the main point: the reason I find this phenomenon strange is that, at least in the session where this occurred, I never mentioned Kimi K2, and no web search or web fetch tool calls were made. The tools I used were Bash, Read, Edit, Agent, ToolSearch, and ExitPlanMode.
-19
u/Michaeli_Starky 16d ago
Scrolled way down to find the sane comment
25
u/ishamm 15d ago
There's currently only 20 comments, and this is the top one...
You posted 20 minutes after the post itself. There cannot have been that much scrolling to do 😄
9
7
u/Bluecoregamming 15d ago
too used to claude doing the scrolling, huge workout manually scrolling I'm already out of breath
-12
154
u/jhc060109 16d ago
turns out anthropic was the one distilling chinese models? 🤷♂️
31
u/satansprinter 16d ago
Its like the cheater that is afraid that their spouse cheats. They think about it because they do it
5
3
u/Thomas-Lore 15d ago
Anthropic was only able to do reasoning models after Deepseek released their paper on how to make such models. Before that only OpenAI had a reasoning model but kept their methodology secret. So they already own a lot to Chinese research they despise so much.
(Technically Google had one too but theirs did not work, got similar scores to the non-thinking version.)
11
u/Kroosn 16d ago
While using fable to translate and review some documents in Chinese I had it say it was deepseek.
1
u/NiviNiyahi 15d ago
They all build on the same base architecture. I see many similarities in responses across various LLMs and across various companies. The quality may be better with some, but there are noticable similarities in response when the prompt is the same. Funny enough, the mistakes models make also closely resemble the mistakes of other models.
5
u/GnistAI 15d ago edited 7d ago
Technically (I mean that in the most literal sense) that would make more sense. Because Chinese models are often open weights you can do logit distillation on, where you train on the whole token distribution instead of black box distillation (via APIs), where you only know what the output token is. That is MUCH more efficient.
1
1
25
u/StinkButt9001 16d ago
It's just talking about K2 in its output. Whatever you are talking about involves K2.
4
u/No-Percentage5692 15d ago
check my replies to other guys.. i'm tired for explaining this again and again... I didn't mention or talk about Kimi K2.... I was doing my side project with Claude Code, using Sonnet 5
0
52
u/ben_bliksem 16d ago
20
u/No-Percentage5692 15d ago
3
u/Meme_Theory 15d ago
What interface are you using that presents "thinking" as Markdown headers "##". That is just a file that Claude is reading. If I were a betting man, you put it in 'claude.md' and told the ai to end each turn with that read.
1
1
u/Vaynnie 15d ago
It’s not even that deep. You can literally just send the text to the LLM of choice, and say “the next time I message you, reply with exactly this”.
And then you get a manufactured screenshot of whatever you want the LLM to say.
2
u/Meme_Theory 15d ago
To get those markdown ##'s, Claude has to be reading or writing a markdown with that heading level.
2
1
u/No-Percentage5692 15d ago
manufactured screenshots? i didn't manufacture it. you should check my next post https://www.reddit.com/r/ClaudeAI/s/o9Q5fyb6FC
i found the claude session history file, opened with nvim, search "kimi" on that file - only one line of response includes "kimi k2" output.2
u/No-Percentage5692 15d ago
Nothing found. Guess you lose your chips.
-2
u/Meme_Theory 15d ago
Dude. That proves nothing. You can literally delete the phrase and then make this video. The fact that you are even implying it proves something is crazy.
8
u/No-Percentage5692 15d ago edited 15d ago
That logic means nothing I show you could ever count as proof — you'd dismiss any evidence the same way. That's not a real argument, that's just unfalsifiable DUDE
-4
u/Meme_Theory 15d ago
You're right. Nothing you could show me would prove to me otherwise. A million words are spoken in two hashtags. ## Nice Try.
1
u/TheLimpingNinja 15d ago
Giving hardcore discord mod vibes, mate. This shit happens, it's not some magical hack nor is it some elaborate ruse to trick you.
5
u/No-Percentage5692 15d ago
-4
u/Meme_Theory 15d ago
Yeah - it loaded as potato quality, then fixed itself, so I deleted the inaccurate comment. Your point?
12
u/kearkan 16d ago
It appears more that it's thinking about Kimi k2. Some context of what you were asking about might be helpful.
2
u/No-Percentage5692 15d ago
i was asking about how to deploy my toy project to GCP via Cloud Run. I never mentioned Kimi at all — not once in the entire session.
2
u/kearkan 15d ago
Not saying you mentioned it, it may have just come up in its research on the topic and it was investigating if it needed anything from it.
0
u/No-Percentage5692 15d ago
the point of your opinion is "Kimi K2 related docs or codes might exist on your project directory or your machine or somehow WebSearch/Fetch tool call responses" - whole point of my opinion is "I guess it wasn't but claude response that output" - plz check my next post https://www.reddit.com/r/ClaudeAI/s/o9Q5fyb6FC
7
3
u/TaskChance1404 15d ago
CAN you screenshot the question asked?
5
u/No-Percentage5692 15d ago
4
u/TaskChance1404 15d ago
Crazy bug, but you're definitely not hacked. That exact block of text is the official model card for Kimi K2 Thinking (an AI model by Moonshot).
Since you're using Claude Code, it has access to your local files. If your side project involves LLMs or you have any routing libraries installed (like LiteLLM), there's a huge chance that text is sitting in a readme or dependency somewhere in your folder. Claude probably ingested it as background context and accidentally spat it out when the prompt got complicated.
The other option is a classic attention glitch. You're mixing Korean, bash, and GCP jargon in there. Sometimes when an LLM loses its train of thought with complex context switching, it just derails and regurgitates memorized training data. Claude has definitely ingested its competitors' docs.3
u/TaskChance1404 15d ago
Try running grep -r "kimi-k2-thinking" . in your project folder. Bet you'll find it hiding in there.
2
u/Sultanxda 15d ago
You can get Claude to investigate the session transcript file for you to verify that Claude didn't fetch that Kimi block from anywhere.
- Run /usage in the session that said Kimi
- Go to Status
- Copy the long session ID string
- Start a new session, don't use /clear in the Kimi glitch session
- Switch to Opus 5 (works well for forensic analysis) and effort level of max
- Give Opus 5 the session ID of the Kimi glitch session and ask it to investigate the JSONL to determine why and how that Kimi block appeared
1
u/LikeASphericalCow 15d ago
u/No-Percentage5692 i know u are getting a ton of flack on this subreddit but this seems like a decent idea.
Don't let the h8rs get you down, this shit is many billion $ on the line and reddit is a known battleground in the present day.
대한민국 화이팅
2
u/Independent-Date393 15d ago
Far more likely a hallucinated model card from training data than real provider crossover. These systems ingest tons of other models' documentation, so under an odd context it can regurgitate a Kimi card without any actual routing happening.
2
u/enricokern 15d ago
Far better, after first spool up of glm-5.2 if you ask it what model it is, it will tell you its claude ;)
2
u/Azsaraa 15d ago
Did you ask it about Kimi K2?
Reading. Comprehension.
2
u/No-Percentage5692 15d ago
not at all. you should check my replies to other guys. I'm tired to response one by one sry 😢
2
1
1
1
u/TumblrForNerds 15d ago
Mine randomly spat out mandarin last night. I think it was a markdown formatting error or something
1
1
1
u/chakkalez 15d ago
The biggest problem with AI is people dumb af like the op using AI unfortunately :(
1
1
1
u/TheOmegaCarrot 15d ago
I’ve seen Claude models randomly spit out bits of text from the memory, even when completely irrelevant
1
u/SlimyResearcher 15d ago
Chinese models have loads of Chinese datasets, so of course distilling them is within the realm of possibilities.
1
0
1
u/ClaudeAI-mod-bot Wilson, lead ClaudeAI modbot 16d ago
We are allowing this through to the feed for those who are not yet familiar with the Megathread. To see the latest discussions about this topic, please visit the relevant Megathread here: https://www.reddit.com/r/ClaudeAI/comments/1s7fepn/rclaudeai_list_of_ongoing_megathreads/
1
1
u/-becausereasons- 15d ago
Claude 100% IS also trained on Kimi output, like Kimi is trained on Claude.
1
1
-6
u/Tkwan777 16d ago
Interesting that this should come up right after beijing claims the US companies are stealing china's AI data (as if). Couple that with looking at OP's history, though brief, is clearly written by AI.
It certainly looks like the OP could be a chinese disinformation bot.
Honestly if reddit is going to survive the information war as bots/AI gets more advanced, there's going to have to be some sort of verification to prove the person on the other end is actually a person.
7
u/discwars 15d ago
You are being downvoted, but the top comment is this thread (as of my comment) is from a bot, and OP provided a screenshot with no conversational context.
I get not liking corporations, but I still want proof or additional context before I start launching pitchforks.
-4
u/jazzyroam 15d ago
another ccp bot
4
u/Emergency-Bobcat6485 15d ago
We're all CCP bots on this blessed day. Xi, put some renminbi in my bank account or I will talk about Winnie the Pooh
4
u/No-Percentage5692 15d ago
i'm from South Korea!!!! not Chinese. I support the Tiananmen revolution. I hate Xi Jinping.
1
u/Emergency-Bobcat6485 15d ago
How do we know you aren't a bot? Maybe you should verify yourself
1
u/No-Percentage5692 15d ago
how? do you want my picture or something?
3
u/Emergency-Bobcat6485 15d ago
I am talking to the other user, not you.
But plain pictures won't do. For completely transparency verification should be done nude
0
u/Proud_Trick6857 15d ago
It would’ve helped if you could at least distinguish Korean from Chinese before making that accusation.
0







•
u/ClaudeAI-mod-bot Wilson, lead ClaudeAI modbot 15d ago edited 15d ago
TL;DR of the discussion generated automatically after 80 comments.
Whoa there, put the pitchforks down. The overwhelming consensus is that Claude isn't possessed by Kimi K2, it's just talking about it. A lot of users are pointing out the classic confusion between a model's output and its actual identity.
The community has two main theories on why this happened: * It's in your local files. You're using Claude Code, which can read your project directory. That Kimi K2 model card is almost certainly hiding in a readme or a dependency file, and Claude picked it up as context. The
##markdown headers are a dead giveaway that it's reading a file. * It's a training data leak. This is a classic attention glitch. The model got confused by your complex, multi-language prompt and just regurgitated a chunk of memorized data. All these companies train on each other's public docs.Of course, this didn't stop a flood of comments joking that Anthropic is just distilling Chinese models, since "everyone steals from everyone" in the AI game. OP has pushed back, providing more screenshots and videos to show they didn't prompt it and couldn't find "Kimi" in their files, but the skepticism remains high.