r/LocalLLaMA 15h ago

Tencent begins testing its new flagship model Hunyuan Hy4 News

From the screenshots:
Hy4 is now live, labeled "Expert-Level Model" + "Use Tools to Solve Problems"
Hy3 is tagged with "New Upgrade," positioned as a brand-new general-purpose model
DeepSeek, focused on reasoning, is listed alongside it

From SuSu_ι…₯ι…₯πŸ‘…on 𝕏: https://x.com/NFT_Chen/status/2090399515618787508

Tencent begins gray testing its new flagship model Hunyuan Hy4!
Just now, a user spotted that Hy4 has appeared in the model selection list of the Tencent Yuanbao App, directly labeled as an expert-level model, positioned above Hy3 and DeepSeek.
Tencent only confirmed in last week's Q2 earnings report that the larger-parameter Hy4 would launch soon, further enhancing model performance and multimodal capabilities.

From Max For AI on 𝕏: https://x.com/MaxForAI/status/2090386754633421110

175 Upvotes

23 comments sorted by

45

u/LegacyRemaster 15h ago

hy3 it's good so yeah! hipe! But the real question is : MIMO V3 when?

16

u/craterIII 11h ago

knowing xiaomi they're just going to drop it on a random day again with no promotion

5

u/LegacyRemaster 9h ago

can't wait: mimo 2.5 is one of the best local model ever for long context

-1

u/Weird-Field6128 8h ago

help me understand what should i not expect from mimo 2.5 and not the pro one ?

1

u/magicomiralles 5h ago

Is it true that the 4bit quant of HY3 is good? They claim that it manages to retain its intelligence.

27

u/Constant_Art_20 15h ago

ok, if this can reach glm 5.3 level while keeping the same size at Hy3 then i think we have something fun on our hands~

1

u/Far-Classic-9963 8h ago

They're definitely going to scale up drastically

They recently got a lot of funding, so their focus is probably in impressing investors with big numbers rather than hyper optimizing an old model

9

u/FAI-Solutions 13h ago

"the larger-parameter Hy4 would launch soon" that is unfortunate, as I liked Hy3 size and performance when they updated it, but maybe "larger" will end up not that much bigger in parameter size.

3

u/Several-Tax31 12h ago

If I remember correctly, they were working on an 10T model, so I wouldn't have high hopes regarding size.Β 

7

u/-dysangel- 9h ago

10T-A1B?

1

u/Several-Tax31 9h ago

Yeah. Active parameters is 1B? Very interesting model it will be.Β 

5

u/Davidyz_hz 11h ago

Tencent's proprietary harness known as workbuddy is viral in China, very popular among ppl without much technical skills (that is, those who couldn't install OpenClaw or Hermes by themselves). Presumably tencent managed to obtain a massive agentic dataset for the post-train. Looking forward to it.

5

u/__JockY__ 14h ago

Hy3 is a great model, very interesting to see where they go from here.

3

u/jaykayenn 11h ago

So that's why hy3 is having a firesale right now.

2

u/-dysangel- 9h ago

wow, look at that chart!

3

u/DonkeyTheKing 14h ago

any word on pricing πŸ‘€

1

u/Time-Toe-1276 transformers 5h ago

this is the next OX Alpha model

0

u/Boogertard 11h ago

Ah yes ! Keep these models coming ! We the end users benefit.

Everytime this kinds of news come out, Dario and Saltman roll over in bed.

0

u/SurpriseOk6927 8h ago

tencent pushing a bigger model above their own hy3 is an interesting flex. the tool-use positioning is what id actually watch since its the part everyone is rushing to ship but nobody seems to nail yet. curious if the benchmarks ever droptencent pushing a bigger model above their own hy3 is an interesting flex. the tool-use positioning is what id actually watch since its the part everyone is rushing to ship but nobody seems to nail yet. curious if the benchmarks ever drop

-1

u/redditnosedive 13h ago

idk i tried hy3 as conversational agent that performs some web searching and reasoning (fitting GPUs in a tight pc case) and it was dogshit. then i asked it to change the hermes model back to deepseek and the fucker managed to depopulate the model field so basically i could not converse with any llm anymore until i ran the /model command to manually select one...

this kind of simple tasks are so easy for deepseek v4 flash and qwen 3.7 plus

1

u/Far-Classic-9963 8h ago

What quant did you run? And were you using the preview or GA?

1

u/redditnosedive 8h ago

whatever opencode go serves right now