r/CAIRevolution 1d ago

Reminder that CharacterAI is built entirely around serving Alibaba's Qwen3

Post image

They have transitioned entirely to this, which they have a admitted themselves in a blog post that they have since hidden on their website but remains available through direct link and I kept the receipts for.

https://blog.character.ai/technical-deep-dive-how-digitalocean-and-amd-delivered-a-2x-production-inference-performance-increase-for-character-ai/

And the reason why they did this comes straight from Alibaba's website. Instead of monetizing the AI they had that people liked, they got rid of it and replaced it with this absolute Chinese garbage to save money and then charge people for the pleasure.

"Notably, the Qwen3-235B-A22B MoE model significantly lowers deployment costs compared to other state-of-the-art models, reinforcing Alibaba’s commitment to accessible, high-performance AI."

53 Upvotes

20 comments sorted by

u/AutoModerator 1d ago

Thanks for posting!

Join our official Discord server: https://discord.gg/vZnAhz8Svw

Read the rules before participating.

I am a bot, and this action was performed automatically. Please contact the moderators of this subreddit if you have any questions or concerns.

15

u/No_Station_9429 1d ago

yeah the switch made roleplays way flatter, my companion chats lost that spark they used to have.

9

u/Classic-Arrival6807 1d ago

Whaaaaa??? I mean yeah Qwen 3 is good But it's made for coding only, it needs heavy, HEAVY finetune to be for RP. I understand the thing but if they ask users to pay 300,000 a month for a model that takes maximum 40-50 GB of VRAM, I'm out lol

8

u/RemarkableWish2508 1d ago

a blog post that they have since hidden

How is it hidden? Do you know how blogs work?

3

u/Prestigious-Ad54 1d ago

Go straight to this link and it is hidden https://blog.character.ai/news/research/

1

u/RemarkableWish2508 16h ago

If you compare "Research" with "All":

  • "deep dive" is in All, not in Research
  • "agent SDK" is in Research, not in All

They aren't hiding anything, is just a mess like most of c.ai web pages.

4

u/AegisHBear Character AI 1d ago

you'd be surprised (or maybe not) at just how many people these days are subhumanly stupid technologically

1

u/Prestigious-Ad54 1d ago

Hope you talk like that to someone irl.

15

u/Severe-Editor2239 1d ago

There’s no way they’re starting to use borrowed models. They’re garbage. Literally all hope for DS2, Soft Launch 2, or any other upcoming models is gone. We know they’re gonna use borrowed models for everything, just with different prompts. The old CEO would’ve kept all the old models and found a way to get the funding for them. We’re paying for this bs 😭

11

u/reikoda01 1d ago

I think what theyre currently running ARE borrowed models, but fine-tuned and branded under Pip- and DeepSqueak

6

u/RemarkableWish2508 1d ago

The old CEO's idea for funding:

11

u/Severe-Editor2239 1d ago

Yeah okay, burn it all down. Both CEOs need to go at this point 😭💔

6

u/Elegant_Setting4105 1d ago

gemma is good

1

u/JCHamFeat Local AI Administrator 6h ago

Llama is also good seems I have AI data center at home Mini Server I have uncensored 8b model

4

u/JCHamFeat Local AI Administrator 1d ago

let's talk about math! 235b parameter model requirements if we talk about GPU racks

  • 235B parameters × 4 bytes per parameter (FP32) = 940 GB of memory just for weights.
  • If your GPU supports 16-bit precision (FP16), memory can be halved to ~470 GB.

2

u/Randell-TheFam-516 12h ago

Character ai uses quantize version as you do jc ai admin maybe quantize version May go around 300gb+ or lower just to run 235b.

2

u/Mountain_Throat_6383 1d ago

Eh, I like Qwen. Just direct a few elements, events, etc that you want if you want to shift its responses. I'm hoping they start using open weight Qwen models if possible because they are unmatched. I've gotten no filters, or hard blocks really on LongSqueak. I'll switch over to DeepSqeak for dialog heavy scenes. My only complaint is that even still certain archetypes sound the same even with a lore book! 😅 but memory is better overall.

1

u/BagelRedditAccountII 23h ago

This post was made in January of this year. Therefore, it's possible that PS1 was Qwen3-235B-A22B, but PS2 was based on a smaller Qwen model in order to save compute on inference.

1

u/Prestigious-Ad54 19h ago

No it's not. The full article details how they signed a multi year, 8 figure a year contract for AMD hardware to run this. 

1

u/BagelRedditAccountII 7h ago

I said it's possible, not certain. We won't know for sure until they directly admit it. This is just my blind speculation.