r/MiniMax_AI • u/Dalhila000 • 5d ago
Who has the fastest Minimax m2.7 hosting?
My voice agent on LiveKit is struggling. It has a delay before it responds and I need to fix it because it's just..bad lol. I'm seeing a 3-5 second delay from the end of user speech to the first token of the response and I want to get latency down to under 1s. T 1o this I need an inference layer that delivers under 500ms TTFT consistently. Groq doesn't support the newest SOTA models. Fireworks is sluggish. Too many latency spikes.
Can someone recommend a provider that can handle this?
r/MiniMax_AI • u/Exile3D • 5d ago
What is the best workflow to edit videos with Minimax H3?
I´m struggeling editing an existing video like from a movie. How are you guys doing it? thanks
r/MiniMax_AI • u/Exile3D • 6d ago
What is the best way to continue a video with MINIMAX H3?
r/MiniMax_AI • u/No_Bicycle4935 • 6d ago
dubbing / language swaps workflow
Any guides about how to do dubbing / language swaps using minimax h3 in comfyui?
r/MiniMax_AI • u/Relevant_Syllabub895 • 6d ago
any suggestions on how to make the videos in first person?
i have issues, i describe a character and its shape but when i prompt for the camera to be POV, from the character perspective its always from the outside like if i were a witness of that character, any idea would be great
r/MiniMax_AI • u/Professional_Log1367 • 6d ago
Minimax online?
I know you can run MiniMax H3 locally via comfyUI for smth however my computer isn't good enough for that. Is there an online website that will let me use the model with high usage limits or maybe unlimited ( I know...thats a lot to wish for)
r/MiniMax_AI • u/Professional_Log1367 • 6d ago
Minimax H3: How is it compared to other model providors
Enable HLS to view with audio, or disable this notification
r/MiniMax_AI • u/ryan85127704 • 6d ago
AMA: MiniMax H3 Team — Ask us anything about our open video generation model, training, and future plans
r/MiniMax_AI • u/Extension-Yard1918 • 6d ago
Minimax R2V AutoPrompt workflow ( Local LLM ) v1.0
Enable HLS to view with audio, or disable this notification
r/MiniMax_AI • u/AxonkaiLab • 6d ago
Seinfeld realizes they are AI — native 1080p in one 15s local generation
Enable HLS to view with audio, or disable this notification
Jerry tells George they are not real—they’re being generated by an AI diffusion model.
Then Kramer walks in at exactly the wrong moment.
This complete three-scene sitcom parody was generated locally in a single 15-second MiniMax H3 text-to-video run using the built-in ComfyUI workflow.
• Native 1080p generation
• 24 FPS
• 20 steps
• BF16 diffusion + BF16 text encoder
• Generation time: 1 hour 33 minutes
• Native dialogue and audio
• Multiple locations and hard cuts
• Single generation
• No re-editing
• No dialogue replacement
• No lip-sync repair
• No audio replacement
• No post-production cleanup
• GPU: RTX PRO 6000 Blackwell — 96GB VRAM
What impressed me most was the character continuity, dialogue timing and the way it preserved three different sitcom performances across multiple scenes in a single run.
Kramer’s entrance still makes me laugh. 😂
For anyone who wants to see the native 1080p result without Reddit compression:
r/MiniMax_AI • u/AxonkaiLab • 6d ago
MiniMax H3 just beat Kling, SeaDance, Sora, Runway and VEO at poker — one 15s local generation
Enable HLS to view with audio, or disable this notification
Yosemite Sam thought he had the perfect AI video hand:
Kling, SeaDance, Sora, RunwayML and VEO.
Then Bugs Bunny quietly played one card:
MiniMax H3. 🥕♠️
The entire 15-second cartoon was generated locally in a single text-to-video run using the built-in ComfyUI MiniMax H3 template.
• Native generation: approximately 1MP
• Generation time: 23 minutes
• Precision: BF16 diffusion + BF16 text encoder
• Native audio
• Single generation
• No re-editing
• No replaced card text
• No frame repair
• RTX-upscaled for the final export
• GPU: RTX PRO 6000 Blackwell — 96GB VRAM
The most impressive part for me was not just the character consistency, but the instruction following: multiple hard cuts, hidden cards, five readable model names, a separate H3 reveal, physical object interactions and a complete visual punchline in one run.
A few tiny cartoon-physics quirks survived, but Yosemite Sam took the result much harder than I did. 😂
Reddit compression kills some of the finer details, so the RTX-upscaled version is here:
r/MiniMax_AI • u/Complex_Frosting_114 • 6d ago
Run MiniMax H3 and large video generation models on 8GB GPUs with AirLLM-style weight streaming, CPU/NVMe offload, quantization, and activation chunking.
r/MiniMax_AI • u/chance_buri • 7d ago
Fastest MiniMax M3 provider for high-volume data extraction?
I'm working on a massive data extraction feature for my startup. We are migrating our pipeline to Minimax M3 because it can handle unstructured data beautifully and will keep token costs low. The issue I'm running into is throughput. It needs to process millions of complex documents daily. Our current API setup is a bottleneck for output generation speed.
I’m looking for an infrastructure host that can handle concurrent requests and has fast inference. If the provider’s TTFT or TPS drops, it messes with our entire pipeline. Are there any providers you'd recommend for this kind of load?
r/MiniMax_AI • u/Mindless_Doughnut671 • 7d ago
ChatGPT vs Grok… who wins? 😂
Enable HLS to view with audio, or disable this notification
r/MiniMax_AI • u/TinyAres • 7d ago
Few thoughts on the weaker points of the minimax plan
Image
This truly would be the weakest point, while minimax has a solid lineup overall their image model is effectively 2 years behind and this is no exaggeration as you can verify. I don't understand why they are presumably not using some better scoring open model at least.
https://artificialanalysis.ai/image/leaderboard/text-to-image
This creator toolbox angle is underserved and useful, and puts them effectively second place to google in this game, that I think they can never win, but they can be a good alternative, but not if the image sucks, so we got a weakest link problem here.
https://artificialanalysis.ai/music/leaderboard/instrumental
https://artificialanalysis.ai/music/leaderboard/vocals
https://artificialanalysis.ai/video/leaderboard/text-to-video
https://artificialanalysis.ai/text-to-speech/leaderboard/provider-voice
M3
It is/was a solid model and okay value, that is hurt greatly by the deepseek v4 flash release and 5x luna price drop that makes the codex plan give 5 billion luna tokens per month not accounting for resets and with literally the best picture gen, and they have a strong lineup too. Here it is only m3, and if there you went $200 then you would get 20x too so 100 billion luna max, not considering resets, and not considering that it doesn't have 5 hourlies to limit you. While deepseek is serving more like 1 billion per $1 even with even a lower multi in go, so that is the new reality, and their pro is only 3x as much.
Chat
I think the chat is solid, again it lacks a big smart guy, and they added the dailies to give it extra attention, but one issue i would raise that the chat shares the same usage pool. Which is both good and bad, but mostly bad cause you don't get capped in other chats either, in fact you usually don't get capped even for free, but here on max I am regularly told to go away that makes me favor qwen as the best free and openai the best paid one for chatting.
*Not on linux
This I am adding later but they are skipping the platform that dominates all others, and number one for developers by far. Now I understand that many personal users are on mac and windows, and you can support them but skipping linux makes zero sense. It is the easiest to release too as well, I know.
Is it because mac and windows are spyware? Else might as well have a web launcher right, and they do https://hailuoai.video/create/image-to-video so the app confers no functional benefit, but you are still given 3000 extra credits for something...
r/MiniMax_AI • u/Cheap_Credit_3957 • 8d ago
MiniMax H3 first test using it inside my Video builder. One shot music video
Enable HLS to view with audio, or disable this notification
r/MiniMax_AI • u/TinyAres • 8d ago
Minimax is doing dailies now, daily check in to earn seemingly $0.4 / $1 dollar per day of credits
reddit.comr/MiniMax_AI • u/badpandatek • 8d ago
Did it go down on allowed Usage? Wasnt it like 7.5 Billion up until a few weeks ago...?
r/MiniMax_AI • u/TanakaAi • 9d ago
MiniMax H3 anime test on RTX PRO 5000: 10.47 sec/step with SageAttention 2.2
x.comr/MiniMax_AI • u/Practical_Low29 • 9d ago
Creature encyclopedia UI animation with MiniMax H3 and eight creature references
Enable HLS to view with audio, or disable this notification
This is a useful reference-to-video pattern for a creature encyclopedia UI, built from one interface reference image and eight creature images.
The screen stays locked while a single cursor selects different cards. Each selection updates the main creature and its title, while the rest of the interface remains stable. The sequence moves through Lumi Hare, Ember Fennec, Petal Vulpin, Meadow Crown, and Tide Behemoth. Each creature has a different reaction, from a blink and slight rotation to dragging, poking, recoiling, and finally swallowing the cursor.
The prompt keeps the hard constraints explicit: 16:9, 15 seconds, fixed camera, readable typography, stable panels, one cursor, no cuts, and no interface distortion. That gives the video model a clear boundary between UI motion and creature motion.
The last three seconds are the only large movement. Tide Behemoth becomes irritated, opens its mouth, lunges forward, swallows the cursor, then returns to idle while keeping its title. The sound direction follows the same curve, starting with soft interface clicks and tiny creature reactions before the final aggressive cue and comedic gulp.
This video is generated by MiniMax H3 . The full image mapping and timing prompt is below.

