Live access is not confirmed for this model. No current purchase price is advertised.
Related provider subscriptions
No current subscription pricing is tracked for this model.
Confirm this specific model, usage limits, and billing terms with the provider. A subscription does not automatically include API credits.
No login needed to compare. Prices are in USD; provider charges are separate from AI Market Cap plans. Context length, caching, tools, taxes, and regional terms can change the final cost. Open weights do not mean free hosting.
---
Quality Score
---
Arena ELO
Undisclosed
Parameters
---
Context
Evidence profile
How complete is this record?
This measures the amount of verifiable public evidence we have, not how capable the model is. A missing field means it has not been verified yet, not that its value is zero.
Recent launch, pricing, benchmark, and API signals linked to this model or its provider.
BenchmarksMiniMaxToday
MiniMax M2.5 - SOTA in Coding and Agent, Designed for Agent Universe
MiniMax M2.5 - SOTA in Coding and Agent, Designed for Agent Universe | MiniMax Models LLM MiniMax M3 MiniMax M2.7 MiniMax M2.5 VIDEO MiniMax H3 SPEECH & MUSIC MiniMax Speech 2.8 MiniMax Music 3.0 Product MiniMax Code MiniMax Design Audio Talkie API Plan Research Company Intelligence with everyone About News Investor Relations Contact Us Models LLM MiniMax M3 NEW MiniMax M2.7 MiniMax M2.5 VIDEO MiniMax H3 NEW SPEECH & MUSIC MiniMax Speech 2.8 NEW MiniMax Music 3.0 NEW
MiniMax M2.5 - SOTA in Coding and Agent, Designed for Agent Universe
MiniMax M2.5 - SOTA in Coding and Agent, Designed for Agent Universe | MiniMax Models LLM MiniMax M3 MiniMax M2.7 MiniMax M2.5 VIDEO MiniMax H3 SPEECH & MUSIC MiniMax Speech 2.8 MiniMax Music 3.0 Product MiniMax Code MiniMax Design Audio Talkie API Plan Research Company Intelligence with everyone About News Investor Relations Contact Us Models LLM MiniMax M3 NEW MiniMax M2.7 MiniMax M2.5 VIDEO MiniMax H3 NEW SPEECH & MUSIC MiniMax Speech 2.8 NEW MiniMax Music 3.0 NEW
MiniMax M2.5 - SOTA in Coding and Agent, Designed for Agent Universe
MiniMax M2.5 - SOTA in Coding and Agent, Designed for Agent Universe | MiniMax Models LLM MiniMax M3 MiniMax M2.7 MiniMax M2.5 VIDEO MiniMax H3 SPEECH & MUSIC MiniMax Speech 2.8 MiniMax Music 3.0 Product MiniMax Code MiniMax Design Audio Talkie API Plan Research Company Intelligence with everyone About News Investor Relations Contact Us Models LLM MiniMax M3 NEW MiniMax M2.7 MiniMax M2.5 VIDEO MiniMax H3 NEW SPEECH & MUSIC MiniMax Speech 2.8 NEW MiniMax Music 3.0 NEW
MiniMax M2.5 - SOTA in Coding and Agent, Designed for Agent Universe
MiniMax M2.5 - SOTA in Coding and Agent, Designed for Agent Universe | MiniMax Models LLM MiniMax M3 MiniMax M2.7 MiniMax M2.5 VIDEO MiniMax H3 SPEECH & MUSIC MiniMax Speech 2.8 MiniMax Music 3.0 Product MiniMax Code MiniMax Design Audio Talkie API Plan Research Company Intelligence with everyone About News Investor Relations Contact Us Models LLM MiniMax M3 NEW MiniMax M2.7 MiniMax M2.5 VIDEO MiniMax H3 NEW SPEECH & MUSIC MiniMax Speech 2.8 NEW MiniMax Music 3.0 NEW
MiniMax M2.5 - SOTA in Coding and Agent, Designed for Agent Universe
MiniMax M2.5 - SOTA in Coding and Agent, Designed for Agent Universe | MiniMax Models LLM MiniMax M3 MiniMax M2.7 MiniMax M2.5 VIDEO MiniMax H3 SPEECH & MUSIC MiniMax Speech 2.8 MiniMax Music 3.0 Product MiniMax Code MiniMax Design Audio Talkie API Plan Research Company Intelligence with everyone About News Investor Relations Contact Us Models LLM MiniMax M3 NEW MiniMax M2.7 MiniMax M2.5 VIDEO MiniMax H3 NEW SPEECH & MUSIC MiniMax Speech 2.8 NEW MiniMax Music 3.0 NEW
MiniMax M2.5 - SOTA in Coding and Agent, Designed for Agent Universe
MiniMax M2.5 - SOTA in Coding and Agent, Designed for Agent Universe | MiniMax Models LLM MiniMax M3 MiniMax M2.7 MiniMax M2.5 VIDEO MiniMax H3 SPEECH & MUSIC MiniMax Speech 2.8 MiniMax Music 3.0 Product MiniMax Code MiniMax Design Audio Talkie API Plan Research Company Intelligence with everyone About News Investor Relations Contact Us Models LLM MiniMax M3 NEW MiniMax M2.7 MiniMax M2.5 VIDEO MiniMax H3 NEW SPEECH & MUSIC MiniMax Speech 2.8 NEW MiniMax Music 3.0 NEW
Video DeltaNet: A Video-Native Hybrid Attention for Livestream Video Generation
Video diffusion models repeatedly process long spatiotemporal token sequences during denoising, making attention a major computational bottleneck. Linear attention offers an appealing alternative and has been widely adopted in recent large language models, but directly applying it to video models often fails to preserve the fine-grained interactions required for high-quality generation. We present Video DeltaNet (VDN), which combines local Softmax attention with bidirectional linear memory for long-range video context. Its linear branch introduces Video Delta Attention (VDA), which updates memory once per frame by jointly incorporating its spatial tokens. Separate output projections and learnable gates calibrate the two branches, while a staged teacher-alignment recipe progressively introduces the new pathway into pretrained models. We instantiate VDN on MiniMax H3, applying the hybrid to video-to-video interactions while retaining Softmax for interactions involving text or audio. With eight-step distillation and an optimized SGLang serving stack, VDN-H3 completes DiT denoising for a 14.3-second, 768p video in 6.70 seconds on eight NVIDIA B200 GPUs, corresponding to a 14.5x speedup over the 50-step dense H3 baseline on the same GPU count.
Can MiniMax-H3 Reason About the Physical World? An Evaluation of Omni-Modal Generative Model
Recent Omni-Modal Generative Models (Omni-Models) have advanced content generation toward unified modeling of text, images, video, and audio. MiniMax-H3 exemplifies this transition by combining multimodal context understanding with joint audio-visual generation in a shared latent framework. Its unified architecture raises a fundamental question: Can multimodal alignment improve the model's world reasoning, and what new evaluation paradigms do omni-modal inputs enable? To investigate this question, this work introduces a comprehensive evaluation framework organized around four complementary dimensions of physical world reasoning. Unlike existing evaluation frameworks for video generation and world models, which are often constrained by limited input modalities and evaluation settings where prompts closely match the target video content, our evaluation is specifically designed to exploit the multimodal inputs of Omni-Model. We construct a diverse set of novel tasks that require models to integrate complementary information across modalities. Specifically, we consider four scenarios, including implicit prompts paired with multiple frames, audio-image, prefix-videos, and audio-video inputs. Every single modality provides only partial evidence about the underlying event, requiring the model to jointly reason over the complementary semantic cues to infer latent event states and future dynamics. Across 517 evaluation instances, MiniMax-H3 achieves an overall success rate of 41.97%. Video-based Decision Reasoning yields the highest success rate at 56.00%, while Audio-based Disambiguation Reasoning is the weakest, reaching only 27.40%. These results indicate that effective multimodal integration remains key to fully exploiting the benefits of diverse input modalities. The project is available at https://github.com/gulucaptain/MiniMax-H3-Reason.
VC-Attention: Value Smoothing and Softmax Casting for Low-bit Attention
Diffusion Transformers deliver state-of-the-art video generation, but their long spatiotemporal sequences make attention the dominant deployment cost, and a deployable low-bit kernel must be accurate and fast. Accuracy is limited by outliers: a block's quantization scale is set by its largest entries, leaving typical entries confined to a narrow range of representable values. Prior work smooths queries and keys, but value outliers follow no fixed channel or spatiotemporal structure and remain the dominant source of output error. Speed is limited by softmax: low-bit Tensor Cores accelerate only the two matrix multiplications, so the high-precision exponential between them becomes the longest pipeline stage on datacenter GPUs. We propose VC-Attention, a training-free low-bit attention framework that addresses both by pairing Value smoothing with a fused probability Cast. V-Smooth reorders value tokens by lightweight online clustering, so the tokens in a hardware block quantize well together. It quantizes only the residual after subtracting the block mean, and restores that mean from the row sum the online softmax already maintains. ExpCast-FP8 maps log-domain scores directly to E4M3 probability codes with one fused multiply-add, eliminating the FP32 exponential and the format conversion. We implement VC-Attention for B200, B300, H200, RTX PRO 6000, and RTX 5090. Across Wan2.2, LongCat-Video, HunyuanVideo-1.5, and MiniMax-H3, VC-Attention improves fidelity over low-bit baselines, speeds up the attention kernel over BF16 FlashAttention-4 by 1.46-1.59x on datacenter Blackwell and Hopper and by 2.3-3.6x on workstation cards, and generates a clip 1.13-1.19x and 1.36-1.70x faster end to end.
MiniMax M2.5 - SOTA in Coding and Agent, Designed for Agent Universe
MiniMax M2.5 - SOTA in Coding and Agent, Designed for Agent Universe | MiniMax Models LLM MiniMax M3 MiniMax M2.7 MiniMax M2.5 VIDEO MiniMax H3 SPEECH & MUSIC MiniMax Speech 2.8 MiniMax Music 3.0 Product MiniMax Code MiniMax Design Audio Talkie API Plan Research Company Intelligence with everyone About News Investor Relations Contact Us Models LLM MiniMax M3 NEW MiniMax M2.7 MiniMax M2.5 VIDEO MiniMax H3 NEW SPEECH & MUSIC MiniMax Speech 2.8 NEW MiniMax Music 3.0 NEW
MiniMax M2.5 - SOTA in Coding and Agent, Designed for Agent Universe
MiniMax M2.5 - SOTA in Coding and Agent, Designed for Agent Universe | MiniMax Models LLM MiniMax M3 MiniMax M2.7 MiniMax M2.5 VIDEO MiniMax H3 SPEECH & MUSIC MiniMax Speech 2.8 MiniMax Music 3.0 Product MiniMax Code MiniMax Design Audio Talkie API Plan Research Company Intelligence with everyone About News Investor Relations Contact Us Models LLM MiniMax M3 NEW MiniMax M2.7 MiniMax M2.5 VIDEO MiniMax H3 NEW SPEECH & MUSIC MiniMax Speech 2.8 NEW MiniMax Music 3.0 NEW
MiniMax M2.5 - SOTA in Coding and Agent, Designed for Agent Universe
MiniMax M2.5 - SOTA in Coding and Agent, Designed for Agent Universe | MiniMax Models LLM MiniMax M3 MiniMax M2.7 MiniMax M2.5 VIDEO MiniMax H3 SPEECH & MUSIC MiniMax Speech 2.8 MiniMax Music 3.0 Product MiniMax Code MiniMax Design Audio Talkie API Plan Research Company Intelligence with everyone About News Investor Relations Contact Us Models LLM MiniMax M3 NEW MiniMax M2.7 MiniMax M2.5 VIDEO MiniMax H3 NEW SPEECH & MUSIC MiniMax Speech 2.8 NEW MiniMax Music 3.0 NEW
MiniMax M2.5 - SOTA in Coding and Agent, Designed for Agent Universe
MiniMax M2.5 - SOTA in Coding and Agent, Designed for Agent Universe | MiniMax Models LLM MiniMax M3 MiniMax M2.7 MiniMax M2.5 VIDEO MiniMax H3 SPEECH & MUSIC MiniMax Speech 2.8 MiniMax Music 3.0 Product MiniMax Code MiniMax Design Audio Talkie API Plan Research Company Intelligence with everyone About News Investor Relations Contact Us Models LLM MiniMax M3 NEW MiniMax M2.7 MiniMax M2.5 VIDEO MiniMax H3 NEW SPEECH & MUSIC MiniMax Speech 2.8 NEW MiniMax Music 3.0 NEW