Chinese AI lab behind MiniMax M3 (1M-context coding/agentic LLM), open-weight H3 video, and Speech models - available via API and MiniMax Code agent.

Overview

MiniMax is a Chinese AI research lab and product company whose model family spans language, video, speech, image, and music. Its current text flagship, MiniMax M3, is built on a novel sparse-attention architecture (MSA) that reaches a 1-million-token context and is positioned for coding and agentic workflows; the company also ships MiniMaxCode, an agent harness that assembles skill teams to solve longer tasks. On the generative side, MiniMax H3 is an open-weight omni-modal model and Hailuo remains a popular text-to-video line. In our evaluation, MiniMax’s differentiator is the long context plus multimodal breadth from one vendor: you can run document analysis, agents, and video generation inside one account. It is a strong choice for teams already operating in the China ecosystem or wanting an alternative to the usual US frontier models. The caveats are regional payment and access friction, plus API pricing that has moved around, so cost modeling should be done against the live rate card.

Key Features

  • MiniMax M3 language model with 1M-token context for coding and long-document agents
  • MiniMaxCode agent harness that learns habits and assembles skill teams
  • Open-weight H3 omni-modal and Hailuo text-to-video generation
  • Speech-2.8 HD/Turbo for natural multilingual text-to-speech
  • Both Anthropic-compatible and OpenAI-compatible API endpoints
  • Self-hostable open-weight models (M3, H3) for on-prem control

Pricing

PlanPriceForNotes
Free credits$0New API accountsTrial allowance; no music paid API for new users
Individual plan$49/moSolo developersSubscription after 2026 repricing (was $29)
M3 API <=512K ctx$4.20 in / $16.80 out / 1M tokPay-as-you-go textStandard rate; ~$0.59 / $2.37 per 1M (approx 7.1 RMB/USD)
M3 API 512K-1M ctx$8.40 in / $33.60 out / 1M tokLong-context callsPremium band for 1M-context workloads
Open-weight self-hostFreeOrgs wanting controlM3/H3 weights on Hugging Face and ModelScope

Comparison

Next to DeepSeek, MiniMax prices higher but leans into long context and multimodal breadth rather than pure cost leadership. Against Qwen, both are Chinese open-weight players, but MiniMax pushes a 1M-token coding model and a separate agent harness. Versus Kimi, which is best known for long-text chat, MiniMax covers more modalities (video, speech, music) under one roof. Choose MiniMax when you want long-context agents or multi-modal generation from a single vendor.

Compare alternatives

Side-by-side with the 3 closest alternatives.

ToolCategoryPricingVisit
MiniMax (this) chat, code, agentsFree $0 · From $4.2/mo Site ↗
DeepSeekchat, codeFrom $0/mo Site ↗
Qwen (通义千问)chat, codeFree $0 Site ↗
KimichatFree $0/mo · From $0.55/mo Site ↗
MiniMax Current

Chinese AI lab behind MiniMax M3 (1M-context coding/agentic LLM), open-weight H3 video, and Speech models - available via API and MiniMax Code agent.

chatcodeagents
Free $0 · From $4.2/mo

Open, low-cost reasoning models with strong math and coding performance. Frontier-level reasoning at low cost Read our hands-on review and compare the top

chatcode
From $0/mo

Alibaba's open-weight Qwen LLM family and Qwen Chat assistant for multilingual conversation, coding, and document Q&A. Genuinely open-weight smaller

chatcode
Free $0

Moonshot AI's assistant powered by the K2.6 MoE model, built for long-horizon coding, agent swarms, and native multimodal chat with a 262K context window.

chat
Free $0/mo · From $0.55/mo
Editor’s Review
4.3/5
Pros
  • +1M-token context handles repository-scale coding and long documents
  • +Broad multimodal coverage (text, video, speech, image) from one vendor
  • +Open-weight models available for self-hosting and on-prem control
Cons
  • API pricing rose in 2026 and can shift, complicating cost forecasts
  • Regional payment and access friction for non-China users
  • Music generation API closed to new users as of August 2026

MiniMax is the most complete Chinese multimodal lab after the recent M3 and H3 releases, with a genuinely long context that helps agentic coding. Budget for its moving API prices and weigh the open-weight option if data residency matters.

See all reviews →