infery
← All models

MiniMax models

45 models on Infery.ai.

H3 Max Camera Controls

h3-max-camera-controls

H3 Max Multi Angle turns a single image into a video with precise, keyframe-based control over the camera's orbit, elevation, and distance…

3.13 cr / sec
Video output

H3 Max Reference to Video

minimax/h3-max/reference-to-video

6.25 cr / sec
Video output

H3 Max Turbo Text to Video

h3-max-turbo

1.56 cr / sec
Video output
Hailuo 02

Hailuo 02

hailuo-02

MiniMax Hailuo 02 generates high-quality video at up to 1080p with realistic physics and strong prompt following. Use Hailuo 02 with an API.

1.88 cr / sec
Video output
Hailuo 2.3

Hailuo 2.3

hailuo-2.3

MiniMax Hailuo 2.3 generates high-fidelity video with realistic human motion, cinematic VFX, and strong style adherence at up to 1080p.

5.83 cr / sec
Video output
Hailuo 2.3 Fast

Hailuo 2.3 Fast

hailuo-2.3-fast

MiniMax Hailuo 2.3 Fast is a lower-latency image-to-video model with core motion quality and visual consistency at up to 1080p.

3.96 cr / sec
Video output

Image 01

image-01

MiniMax Image 01 generates images from text with character reference support, detailed lighting, and realistic human subjects.

1.25 cr / image
Image output

Minimax

minimax-hailuo-02-fast-image-to-video

Create blazing fast and economical videos with MiniMax Hailuo-02 Image To Video API at 512p resolution

2.13 cr / sec
Video output

Minimax

minimax-preview-speech-2.5-hd

Generate speech from text prompts and different voices using the MiniMax Speech-02 HD model, which leverages advanced AI techniques to…

0.013 cr / char
Audio output

Minimax

minimax-preview-speech-2.5-turbo

Generate fast speech from text prompts and different voices using the MiniMax Speech-02 Turbo model, which leverages advanced AI techniques…

0.00750 cr / char
Audio output
MiniMax-01

MiniMax-01

minimax-01

MiniMax-01 is a combines MiniMax-Text-01 for text generation and MiniMax-VL-01 for image understanding.

1M ctx

25 cr in / 1M138 cr out / 1M
ChatStreamingVision

MiniMax H3 Image to Video LoRA

h3-image-to-video-lora

Images into video with synchronized audio using MiniMax H3; your image becomes the first frame, prompt optional, with trained LoRA support…

7.81 cr / sec
Video output

MiniMax H3 Max Text to Video

h3-max

3.13 cr / sec
Video output

MiniMax H3 Reference to Video LoRA

h3-reference-to-video-lora

References into video with synchronized audio using MiniMax H3

7.81 cr / sec
Video output

MiniMax H3 Text to Video

h3-text-to-video

MiniMax H3 is a frontier video model.

6.25 cr / sec
Video output

MiniMax H3 Text to Video LoRA

h3-text-to-video-lora

Generate video with synchronized audio from a text prompt using MiniMax H3; load a trained LoRA at adjustable strength to lock in style…

7.81 cr / sec
Video output

MiniMax Hailuo 02 [Pro] (Text to Video)

minimax-hailuo-02-pro-text-to-video

MiniMax Hailuo-02 Text To Video API (Pro, 1080p): Advanced video generation model with 1080p resolution

10 cr / sec
Video output

MiniMax Hailuo 02 [Standard] (Text to Video)

minimax-hailuo-02-standard-text-to-video

MiniMax Hailuo-02 Text To Video API (Standard, 768p): Advanced video generation model with 768p resolution

5.63 cr / sec
Video output

MiniMax Hailuo 2.3 Fast [Pro] (Image to Video)

minimax-hailuo-2.3-fast-pro-image-to-video

MiniMax Hailuo-2.3-Fast Image To Video API (Pro, 1080p): Advanced fast image-to-video generation model with 1080p resolution

41.25 cr / video
Video output

MiniMax (Hailuo AI) Music

minimax-music

Generate music from text prompts using the MiniMax model, which leverages advanced AI techniques to create high-quality, diverse musical…

4.38 cr / req
Music output

MiniMax (Hailuo AI) Music v1.5

minimax-music-v1.5

Generate music from text prompts using the MiniMax model, which leverages advanced AI techniques to create high-quality, diverse musical…

3.75 cr / req
Music output

MiniMax (Hailuo AI) Text to Image

minimax-image-01

Generate high quality images from text prompts using MiniMax Image-01. Longer text prompts will result in better quality images.

1.25 cr / image
Image output

MiniMax (Hailuo AI) Video 01

minimax-video-01-live-image-to-video

Generate video clips from your images using MiniMax Video model

62.5 cr / video
Video output

MiniMax (Hailuo AI) Video 01

minimax-video-01-image-to-video

Generate video clips from your images using MiniMax Video model

62.5 cr / video
Video output

Minimax Image Subject Reference

minimax-image-01-subject-reference

Generate images from text and a reference image using MiniMax Image-01 for consistent character appearance.

1.25 cr / image
Image output
MiniMax M1

MiniMax M1

minimax-m1

MiniMax-M1 is a large-scale, open-weight reasoning model designed for extended context and high-efficiency inference.

1M ctx

50 cr in / 1M275 cr out / 1M
ChatStreamingTools
MiniMax M2

MiniMax M2

minimax-m2

MiniMax-M2 is a compact, high-efficiency large language model optimized for end-to-end coding and agentic workflows.

197K ctx

31.88 cr in / 1M128 cr out / 1M
ChatStreamingToolsJSON
MiniMax M2.1

MiniMax M2.1

minimax-m2.1

MiniMax-M2.1 is a lightweight, state-of-the-art large language model optimized for coding, agentic workflows, and modern application…

197K ctx

37.5 cr in / 1M150 cr out / 1M3.75 cr cached in / 1M
ChatStreamingToolsJSON
MiniMax M2.5

MiniMax M2.5

minimax-m2.5

MiniMax-M2.5 is a SOTA large language model designed for real-world productivity.

197K ctx

33.75 cr in / 1M135 cr out / 1M3.38 cr cached in / 1M
ChatStreamingToolsJSON
MiniMax M2.7

MiniMax M2.7

minimax-m2.7

MiniMax-M2.7 is a next-generation large language model designed for autonomous, real-world productivity and continuous improvement.

197K ctx

37.5 cr in / 1M150 cr out / 1M7.5 cr cached in / 1M
ChatStreamingToolsJSON
MiniMax M2-her

MiniMax M2-her

minimax-m2-her

MiniMax M2-her is a dialogue-first large language model built for immersive roleplay, character-driven chat, and expressive multi-turn…

66K ctx

37.5 cr in / 1M150 cr out / 1M3.75 cr cached in / 1M
ChatStreaming
MiniMax M3

MiniMax M3

minimax-m3

MiniMax-M3 is a multimodal foundation model from MiniMax. 1,048,576 token context window, maximum output of 131,072 tokens.

1M ctx

37.5 cr in / 1M150 cr out / 1M7.5 cr cached in / 1M
ChatStreamingVisionToolsJSON

Minimax Music

minimax-music-v2

Generate music from text prompts using the MiniMax Music 2.0 model, which leverages advanced AI techniques to create high-quality, diverse…

3.75 cr / req
Music output

Minimax Music 2.5

minimax-music-v2.5

MiniMax Music 2.5 creates complete tracks with singing, backing music, and detailed arrangements from lyrics and a style description.

18.75 cr / track
Music output

Minimax Music 2.6

minimax-music-v2.6

MiniMax Music 2.6 creates complete tracks with singing, backing music, and detailed arrangements from lyrics and a style description.

18.75 cr / track
Music output

MiniMax Music 3

music-3

MiniMax Music 3 is a high-performance music generation model for creating complete songs up to five minutes long

0.250 cr / sec
Music output

MiniMax Speech-02 HD

minimax-speech-02-hd

Generate speech from text prompts and different voices using the MiniMax Speech-02 HD model, which leverages advanced AI techniques to…

0.013 cr / char
Audio output

MiniMax Speech-02 Turbo

minimax-speech-02-turbo

Generate fast speech from text prompts and different voices using the MiniMax Speech-02 Turbo model, which leverages advanced AI techniques…

0.00750 cr / char
Audio output

MiniMax Speech 2.6 [HD]

minimax-speech-2.6-hd

Generate speech from text prompts and different voices using the MiniMax Speech-2.6 HD model, which leverages advanced AI techniques to…

0.013 cr / char
Audio output

MiniMax Speech 2.6 [Turbo]

minimax-speech-2.6-turbo

Generate speech from text prompts and different voices using the MiniMax Speech-2.6 HD model, which leverages advanced AI techniques to…

0.00750 cr / char
Audio output

MiniMax Speech 2.8 [HD]

minimax-speech-2.8-hd

Generate speech from text prompts and different voices using the MiniMax Speech-2.8 HD model, which leverages advanced AI techniques to…

0.013 cr / char
Audio output

MiniMax Speech 2.8 [Turbo]

minimax-speech-2.8-turbo

Generate speech from text prompts and different voices using the MiniMax Speech-2.8 Turbo model, which leverages advanced AI techniques to…

0.00750 cr / char
Audio output

Video 01

video-01

MiniMax Video 01 (Hailuo) generates 6-second videos from text or images at 720p with cinematic camera movement. Use Video 01 with an API.

2.08 cr / sec
Video output

Video 01 Director

video-01-director

MiniMax Video 01 Director generates 720p videos with precise camera movement control using bracketed commands.

2.08 cr / sec
Video output

Video 01 Live

video-01-live

MiniMax Hailuo Live is an image-to-video model trained for Live2D and animation, with smooth motion and facial expression control.

2.08 cr / sec
Video output