Image-Text-to-Video
MiniMax H3
Diffusers
Safetensors
text-to-video
image-to-video
video-to-video
text-to-audio-video
image-to-audio-video
image-text-to-audio-video
video-to-audio-video
audio-to-audio-video
audio-video-generation
multimodal
synchronized-audio-video
reference-to-audio-video
Instructions to use MiniMaxAI/MiniMax-H3 with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Diffusers
How to use MiniMaxAI/MiniMax-H3 with Diffusers:
pip install -U diffusers transformers accelerate
import torch from diffusers import DiffusionPipeline # switch to "mps" for apple devices pipe = DiffusionPipeline.from_pretrained("MiniMaxAI/MiniMax-H3", dtype=torch.bfloat16, device_map="cuda") prompt = "Astronaut in a jungle, cold color palette, muted colors, detailed, 8k" image = pipe(prompt).images[0] - Inference
- Notebooks
- Google Colab
- Kaggle
ใๆ่ใ MiniMax ่ง้ขๆจกๅๅ จ็ๅผๆบ็ๆ
pinnedโค๏ธ 4
3
#61 opened 21 days ago
by
MiniMax-AI
Any License Question Ask here!
pinned๐ 17
54
#12 opened 25 days ago
by
ryanlee-dev
ๅ็ฐminimax็ๅฐ็็ต๏ผ็จ็ผ้ฎ้ข๏ผไบบ็ฉๅ ไนไธไผ็จ็ผ็
1
#96 opened 3 days ago
by
easonchow0419
Please provide more in-depth documentation for Full-Reference Mode
๐ 5
#95 opened 5 days ago
by
infearia
Running on <4gb Vram
๐ฅ 1
#93 opened 8 days ago
by
Jit2024
H3-Regenerate-2K open weights?
โ 2
7
#92 opened 9 days ago
by deleted
MiniMax H3 Ref2VA + audio recasts identity, while FL2VA + Ref2V Turbo preserves it โ expected behavior?
#91 opened 9 days ago
by
lorentaken
PURE MOJO STACK
#90 opened 10 days ago
by
alexnvo
่ธ้ฆๅ ้็ๆฌ
#89 opened 10 days ago
by
ReSerendipity
Any plans to release a non-distilled base model for training?
๐๐ 4
1
#88 opened 10 days ago
by
AndroYD
่ธๅดฉ็ๅฆๆญคๅๅฎณ๏ผไฝๆถไฟฎๅค~
1
#87 opened 11 days ago
by
criuslv-cn
This is currently the best openโsource video model, yet the conservative algorithm of the VAE seems to create a computational bottleneck?
๐โค๏ธ 9
#85 opened 12 days ago
by
ass5002
Getting different Accents when speaking the same language
1
#84 opened 12 days ago
by
Hcheyne
Ref2VA environmental audio feels too quiet in dialogue scenes
โ 1
2
#82 opened 12 days ago
by
szczypen
Question about prompting on static elements
#81 opened 13 days ago
by
leonnn1
Thanks and Suggestions
#80 opened 15 days ago
by
Pvok
Omni-Rewriter Replay: observe a clip into validated H3 PE (open harness)
#79 opened 15 days ago
by
Wayne-King
Color distortion occurs when encoding and decoding images using a video VAE.
2
#78 opened 15 days ago
by
l13462580123
checkpoint inconsistent between original and diffusers
#77 opened 15 days ago
by
l13462580123
Weird gibberish / clipped voice at the very start of H3 videos (reference to video)
๐ 4
10
#76 opened 15 days ago
by
jesleocizi
[FL2VA] Anime style videos have half the FPS
2
#74 opened 16 days ago
by
ChiNoel
Call MiniMax H3 with your existing OpenAI SDK
#72 opened 16 days ago
by
irene891107
Some question about VIDEO_PROMPT_WRITING_GUIDE_ref_en.md
2
#71 opened 18 days ago
by
iouzzr
Minimax H3 prompt adherence really varies depending on the resolution
5
#65 opened 20 days ago
by
TheBobun
How define audio ref to the Subject 1 - Ref2V
10
#64 opened 20 days ago
by
Gamb
็ฌฌไธๆฌกๆฅ่งฆๅฑไปฌ่ฟไธช็ฝ็ซ๏ผๆๅฅฝๅ ไธช้ฎ้ขๆณ่ฆ้ฎ
3
#63 opened 20 days ago
by
zhao7259
Update documentation to Please add how to character swap for Ref2V/Image to Video.
1
#60 opened 21 days ago
by
CoolaidFun
่ฑ้ขๅธฆ๏ผ่ฑไธ่กฃ๏ผ่ฑๅธฆๆญ่ขข็้็ญ่ฟๆ ๆณๆญฃๅธธ่กจ็ฐ
1
#59 opened 21 days ago
by
sumirecccp
MiniMax H3 Prompt Enhancer, powered by a fine-tuned 350M-parameter model
โค๏ธ 3
1
#58 opened 21 days ago
by
geocine
Will RunPod work??
#56 opened 22 days ago
by
tonyface
Almost VR support
๐ 4
#55 opened 22 days ago
by
Ddfgddsd
่ฏทๆไธไธ๏ผ่ฟไธชๆจกๅๅฆไฝ็ๆไธไธชๅฏไปฅ้ฆๅฐพๅพช็ฏ็่ง้ข๏ผ
5
#54 opened 22 days ago
by
oioitff
lora่ฎญ็ปๆไปไน้่ฆๆณจๆ็ๅฐๆน๏ผ
#53 opened 22 days ago
by
wuyuetiger
Ref2va always have some noise
๐โค๏ธ 2
3
#50 opened 23 days ago
by
jiangjihua
Vllmomni minmaxh3 roadmap
#49 opened 23 days ago
by
feizhai123
sparse-attention inference
3
#48 opened 23 days ago
by
mzbac
The actual correct system prompt for IT2V - (Took me a while)
๐ 6
1
#47 opened 24 days ago
by
cushycrux
Where is the money, Lebowski? The "duct-tape" architecture review)))
โ 3
26
#46 opened 24 days ago
by
Qozimo
Possibly a small mistake in docs/VIDEO_PROMPT_WRITING_GUIDE_base_en.md
#45 opened 24 days ago
by
hum-ma
Is this a CFG distilled model or were the weights trained without CFG from the start?
๐๐ 4
#44 opened 24 days ago
by
natalie5
LoRA Training with MiniMax H3 - Question
1
#43 opened 24 days ago
by
Tomcat2048
Thanks a lot
โค๏ธ 14
#42 opened 24 days ago
by
DodoPapa
ๆๆฏ้ฟ้็้ซ็ฎก
๐ 5
9
#41 opened 24 days ago
by
wjm17173
Amazing job ! really thank you for sharing and pushing forward the video generation technology!
๐ 1
#40 opened 24 days ago
by
aiclouddaily
Minimax H3 2K Upscaler?
4
#39 opened 24 days ago
by
AlperKTS
Only MiniMax-H3 can do !
๐ฅ 5
#38 opened 24 days ago
by
sunnyboxs
้ๅธธๆฃ็ๅคๆจกๆ่ฝๅใๆฏๅฆๅฏไปฅๅบไบ่ฟไธชๆฉๅฑ็ๅพ่ฝๅ๏ผ
5
#37 opened 24 days ago
by
wdtd
vLLM-Omni ComfyUI integration for MiniMax-H3 (T2VA / FL2VA / Ref2VA)
๐ 2
#36 opened 24 days ago
by
shunyang90
my opinion about nsfw
๐ 14
7
#35 opened 24 days ago
by
Arun63