Image-Text-to-Video
MiniMax H3
Diffusers
Safetensors
text-to-video
image-to-video
video-to-video
text-to-audio-video
image-to-audio-video
image-text-to-audio-video
video-to-audio-video
audio-to-audio-video
audio-video-generation
multimodal
synchronized-audio-video
reference-to-audio-video
Instructions to use MiniMaxAI/MiniMax-H3 with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- MiniMax H3
How to use MiniMaxAI/MiniMax-H3 with MiniMax H3:
# No code snippets available yet for this library. # To use this model, check the repository files and the library's documentation. # Want to help? PRs adding snippets are welcome at: # https://github.com/huggingface/huggingface.js
- Diffusers
How to use MiniMaxAI/MiniMax-H3 with Diffusers:
pip install -U diffusers transformers accelerate
import torch from diffusers import DiffusionPipeline # switch to "mps" for apple devices pipe = DiffusionPipeline.from_pretrained("MiniMaxAI/MiniMax-H3", dtype=torch.bfloat16, device_map="cuda") prompt = "Astronaut in a jungle, cold color palette, muted colors, detailed, 8k" image = pipe(prompt).images[0] - Inference
- Notebooks
- Google Colab
- Kaggle
[FL2VA] Anime style videos have half the FPS
#74
by ChiNoel - opened
Anime style videos generated by MiniMax-H3 are likely to be 12 FPS instead of 24 (the resulting 24 FPS video will have each frame repeated twice), or they can be of variable framerate (where one part is smooth and the other part is stuttering). I can only get like 1 out of 10+ videos that are actually smooth 24 FPS.
I know this is mostly contributed by the training data where most anime is just like this. But I struggle to generate one in real 24 FPS, even by using prompts like "Smooth motion", "Fluid", "High framerate".
Does anyone have any success in generating real 24 FPS anime style videos consistently?