Minimax-h3_Singularity
Singularity fine-tune of MiniMax-H3 for HDR text-, image- and reference-to-video generation in ComfyUI.
Model Description
π Model Overview
Minimax-h3_Singularity is a comprehensive fine-tuned fusion model specialized in enhancing the capabilities of MiniMax-H3. Designed as a versatile multimodal video generation model, it natively supports Text-to-Video (T2V), Image-to-Video (I2V), Reference-to-Video (Ref2V), and Video-to-Video (V2V) workflows within ComfyUI.
Built upon a strategic fusion of key checkpoints (including ref, fl, b25-49, etc.), this model underwent deep high-step fine-tuning. To preserve the original model's foundational strengths and broad generalization while solving artifacts introduced by high-step training, we spent 3 full days on precise model pruning and weight optimization. The result is a clean, sharp, and highly dynamic video generation model.
β¨ Key Improvements & Features
- π¬ HDR Image Quality & Blur Reduction: Fine-tuned on high-dynamic-range (HDR) video datasets to significantly enhance visual clarity and eliminate motion blur during high-speed action.
- π€ Distant Face Restoration: Drastically reduces facial distortion, blurriness, and collapsing in medium-to-long shots.
- π¨ Clean & De-Oiled Aesthetic: Removes heavy, unnatural skin shine and glossy textures, rendering natural lighting and photorealistic materials.
- βοΈ Enhanced Dynamic Motion: Boosts motion fluidity and physical impact, excels in complex action sequences such as sword fighting and martial arts/melee combat.
- π VFX & Fantasy Effects: Specifically optimized for fantasy spellcasting, particle aura, and magical combat visual effects.
- π Expressive Facial Dynamics: Captures subtle facial expressions and emotional nuances more vividly.
- πΉ Cinematography & Camera Control: Strengthens responsiveness to camera movements (pan, tilt, zoom, tracking shots) for cinematic storytelling.
- π‘οΈ Full Base Capability Retention: 100% preserves MiniMax-H3's original prompt adherence, style adaptability, and base multimodal generation strength.
π¬ Showcase
π‘ Usage Guide
Multimodal Pipeline Support
This model is fully compatible with ComfyUI and supports:
- Text-to-Video (T2V)
- Image-to-Video (I2V)
- Reference-to-Video (Ref2V)
- Video-to-Video (V2V)
π Recommended Acceleration LoRA
For high-speed generation with minimal quality loss, we strongly recommend pairing with:
minimax_h3_ref2v_turbo_4step_v0.1(Enables 4-step fast inference)
π Online Interactive Demo
Test the model directly in your browser without local GPU setup: π Try it on RunningHub Workflows
π Acknowledgements
Special thanks to the MiniMax open-source team for creating and releasing the powerful MiniMax-H3 multimodal video model, providing a solid foundation for the open-source community! π€
π€ Community & Commercial Inquiries
Feel free to connect for tutorials, community discussions, workflow sharing, or commercial collaborations:
- YouTube Channel: AIGC-Singularity
- Bilibili Channel: AIGC-Singularity Space
- QQ Group 1:
1058747239(Request to join) - QQ Group 2:
1072010342(Request to join) - Business Inquiries (WeChat):
aigctyd - Email:
a592991299@gmail.com
Sign up to read complete case studies, access detailed metrics, and unlock all use cases.
Sign up to read complete case studies, access detailed metrics, and unlock all use cases.