PaddlePaddle/PaddleMIX
PaddleMIX is a PaddlePaddle-based multimodal AI framework providing diffusion models and vision-language models for image, video, and text generation tasks.

Not currently ranked — collecting fresh signals.
star history
PaddleMIX offers a comprehensive multimodal model library including end-to-end large-scale vision-language models (LLaVA, Qwen2-VL, DeepSeek-VL, InternVL2, MiniCPM-V) and a diffusion model toolbox (ppdiffusers) for text-to-image, text-to-video, and image-to-text generation. It provides full-pipeline development tools and high-performance distributed training capabilities for multimodal AI tasks.
Frequently asked
- What is PaddlePaddle/PaddleMIX?
- PaddleMIX is a PaddlePaddle-based multimodal AI framework providing diffusion models and vision-language models for image, video, and text generation tasks.
- Is PaddleMIX open source?
- Yes — PaddlePaddle/PaddleMIX is open source, released under the Apache-2.0 license.
- What language is PaddleMIX written in?
- PaddlePaddle/PaddleMIX is primarily written in Python.
- How popular is PaddleMIX?
- PaddlePaddle/PaddleMIX has 724 stars on GitHub.
- Where can I find PaddleMIX?
- PaddlePaddle/PaddleMIX is on GitHub at https://github.com/PaddlePaddle/PaddleMIX.