Awesome-Unified-Multimodal-Models

📖 This is a repository for organizing papers, codes and other resources related to unified multimodal models.

S

showlab

Dernière activité 10 oct. 2025
showlab/Awesome-Unified-Multimodal-Models

834

étoiles

41

forks

8

issues ouvertes

Ce README est souvent en anglais.

Awesome Unified Multimodal Models Awesome

This is a repository for organizing papers, codes and other resources related to unified multimodal models.

TAX

🤔 What are unified multimodal models?

Traditional multimodal models can be broadly categorized into two types: multimodal understanding and multimodal generation. Unified multimodal models aim to integrate these two tasks within a single framework. Such models are also referred to as Any-to-Any generation in the community. These models operate on the principle of multimodal input and multimodal output, enabling them to process and generate content across various modalities seamlessly.

🔆 This project is still on-going, pull requests are welcomed!!

If you have any suggestions (missing papers, new papers, or typos), please feel free to edit and pull a request. Just letting us know the title of papers can also be a great contribution to us. You can do this by open issue or contact us directly via email.

⭐ If you find this repo useful, please star it!!!

Unified Multimodal Understanding and Generation

Acknowledgements

This template is provided by Awesome-Video-Diffusion and Awesome-MLLM-Hallucination.

Projets similaires

📖 This is a repository for organizing papers, codes, and other resources related to unified multimodal models.

PPurshow
365 étoiles14

Awesome Unified Multimodal Models

multimodal-large-language-modelsmultimodal-modelstext-to-image-generation
AATH-MaaS
1,3 k étoiles48

📖 This is a repository for organizing papers, codes and other resources related to Visual Reinforcement Learning.

Wweijiawu
455 étoiles24