text-to-image
160 个项目 · ⭐ 165.7kOfficial implementation of OneDiffusion paper (CVPR 2025)
[ICCV 2023] A latent space for stochastic diffusion models
(Accepted by IJCV) Liquid: Language Models are Scalable and Unified Multi-modal Generators
Official implementation for "Blended Latent Diffusion" [SIGGRAPH 2023]
[CVPR 2024 Highlight] MIGC and [TPAMI 2024] MIGC++ (Official Implementation)
Official implementation for "Blended Diffusion for Text-driven Editing of Natural Images" [CVPR 2022]
Official code for the CVPR 2025 paper "SemanticDraw: Towards Real-Time Interactive Content Creation from Image Diffusion Models."
Multimodal AI Story Teller, built with Stable Diffusion, GPT, and neural text-to-speech
[NeurIPS 2025 Spotlight] A Unified Tokenizer for Visual Generation and Understanding
Official implementation for "Break-A-Scene: Extracting Multiple Concepts from a Single Image" [SIGGRAPH Asia 2023]
Official implementation of AsymFlow, pi-Flow, GMFlow
[ ICLR 2024 ] Official Codebase for "InstructCV: Instruction-Tuned Text-to-Image Diffusion Models as Vision Generalists"
共 160 条 · 第 3 / 8 页