-
Compose and Conquer: Diffusion-Based 3D Depth Aware Composable Image Synthesis
Paper • 2401.09048 • Published • 10 -
Improving fine-grained understanding in image-text pre-training
Paper • 2401.09865 • Published • 17 -
Depth Anything: Unleashing the Power of Large-Scale Unlabeled Data
Paper • 2401.10891 • Published • 60 -
Scaling Up to Excellence: Practicing Model Scaling for Photo-Realistic Image Restoration In the Wild
Paper • 2401.13627 • Published • 74
Collections
Discover the best community collections!
Collections including paper arxiv:2501.08332
-
Go-with-the-Flow: Motion-Controllable Video Diffusion Models Using Real-Time Warped Noise
Paper • 2501.08331 • Published • 20 -
MangaNinja: Line Art Colorization with Precise Reference Following
Paper • 2501.08332 • Published • 55 -
GameFactory: Creating New Games with Generative Interactive Videos
Paper • 2501.08325 • Published • 60 -
DiffuEraser: A Diffusion Model for Video Inpainting
Paper • 2501.10018 • Published • 13
-
Neural LightRig: Unlocking Accurate Object Normal and Material Estimation with Multi-Light Diffusion
Paper • 2412.09593 • Published • 18 -
CLEAR: Conv-Like Linearization Revs Pre-Trained Diffusion Transformers Up
Paper • 2412.16112 • Published • 21 -
Thinking in Space: How Multimodal Large Language Models See, Remember, and Recall Spaces
Paper • 2412.14171 • Published • 24 -
DiffSensei: Bridging Multi-Modal LLMs and Diffusion Models for Customized Manga Generation
Paper • 2412.07589 • Published • 45
-
Cinemo: Consistent and Controllable Image Animation with Motion Diffusion Models
Paper • 2407.15642 • Published • 11 -
ashawkey/LGM
Text-to-3D • Updated • 117 -
Cycle3D: High-quality and Consistent Image-to-3D Generation via Generation-Reconstruction Cycle
Paper • 2407.19548 • Published • 25 -
DreamCinema: Cinematic Transfer with Free Camera and 3D Character
Paper • 2408.12601 • Published • 30
-
Magic Insert: Style-Aware Drag-and-Drop
Paper • 2407.02489 • Published • 20 -
ZePo: Zero-Shot Portrait Stylization with Faster Sampling
Paper • 2408.05492 • Published • 7 -
CSGO: Content-Style Composition in Text-to-Image Generation
Paper • 2408.16766 • Published • 18 -
Style-Friendly SNR Sampler for Style-Driven Generation
Paper • 2411.14793 • Published • 36