Collections
Discover the best community collections!
Collections including paper arxiv:2501.08332
-
Cinemo: Consistent and Controllable Image Animation with Motion Diffusion Models
Paper • 2407.15642 • Published • 11 -
ashawkey/LGM
Text-to-3D • Updated • 116 -
Cycle3D: High-quality and Consistent Image-to-3D Generation via Generation-Reconstruction Cycle
Paper • 2407.19548 • Published • 25 -
DreamCinema: Cinematic Transfer with Free Camera and 3D Character
Paper • 2408.12601 • Published • 29
-
EfficientViT-SAM: Accelerated Segment Anything Model Without Performance Loss
Paper • 2402.05008 • Published • 22 -
PALO: A Polyglot Large Multimodal Model for 5B People
Paper • 2402.14818 • Published • 23 -
MM1: Methods, Analysis & Insights from Multimodal LLM Pre-training
Paper • 2403.09611 • Published • 125 -
InternLM-XComposer2-4KHD: A Pioneering Large Vision-Language Model Handling Resolutions from 336 Pixels to 4K HD
Paper • 2404.06512 • Published • 30
-
Magic Insert: Style-Aware Drag-and-Drop
Paper • 2407.02489 • Published • 20 -
ZePo: Zero-Shot Portrait Stylization with Faster Sampling
Paper • 2408.05492 • Published • 7 -
CSGO: Content-Style Composition in Text-to-Image Generation
Paper • 2408.16766 • Published • 18 -
Style-Friendly SNR Sampler for Style-Driven Generation
Paper • 2411.14793 • Published • 36
-
Compose and Conquer: Diffusion-Based 3D Depth Aware Composable Image Synthesis
Paper • 2401.09048 • Published • 10 -
Improving fine-grained understanding in image-text pre-training
Paper • 2401.09865 • Published • 16 -
Depth Anything: Unleashing the Power of Large-Scale Unlabeled Data
Paper • 2401.10891 • Published • 60 -
Scaling Up to Excellence: Practicing Model Scaling for Photo-Realistic Image Restoration In the Wild
Paper • 2401.13627 • Published • 73