Collections
Discover the best community collections!
Collections including paper arxiv:2404.00987
-
3D Congealing: 3D-Aware Image Alignment in the Wild
Paper • 2404.02125 • Published • 10 -
FlexiDreamer: Single Image-to-3D Generation with FlexiCubes
Paper • 2404.00987 • Published • 22 -
ST-LLM: Large Language Models Are Effective Temporal Learners
Paper • 2404.00308 • Published • 8 -
InstantSplat: Unbounded Sparse-view Pose-free Gaussian Splatting in 40 Seconds
Paper • 2403.20309 • Published • 19
-
Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction
Paper • 2404.02905 • Published • 69 -
InstantStyle: Free Lunch towards Style-Preserving in Text-to-Image Generation
Paper • 2404.02733 • Published • 22 -
Cross-Attention Makes Inference Cumbersome in Text-to-Image Diffusion Models
Paper • 2404.02747 • Published • 13 -
Bigger is not Always Better: Scaling Properties of Latent Diffusion Models
Paper • 2404.01367 • Published • 22
-
Condition-Aware Neural Network for Controlled Image Generation
Paper • 2404.01143 • Published • 13 -
FlexiDreamer: Single Image-to-3D Generation with FlexiCubes
Paper • 2404.00987 • Published • 22 -
Advancing LLM Reasoning Generalists with Preference Trees
Paper • 2404.02078 • Published • 45 -
ChatGLM-Math: Improving Math Problem-Solving in Large Language Models with a Self-Critique Pipeline
Paper • 2404.02893 • Published • 22
-
ThemeStation: Generating Theme-Aware 3D Assets from Few Exemplars
Paper • 2403.15383 • Published • 15 -
FlexiDreamer: Single Image-to-3D Generation with FlexiCubes
Paper • 2404.00987 • Published • 22 -
MegaScale: Scaling Large Language Model Training to More Than 10,000 GPUs
Paper • 2402.15627 • Published • 37 -
Interactive3D: Create What You Want by Interactive 3D Generation
Paper • 2404.16510 • Published • 20