CVJul 1

Ink3D: Sculpting 3D Assets with Extremely Complex Textures via Video Generative Models

arXiv:2607.0122220.6ECCV
Predicted impact top 7% in CV · last 90 daysOriginality Incremental advance
AI Analysis

It addresses the problem of generating complex textures for 3D assets, which is a bottleneck in 3D content creation due to limited training data.

Ink3D decouples geometry and texture synthesis, using a video generative model to produce dense orbit-scan videos from which it bakes coherent textures onto 3D geometry, enabling significantly richer and more faithful texture generation than prior approaches.

Recent 3D generative models can synthesize high-quality geometry but often struggle to reproduce intricate textures from reference images, largely due to the scarcity of large-scale 3D training data with rich surface appearance. In contrast, visual generative models are trained on datasets several orders of magnitude larger and excel at modeling complex visual patterns. Motivated by this gap, we introduce Ink3D, a framework that bridges 3D generation with large-scale video generative models to synthesize extremely complex textures. Ink3D first reconstructs a white-mesh geometry using an off-the-shelf 3D generation model. It then employs OrbitPainter, a conditional video generative model, to produce dense orbit-scan videos capturing object appearance across viewpoints. To convert these views into coherent textures, we introduce TextureOptimizer, a neural baking module that integrates dense multi-view observations while mitigating geometry inconsistencies arising from video generation. By decoupling geometry and texture synthesis and leveraging large-scale pretrained video priors, Ink3D enables significantly richer and more faithful texture generation than prior approaches.

Foundations

The foundational work for this paper's niche, ranked by how specifically the neighbourhood builds on it — not by global fame.

Your Notes