跳到主要导航 跳到搜索 跳到主要内容

3D mesh colorization from a single image via geometry prior modulation

  • Rama Bastola Neupane
  • , Kan Li*
  • , Zhuqing Mao
  • *此作品的通讯作者
  • Beijing Institute of Technology
  • Tribhuvan University

科研成果: 期刊稿件文章同行评审

摘要

We propose a novel framework for colorizing 3D meshes from a single RGB image, utilizing a triplane-based representation that integrates both geometric and image features. Unlike traditional texture mapping or view-dependent neural rendering approaches, our method directly predicts per-vertex colors without requiring camera pose information. To capture geometric context, we extract features from an uncolored mesh using a point-based encoder and project them onto three orthogonal planes, aligning them with the image space. Simultaneously, semantic features are extracted from the input image using a vision transformer. These image features are decoded into a triplane representation using a transformer-based decoder, where mesh features modulate the attention and feed-forward mechanisms, enriching the representation with geometric and appearance cues. Each mesh vertex then samples the refined triplane via bilinear interpolation to obtain a descriptive feature, which is decoded into a view-independent RGB color. The model is trained using a combination of 2D photometric loss computed from renderings of the predicted and ground-truth colored meshes, and a 3D vertex color loss. At inference, the method operates using a single RGB image and an uncolored mesh, without requiring camera pose, generating a colored mesh in under one second. Experiments on standard benchmarks demonstrate that our approach produces high-quality and consistent per-vertex colorization, outperforming existing single-view methods in both visual fidelity and generalization.

源语言英语
文章编号114578
期刊Knowledge-Based Systems
330
DOI
出版状态已出版 - 25 11月 2025
已对外发布

指纹

探究 '3D mesh colorization from a single image via geometry prior modulation' 的科研主题。它们共同构成独一无二的指纹。

引用此