Beyond the V-PCC Reference Model: A Holistic Approach to Low-Complexity Attribute Map Generation
Video-based Point Cloud Compression (V-PCC) efficiently compresses dynamic point clouds by projecting 3D patches into 2D occupancy, geometry, and attribute maps before video encoding. The reference attribute map generation pipeline relies on assumptions inherited from the standardization process, notably that global image smoothing and tighter patch packing improve coding performance. This work systematically reassesses these assumptions under practical encoding constraints using uvgVPCCenc, an open-source V-PCC framework designed for real-time encoding. Rather than optimizing individual map generation stages independently, we adopt a holistic approach that analyzes the interactions between attribute background filling, attribute map conversion, and patch packing. To this end, we examine three straightforward methods: a local block-based background filling algorithm, a local RGB444-to-YUV420 attribute map conversion, and patch spacing. Their impact is evaluated on both uvgVPCCenc and the V-PCC reference encoder TMC2 using coding efficiency, computational complexity, and subjective visual quality. The results show that these low-complexity methods can substantially reduce map generation complexity. The proposed background filling accelerates the overall encoding by up to 1.41 \\(\\times\\) , while the proposed attribute map conversion increases this speedup to 1.50 \\(\\times\\) and eliminates visual artifacts caused by the interaction between background filling and the reference conversion. Furthermore, moderate patch spacing improves reconstructed quality by mitigating inter-patch color bleeding with only a marginal bitrate increase. Overall, our findings demonstrate that several assumptions underlying the reference V-PCC map generation pipeline no longer hold under practical encoding constraints. More importantly, it demonstrates that evaluating map generation stages in isolation can lead to misleading conclusions, advocating a holistic methodology for the design of future practical V-PCC encoders.
Authors
- Jarno Vanne (ORCID: https://orcid.org/0000-0002-7944-1938)
- Guillaume Gautier (ORCID: https://orcid.org/0000-0002-0309-6381)
- Alexandre Mercat (ORCID: https://orcid.org/0000-0003-2211-970X)
- Louis Fréneau (ORCID: https://orcid.org/0009-0002-6289-7741)
Institutions
- Tampere University (FI)
Publication Details
- Journal
- ACM Transactions on Multimedia Computing Communications and Applications
- Published
- 2026-09-09
- DOI
- https://doi.org/10.1145/3846172
- Primary Topic
- 3D Shape Modeling and Analysis
- Type
- article
- Field-Weighted Citation Impact
- 0.00