AI Chat Paper
Note: Please note that the following content is generated by AMiner AI. SciOpen does not take any responsibility related to this content.
{{lang === 'zh_CN' ? '文章概述' : 'Summary'}}
{{lang === 'en_US' ? '中' : 'Eng'}}
Chat more with AI
Article Link
Collect
Submit Manuscript
Show Outline
Outline
Show full outline
Hide outline
Outline
Show full outline
Hide outline
Research | Open Access

A general framework for Gaussian Splatting-based human-centric volumetric videos

Shengkun Zhu1 Chengcheng Guo1 Yuanji Lu1 Zhehao Shen1 Yize Wu1 Yu Hong1 Yiwen Cai1 Meihan Zheng1 Yingliang Zhang1,2 Lan Xu1 ( )Jingyi Yu1 ( )
School of Information Science and Technology, ShanghaiTech University, Shanghai, China
DGene, Shanghai, China
Show Author Information

Abstract

Volumetric video is revolutionizing immersive media, with 3D Gaussian Splatting emerging as a key technology due to its unprecedented real-time rendering quality. However, extending this technology to dynamic scenes presents two major challenges: the massive storage and transmission overhead associated with temporal sequences, and a fragmented toolchain ecosystem that hinders efficient research and development. Existing solutions typically focus on isolated stages such as reconstruction or compression, lacking a unified, end-to-end workflow from data acquisition to final viewing. To address these challenges, we propose a comprehensive dynamic Gaussian processing framework that provides a complete, end-to-end pipeline. This framework systematically integrates the entire process, from data acquisition and standardized preprocessing to a suite of diverse dynamic Gaussian reconstruction algorithms. One of its core contributions is a general-purpose compression framework, compatible with the outputs of various reconstruction methods, which significantly reduces the storage footprint of dynamic sequences while maintaining high visual fidelity. To ensure broad accessibility, we have also developed a cross-platform rendering solution that supports high-quality, interactive free-viewpoint experiences on desktop, mobile, and XR devices. Furthermore, to advance the field, we contribute a large-scale, high-quality dynamic human performance capture dataset. Captured with a dense 81-camera array, the dataset comprises over 130 sequences of diverse human motions, including complex interactions with topological changes. Our integrated framework and dataset aim to bridge the entire pipeline from data creation to end-user application, providing a solid foundation for the large-scale adoption and future research of Gaussian Splatting technology.

References

【1】
【1】
 
 
Visual Intelligence
Article number: 8

{{item.num}}

Comments on this article

Go to comment

< Back to all reports

Review Status: {{reviewData.commendedNum}} Commended , {{reviewData.revisionRequiredNum}} Revision Required , {{reviewData.notCommendedNum}} Not Commended Under Peer Review

Review Comment

Close
Close
Cite this article:
Zhu S, Guo C, Lu Y, et al. A general framework for Gaussian Splatting-based human-centric volumetric videos. Visual Intelligence, 2026, 4: 8. https://doi.org/10.1007/s44267-026-00111-7

897

Views

0

Crossref

Received: 03 September 2025
Revised: 12 February 2026
Accepted: 16 February 2026
Published: 01 April 2026
© The Author(s) 2026.

This article is licensed under a Creative Commons Attribution 4.0 International License, which permits use, sharing, adaptation, distribution and reproduction in any medium or format, as long as you give appropriate credit to the original author(s) and the source, provide a link to the Creative Commons licence, and indicate if changes were made. The images or other third party material in this article are included in the article’s Creative Commons licence, unless indicated otherwise in a credit line to the material. If material is not included in the article’s Creative Commons licence and your intended use is not permitted by statutory regulation or exceeds the permitted use, you will need to obtain permission directly from the copyright holder. To view a copy of this licence, visit http://creativecommons.org/licenses/by/4.0/.