arrow
返回

Exploring Spherical Autoencoder for Spherical Video Content Processing

delete2022-10-10
delete1
delete
OA
AI
J
Jin Zhou *
N
Na Li
Y
Yao Liu
S
Shuochao Yao
S
Songqing Chen
DOI:10.1145/3503161.3548364delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
3D spherical content is increasingly presented in various applications (e.g., AR/MR/VR) for better users' immersiveness experience, yet today processing such spherical 3D content still mainly relies on the traditional 2D approaches after projection, leading to the distortion and/or loss of critical information. This study sets to explore methods to process spherical 3D content directly and more effectively. Using 360-degree videos as an example, we propose a novel approach called Spherical Autoencoder (SAE) for spherical video processing. Instead of projecting to a 2D space, SAE represents the 360-degree video content as a spherical object and employs encoding and decoding on the 360-degree video directly. Furthermore, to support the adoption of SAE on pervasive mobile devices that often have resource constraints, we further propose two optimizations on top of SAE. First, since the FoV (Field of View) prediction is widely studied and leveraged to transport only a portion of the content to the mobile device to save bandwidth and battery consumption, we design p-SAE, a SAE scheme with the partial view support that can utilize such FoV prediction. Second, since machine learning models are often compressed when running on mobile devices in order to reduce the processing load, which usually leads to degradation of output (e.g., video quality in SAE), we propose c-SAE by applying the compressive sensing theory into SAE to maintain the video quality when the model is compressed. Our extensive experiments show that directly incorporating and processing spherical signals is promising, and it outperforms the traditional approaches by a large margin. Both p-SAE and c-SAE show their effectiveness in delivering high quality videos (e.g., PSNR results) when used alone or combined together with model compression.
Keyword:
Spherical Autoencoder
Partial View
Compressive Sensing

期刊

P
PROCEEDINGS OF THE ACM CONFERENCE ON SECURITY AND PRIVACY IN WIRELESS AND MOBILE NETWORKS
IF:
0
论文数:
1.8K
被引数:
0

机构

G
George Mason University
学者数:
7.7K
论文数: 7.9K
被引数: 1.0W
R
rutgers university system
学者数:
4.1W
论文数: 3.7W
被引数: 53
引用论文

引用论文

err分享
err收藏
err分享
err收藏
The Prevalence of Trigeminal Neuralgia in Turkey: A Population-Based Study
err2020-07-14
err0
PREAI
errCem Bölük; Ülkü Türk Börü; Mustafa Taşdemir
err分享
err收藏
Nature-Based Solutions and Climate Change – Four Shades of Green
err2017-09-02
err0
PREAI
errStephan Pauleit; Teresa Zölch; Rieke Hansen; Thomas B. Randrup; Cecil Konijnendijk van den Bosch
err分享
err收藏
Personal identifiability of user tracking data during observation of 360-degree VR video
err2020-10-15
err95
errOAAI
errMiller, Mark Roman; Herrera, Fernanda; Jun, Hanseul; Landay, James A.; Bailenson, Jeremy N.
err分享
err收藏
学者 查看更多内容