摘要
360-degree video has shown great potential to the mainstream since its immersive experience. However, 360-degree video streaming requires ultrahigh bandwidth and low latency, which limit the improvement of user quality of experience (QoE). Currently, methods combining field of view (FoV) prediction and adaptive video streaming provide an effective method for addressing the above issues. However, existing FoV prediction methods based on recurrent neural networks (RNN) cannot capture long-range dependency from input to output. Current deep reinforcement learning (DRL)-based adaptive strategies fail to estimate the future bandwidth with high accuracy and fully explore the capability of VR devices. To ameliorate these limitations, we design a DRL-based 360-degree video streaming method named VRFormer with FoV combined prediction and super resolution (SR). First, we adopt a content-aware transformer-based encoder-decoder network to make the long-term FoV prediction. It combines the user's head movement history, eye-tracking history, and user attention extracted from a convolutional neural network (CNN)-based network. Second, we introduce a DNN-based SR network running on VR devices to reconstruct high-definition video content. Finally, we apply a DRL-based network to adaptively allocate rates for future tiles and dynamically control video content reconstruction. Experiments have verified that the proposed method can effectively improve the quality of experience (QoE) of the user's viewing experience compared to the state-of-the-art methods.
| 源语言 | 英语 |
|---|---|
| 主期刊名 | Proceedings - 20th IEEE International Symposium on Parallel and Distributed Processing with Applications, 12th IEEE International Conference on Big Data and Cloud Computing, 12th IEEE International Conference on Sustainable Computing and Communications and 15th IEEE International Conference on Social Computing and Networking, ISPA/BDCloud/SocialCom/SustainCom 2022 |
| 出版商 | Institute of Electrical and Electronics Engineers Inc. |
| 页 | 531-538 |
| 页数 | 8 |
| ISBN(电子版) | 9781665464970 |
| DOI | |
| 出版状态 | 已出版 - 2022 |
| 活动 | 20th IEEE International Symposium on Parallel and Distributed Processing with Applications, 12th IEEE International Conference on Big Data and Cloud Computing, 12th IEEE International Conference on Sustainable Computing and Communications and 15th IEEE International Conference on Social Computing and Networking, ISPA/BDCloud/SocialCom/SustainCom 2022 - Melbourne, 澳大利亚 期限: 17 12月 2022 → 19 12月 2022 |
出版系列
| 姓名 | Proceedings - 20th IEEE International Symposium on Parallel and Distributed Processing with Applications, 12th IEEE International Conference on Big Data and Cloud Computing, 12th IEEE International Conference on Sustainable Computing and Communications and 15th IEEE International Conference on Social Computing and Networking, ISPA/BDCloud/SocialCom/SustainCom 2022 |
|---|
会议
| 会议 | 20th IEEE International Symposium on Parallel and Distributed Processing with Applications, 12th IEEE International Conference on Big Data and Cloud Computing, 12th IEEE International Conference on Sustainable Computing and Communications and 15th IEEE International Conference on Social Computing and Networking, ISPA/BDCloud/SocialCom/SustainCom 2022 |
|---|---|
| 国家/地区 | 澳大利亚 |
| 市 | Melbourne |
| 时期 | 17/12/22 → 19/12/22 |
联合国可持续发展目标
此成果有助于实现下列可持续发展目标:
-
可持续发展目标 7 经济适用的清洁能源
学术指纹
探究 'VRFormer: 360-Degree Video Streaming with FoV Combined Prediction and Super resolution' 的科研主题。它们共同构成独一无二的指纹。引用此
- APA
- Author
- BIBTEX
- Harvard
- Standard
- RIS
- Vancouver