arrow
返回

VidQ: Video Query Using Optimized Audio-Visual Processing

delete2023-06-01
delete0
PRE
AI
N
Noor Felemban *
F
Fidan Mehmeti
P
Porta, Thomas F.
DOI:10.1109/TNET.2022.3215601delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
As mobile devices become more prevalent in everyday life and the amount of recorded and stored videos increases, efficient techniques for searching video content become more important. When a user sends a query searching for a specific action in a large amount of data, the goal is to respond to the query accurately and fast. In this paper, we address the problem of responding to queries which search for specific actions in mobile devices in a timely manner by utilizing both visual and audio processing approaches. We build a system, called VidQ, which consists of several stages, and that uses various Convolutional Neural Networks (CNNs) and Speech APIs to respond to such queries. As the state-of-the-art computer vision and speech algorithms are computationally intensive, we use servers with GPUs to assist mobile users in the process. After a query is issued, we identify the different stages of processing that will take place. Then, we identify the order of these stages. Finally, solving an optimization problem that captures the system behavior, we distribute the process among the available network resources to minimize the processing time. Results show that VidQ reduces the completion time by at least 50% compared to other approaches.
Keyword:
Mobile networks
deep learning
convolutional neural networks
performance optimization
heuristics

期刊

I
IEEE-ACM Transactions on Networking
IF:
3.6
论文数:
4.4K
被引数:
9.5K

机构

I
Imam Abdulrahman Bin Faisal University
学者数:
6.3K
论文数: 4.5K
被引数: 5.4K
P
pennsylvania commonwealth system of higher education (pcshe)
学者数:
12.9W
论文数: 11.7W
被引数: 177
T
Technical University of Munich
学者数:
5.2W
论文数: 3.9W
被引数: 6.2W
学者 查看更多机构