English
Related papers

Related papers: BiSwift: Bandwidth Orchestrator for Multi-Stream V…

200 papers

Channel estimation is crucial in 5G communication networks for optimizing transmission parameters and ensuring reliable, high-speed communication. However, the use of multiple-input and multiple-output (MIMO) and millimeter-wave (mmWave) in…

Information Theory · Computer Science 2026-05-05 Shengzhe Lyu , Yuhan She , Di Duan , Tao Ni , Yu Hin Chan , Chengwen Luo , Ray C. C. Cheung , Weitao Xu

Multicast/broadcast services (MBS) are able to provide video services for many users simultaneously. Fixed amount of bandwidth allocation for all of the MBS videos is not effective in terms of bandwidth utilization, overall forced call…

Networking and Internet Architecture · Computer Science 2018-12-27 Mostafa Zaman Chowdhury , Tuan Nguyena , Young-Il Kimb , Won Ryub , Yeong Min Jang

Emergency response missions depend on the fast relay of visual information, a task to which unmanned aerial vehicles are well adapted. However, the effective use of unmanned aerial vehicles is often compromised by bandwidth limitations that…

Image and Video Processing · Electrical Eng. & Systems 2026-02-03 Karim El Khoury , Tiffanie Godelaine , Simon Delvaux , Sebastien Lugan , Benoit Macq

We present a novel bi-directional Transformer architecture (BiXT) which scales linearly with input size in terms of computational cost and memory consumption, but does not suffer the drop in performance or limitation to only one input…

Computer Vision and Pattern Recognition · Computer Science 2024-11-01 Markus Hiller , Krista A. Ehinger , Tom Drummond

Autonomous driving requires the inference of actionable information such as detecting and classifying objects, and determining the drivable space. To this end, we present Multi-View LidarNet (MVLidarNet), a two-stage deep neural network for…

Computer Vision and Pattern Recognition · Computer Science 2020-08-19 Ke Chen , Ryan Oldja , Nikolai Smolyanskiy , Stan Birchfield , Alexander Popov , David Wehr , Ibrahim Eden , Joachim Pehserl

Deep neural networks with large model sizes achieve state-of-the-art results for tasks in computer vision (CV) and natural language processing (NLP). However, these large-scale models are too compute- or memory-intensive for…

Distributed, Parallel, and Cluster Computing · Computer Science 2021-10-29 Yang Hu , Connor Imes , Xuanang Zhao , Souvik Kundu , Peter A. Beerel , Stephen P. Crago , John Paul N. Walters

Video service providers need their delivery systems to be able to adapt to network conditions, user preferences, display settings, and other factors. HTTP Adaptive Streaming (HAS) offers dynamic switching between different video…

Image and Video Processing · Electrical Eng. & Systems 2025-12-16 Krishna Srikar Durbha , Alan C. Bovik

The creation of practical deep learning data-products often requires parallelization across processors and computers to make deep learning feasible on large data sets, but bottlenecks in communication bandwidth make it difficult to attain…

Neural and Evolutionary Computing · Computer Science 2016-02-22 Tim Dettmers

Perceiving and understanding 3D motion is a core technology in fields such as autonomous driving, robots, and motion prediction. This paper proposes a 3D motion perception method called ScaleFlow++ that is easy to generalize. With just a…

Computer Vision and Pattern Recognition · Computer Science 2024-10-15 Han Ling , Yinghui Sun , Quansen Sun , Yuhui Zheng

Referring video object segmentation (RVOS) aims to segment the target object in a video sequence described by a language expression. Typical multimodal Transformer based RVOS approaches process video sequence in a frame-independent manner…

Computer Vision and Pattern Recognition · Computer Science 2023-09-19 Meng Lan , Fu Rong , Zuchao Li , Wei Yu , Lefei Zhang

Perceiving and understanding 3D motion is a core technology in fields such as autonomous driving, robots, and motion prediction. This paper proposes a 3D motion perception method called ScaleFlow++ that is easy to generalize. With just a…

Computer Vision and Pattern Recognition · Computer Science 2024-10-17 Han Ling , Quansen Sun

Bi-CamoDiffusion is introduced, an evolution of the CamoDiffusion framework for camouflaged object detection. It integrates edge priors into early-stage embeddings via a parameter-free injection process, which enhances boundary sharpness…

Computer Vision and Pattern Recognition · Computer Science 2026-03-17 Patricia L. Suarez , Leo Thomas Ramos , Angel D. Sappa

Accurate 3D object perception and multi-target multi-camera (MTMC) tracking are fundamental for the digital transformation of industrial infrastructure. However, transitioning "inside-out" autonomous driving models to "outside-in" static…

Computer Vision and Pattern Recognition · Computer Science 2026-01-19 Yizhou Wang , Sameer Pusegaonkar , Yuxing Wang , Anqi Li , Vishal Kumar , Chetan Sethi , Ganapathy Aiyer , Yun He , Kartikay Thakkar , Swapnil Rathi , Bhushan Rupde , Zheng Tang , Sujit Biswas

Binospec is a high-throughput, 370 to 1000 nm, imaging spectrograph that addresses two adjacent 8' by 15' fields of view. Binospec was commissioned in late 2017 at the f/5 focus of the 6.5m MMT and is now available to all MMT observers.…

Instrumentation and Methods for Astrophysics · Physics 2019-06-12 Jan Kansky , Igor Chilingarian , Daniel Fabricant , Anne Matthews , Sean Moran , Martin Paegert , J. Duane Gibson , Dallan Porter , John Roll

The perception system is a a critical role of an autonomous driving system for ensuring safety. The driving scene perception system fundamentally represents an object detection task that requires achieving a balance between accuracy and…

Computer Vision and Pattern Recognition · Computer Science 2025-02-12 Novendra Setyawan , Ghufron Wahyu Kurniawan , Chi-Chia Sun , Wen-Kai Kuo , Jun-Wei Hsieh

Model binarization has made significant progress in enabling real-time and energy-efficient computation for convolutional neural networks (CNN), offering a potential solution to the deployment challenges faced by Vision Transformers (ViTs)…

Computer Vision and Pattern Recognition · Computer Science 2025-03-07 Tian Gao , Zhiyuan Zhang , Yu Zhang , Huajun Liu , Kaijie Yin , Chengzhong Xu , Hui Kong

Current video diffusion models achieve impressive generation quality but struggle in interactive applications due to bidirectional attention dependencies. The generation of a single frame requires the model to process the entire sequence,…

Computer Vision and Pattern Recognition · Computer Science 2025-09-25 Tianwei Yin , Qiang Zhang , Richard Zhang , William T. Freeman , Fredo Durand , Eli Shechtman , Xun Huang

We deal with the problem of streaming multiple video streams between pairs of nodes in a multi-hop wireless ad hoc network. The nodes are static, know their locations, and are synchronized (via GPS). We introduce a new interference model…

Networking and Internet Architecture · Computer Science 2011-08-09 Guy Even , Yaniv Fais , Moti Medina , Shimon , Shahar , Alexander Zadorojniy

The rise of Extended Reality (XR) requires efficient streaming of 3D online worlds, challenging current 3DGS representations to adapt to bandwidth-constrained environments. This paper proposes LapisGS, a layered 3DGS that supports adaptive…

Computer Vision and Pattern Recognition · Computer Science 2025-09-24 Yuang Shi , Géraldine Morin , Simone Gasparini , Wei Tsang Ooi

Extending state-of-the-art object detectors from image to video is challenging. The accuracy of detection suffers from degenerated object appearances in videos, e.g., motion blur, video defocus, rare poses, etc. Existing work attempts to…

Computer Vision and Pattern Recognition · Computer Science 2017-08-21 Xizhou Zhu , Yujie Wang , Jifeng Dai , Lu Yuan , Yichen Wei