English
Related papers

Related papers: C-RADIOv4 (Tech Report)

200 papers

4D radar super-resolution, which aims to reconstruct sparse and noisy point clouds into dense and geometrically consistent representations, is a foundational problem in autonomous perception. However, existing methods often suffer from high…

Computer Vision and Pattern Recognition · Computer Science 2025-09-17 Minqing Huang , Shouyi Lu , Boyuan Zheng , Ziyao Li , Xiao Tang , Guirong Zhuo

We present Seedream 3.0, a high-performance Chinese-English bilingual image generation foundation model. We develop several technical improvements to address existing challenges in Seedream 2.0, including alignment with complicated prompts,…

In recent years, deep learning has attracted increasing attention in the field of Cardiac MRI (CMR) reconstruction due to its superior performance over traditional methods, particularly in handling higher acceleration factors, highlighting…

Computer Vision and Pattern Recognition · Computer Science 2026-01-09 Donghang Lyu , Marius Staring , Hildo Lamb , Mariya Doneva

Given the incomplete sampling of spatial frequencies by radio interferometers, achieving precise restoration of astrophysical information remains challenging. To address this ill-posed problem, compressive sensing(CS) provides a robust…

Instrumentation and Methods for Astrophysics · Physics 2025-05-09 Lei Yu , Bin Liu , Cheng-Jin Jin , Ru-Rong Chen , Hong-Wei Xi , Bo Peng

In future 6G cellular networks, a joint communication and sensing protocol will allow the network to perceive the environment, opening the door for many new applications atop a unified communication-perception infrastructure. However,…

Networking and Internet Architecture · Computer Science 2022-02-25 Mohammed Alloulah , Akash Deep Singh , Maximilian Arnold

The C-V2X or LTE-V standard has been designed to support V2X (Vehicle to Everything) communications. The standard is an evolution of LTE, and it has been published by the 3GPP in Release 14. This new standard introduces the C-V2X or LTE-V…

Networking and Internet Architecture · Computer Science 2019-04-22 Manuel Gonzalez-Martin , Miguel Sepulcre , Rafael Molina-Masegosa , Javier Gozalvez

In this study, we dive deep into the inconsistency of pseudo targets in semi-supervised object detection (SSOD). Our core observation is that the oscillating pseudo-targets undermine the training of an accurate detector. It injects noise…

Computer Vision and Pattern Recognition · Computer Science 2023-03-29 Xinjiang Wang , Xingyi Yang , Shilong Zhang , Yijiang Li , Litong Feng , Shijie Fang , Chengqi Lyu , Kai Chen , Wayne Zhang

In this work, we present VARGPT-v1.1, an advanced unified visual autoregressive model that builds upon our previous framework VARGPT. The model preserves the dual paradigm of next-token prediction for visual understanding and next-scale…

Computer Vision and Pattern Recognition · Computer Science 2025-04-07 Xianwei Zhuang , Yuxin Xie , Yufan Deng , Dongchao Yang , Liming Liang , Jinghan Ru , Yuguo Yin , Yuexian Zou

Voice activity detection (VAD) is essential for speech-driven applications, but remains far from perfect in noisy and resource-limited environments. Existing methods often lack robustness to noise, and their frame-wise classification losses…

Sound · Computer Science 2025-08-29 Chien-Chun Wang , En-Lun Yu , Jeih-Weih Hung , Shih-Chieh Huang , Berlin Chen

We introduce a novel co-learning paradigm for manifolds naturally equipped with a group action, motivated by recent developments on learning a manifold from attached fibre bundle structures. We utilize a representation theoretic mechanism…

Machine Learning · Computer Science 2019-12-10 Yifeng Fan , Tingran Gao , Zhizhen Zhao

Classroom environments are particularly challenging for children with hearing impairments, where background noise, multiple talkers, and reverberation degrade speech perception. These difficulties are greater for children than adults, yet…

The emergence of multi-modal deep learning models has made significant impacts on clinical applications in the last decade. However, the majority of models are limited to single-tasking, without considering disease diagnosis is indeed a…

Computer Vision and Pattern Recognition · Computer Science 2024-03-05 Lijian Xu , Ziyu Ni , Xinglong Liu , Xiaosong Wang , Hongsheng Li , Shaoting Zhang

Answer verification methods are widely employed in language model training pipelines spanning data curation, evaluation, and reinforcement learning with verifiable rewards (RLVR). While prior work focus on developing unified verifiers…

Machine Learning · Computer Science 2025-12-02 Ruixiang Feng , Zhenwei An , Yuntao Wen , Ran Le , Yiming Jia , Chen Yang , Zongchao Chen , Lisi Chen , Shen Gao , Shuo Shang , Yang Song , Tao Zhang

As one of the most important underwater sensing technologies, forward-looking sonar exhibits unique imaging characteristics. Sonar images are often affected by severe speckle noise, low texture contrast, acoustic shadows, and geometric…

Computer Vision and Pattern Recognition · Computer Science 2026-05-19 Ping Guo , Chengzhou Li , Guanchen Meng , Qi Jia , Jinyuan Liu , Zhu Liu , Yu Liu , Zhongxuan Luo , Xin Fan

This report describes the UNISOUND submission for Track1 and Track2 of VoxCeleb Speaker Recognition Challenge 2023 (VoxSRC 2023). We submit the same system on Track 1 and Track 2, which is trained with only VoxCeleb2-dev. Large-scale ResNet…

Audio and Speech Processing · Electrical Eng. & Systems 2023-08-25 Yu Zheng , Yajun Zhang , Chuanying Niu , Yibin Zhan , Yanhua Long , Dongxing Xu

This technical report presents the training methodology and evaluation results of the open-source Jasper-Token-Compression-600M model, released in November 2025. Building on previous distillation-based recipes from the English Stella and…

Information Retrieval · Computer Science 2025-11-20 Dun Zhang , Ziyang Zeng , Yudong Zhou , Shuyang Lu

We present a compact, quantization-ready acoustic scene classification (ASC) framework that couples an efficient student network with a learned teacher ensemble and knowledge distillation. The student backbone uses stacked…

While the deep learning techniques promote the rapid development of the speech enhancement (SE) community, most schemes only pursue the performance in a black-box manner and lack adequate model interpretability. Inspired by Taylor's…

Sound · Computer Science 2022-05-03 Andong Li , Shan You , Guochen Yu , Chengshi Zheng , Xiaodong Li

Recently, deep learning has been proposed as a potential technique for improving the physical layer performance of radio receivers. Despite the large amount of encouraging results, most works have not considered spatial multiplexing in the…

Signal Processing · Electrical Eng. & Systems 2020-11-02 Dani Korpi , Mikko Honkala , Janne M. J. Huttunen , Vesa Starck
‹ Prev 1 8 9 10 Next ›