中文
相关论文

相关论文: AP19-OLR Challenge: Three Tasks and Their Baseline…

200 篇论文

Online reinforcement learning (RL) methods are often data-inefficient or unreliable, making them difficult to train on real robotic hardware, especially quadruped robots. Learning robotic tasks from pre-collected data is a promising…

机器人学 · 计算机科学 2024-10-28 Hongyin Zhang , Shuyu Yang , Donglin Wang

Offline reinforcement learning (RL) is challenged by the distributional shift problem. To address this problem, existing works mainly focus on designing sophisticated policy constraints between the learned policy and the behavior policy.…

机器学习 · 计算机科学 2025-01-09 Yang Yue , Bingyi Kang , Xiao Ma , Qisen Yang , Gao Huang , Shiji Song , Shuicheng Yan

The LEAP submission for DIHARD-III challenge is described in this paper. The proposed system is composed of a speech bandwidth classifier, and diarization systems fine-tuned for narrowband and wideband speech separately. We use an…

音频与语音处理 · 电气工程与系统科学 2021-06-15 Prachi Singh , Rajat Varma , Venkat Krishnamohan , Srikanth Raj Chetupalli , Sriram Ganapathy

Overlapped speech detection (OSD) is critical for speech applications in scenario of multi-party conversion. Despite numerous research efforts and progresses, comparing with speech activity detection (VAD), OSD remains an open challenge and…

声音 · 计算机科学 2022-09-27 Ziqing Du , Kai Liu , Xucheng Wan , Huan Zhou

The NIST Speaker Recognition Evaluation - Conversational Telephone Speech (CTS) challenge 2019 was an open evaluation for the task of speaker verification in challenging conditions. In this paper, we provide a detailed account of the LEAP…

音频与语音处理 · 电气工程与系统科学 2020-05-26 Shreyas Ramoji , Prashant Krishnan , Bhargavram Mysore , Prachi Singh , Sriram Ganapathy

This paper describes our participation in the 2022 TREC NeuCLIR challenge. We submitted runs to two out of the three languages (Farsi and Russian), with a focus on first-stage rankers and comparing mono-lingual strategies to Adhoc ones. For…

信息检索 · 计算机科学 2023-03-21 Carlos Lassance , Stéphane Clinchant

Three years ago, we released the Omniglot dataset for one-shot learning, along with five challenge tasks and a computational model that addresses these tasks. The model was not meant to be the final word on Omniglot; we hoped that the…

人工智能 · 计算机科学 2019-06-04 Brenden M. Lake , Ruslan Salakhutdinov , Joshua B. Tenenbaum

The paradigm of large language model (LLM) reasoning is shifting from parameter scaling to test-time compute scaling, yet many existing approaches still rely on uniform brute-force sampling (for example, fixed best-of-N or self-consistency)…

人工智能 · 计算机科学 2026-03-02 Siyuan Ma , Bo Gao , Xiaojun Jia , Simeng Qin , Tianlin Li , Ke Ma , Xiaoshuang Jia , Wenqi Ren , Yang Liu

Search agents, which integrate language models (LMs) with web search, are becoming crucial for answering complex user queries. Constructing training datasets for deep research tasks, involving multi-step retrieval and reasoning, remains…

计算与语言 · 计算机科学 2026-04-03 Nandan Thakur , Zijian Chen , Xueguang Ma , Jimmy Lin

The "VOiCES from a Distance Challenge 2019" is designed to foster research in the area of speaker recognition and automatic speech recognition (ASR) with the special focus on single channel distant/far-field audio, under noisy conditions.…

音频与语音处理 · 电气工程与系统科学 2019-03-01 Mahesh Kumar Nandwana , Julien van Hout , Mitchell McLaren , Colleen Richey , Aaron Lawson , Maria Alejandra Barrios

We investigate the capabilities and scalability of Large Language Models (LLMs) in optimization modeling, a domain requiring structured reasoning and precise formulation. To this end, we introduce OPT-ENGINE, an extensible benchmark…

计算与语言 · 计算机科学 2026-05-15 Yitian Chen , Cheng Cheng , Yinan Sun , Zi Ling , Dongdong Ge

This paper presents the NTIRE 2026 image super-resolution ($\times$4) challenge, one of the associated competitions of the NTIRE 2026 Workshop at CVPR 2026. The challenge aims to reconstruct high-resolution (HR) images from low-resolution…

计算机视觉与模式识别 · 计算机科学 2026-04-17 Zheng Chen , Kai Liu , Jingkai Wang , Xianglong Yan , Jianze Li , Ziqing Zhang , Jue Gong , Jiatong Li , Lei Sun , Xiaoyang Liu , Radu Timofte , Yulun Zhang , Jihye Park , Yoonjin Im , Hyungju Chun , Hyunhee Park , MinKyu Park , Zheng Xie , Xiangyu Kong , Weijun Yuan , Zhan Li , Qiurong Song , Luen Zhu , Fengkai Zhang , Xinzhe Zhu , Junyang Chen , Congyu Wang , Yixin Yang , Zhaorun Zhou , Jiangxin Dong , Jinshan Pan , Shengwei Wang , Jiajie Ou , Baiang Li , Sizhuo Ma , Qiang Gao , Jusheng Zhang , Jian Wang , Keze Wang , Yijiao Liu , Yingsi Chen , Hui Li , Yu Wang , Congchao Zhu , Saeed Ahmad , Ik Hyun Lee , Jun Young Park , Ji Hwan Yoon , Kainan Yan , Zian Wang , Weibo Wang , Shihao Zou , Chao Dong , Wei Zhou , Linfeng Li , Jaeseong Lee , Jaeho Chae , Jinwoo Kim , Seonjoo Kim , Yucong Hong , Zhenming Yan , Junye Chen , Ruize Han , Song Wang , Yuxuan Jiang , Chengxi Zeng , Tianhao Peng , Fan Zhang , David Bull , Tongyao Mu , Qiong Cao , Yifan Wang , Youwei Pan , Leilei Cao , Xiaoping Peng , Wei Deng , Yifei Chen , Wenbo Xiong , Xian Hu , Yuxin Zhang , Xiaoyun Cheng , Yang Ji , Zonghao Chen , Zhihao Xue , Junqin Hu , Nihal Kumar , Snehal Singh Tomar , Klaus Mueller , Surya Vashisth , Prateek Shaily , Jayant Kumar , Hardik Sharma , Ashish Negi , Sachin Chaudhary , Akshay Dudhane , Praful Hambarde , Amit Shukla , Shijun Shi , Jiangning Zhang , Yong Liu , Kai Hu , Jing Xu , Xianfang Zeng , Amitesh M , Hariharan S , Chia-Ming Lee , Yu-Fan Lin , Chih-Chung Hsu , Nishalini K , Sreenath K A , Bilel Benjdira , Anas M. Ali , Wadii Boulila , Shuling Zheng , Zhiheng Fu , Feng Zhang , Zhanglu Chen , Boyang Yao , Nikhil Pathak , Aagam Jain , Milan Kumar , Kishor Upla , Vivek Chavda , Sarang N S , Raghavendra Ramachandra , Zhipeng Zhang , Qi Wang , Shiyu Wang , Jiachen Tu , Guoyi Xu , Yaoxin Jiang , Jiajia Liu , Yaokun Shi , Yuqi Li , Chuanguang Yang , Weilun Feng , Zhuzhi Hong , Hao Wu , Junming Liu , Yingli Tian , Amish Bhushan Kulkarni , Tejas R R Shet , Saakshi M Vernekar , Nikhil Akalwadi , Kaushik Mallibhat , Ramesh Ashok Tabib , Uma Mudenagudi , Yuwen Pan , Tianrun Chen , Deyi Ji , Qi Zhu , Lanyun Zhu , Heyan Zhangyi

This paper introduces an innovative Applicant Tracking System (ATS) enhanced by a novel Robotic process automation (RPA) framework or as further referred to as MLAR. Traditional recruitment processes often encounter bottlenecks in resume…

计算与语言 · 计算机科学 2025-07-15 Mohamed T. Younes , Omar Walid , Mai Hassan , Ali Hamdi

Large language models increasingly operate in interactive settings where solving a task requires multiple rounds of information exchange with a user. However, most current systems treat dialogue reactively and lack a principled mechanism to…

人工智能 · 计算机科学 2026-05-08 Aymen Echarghaoui , Dongxia Wu , Emily B. Fox

ASR has achieved remarkable global progress, yet African low-resource languages remain rigorously underrepresented, producing barriers to digital inclusion across the continent with more than +2000 languages. This systematic literature…

The paper announces the new long-term challenge for improving the performance of automatic speech recognition systems. The goal of the challenge is to investigate methods of correcting the recognition results on the basis of previously made…

计算与语言 · 计算机科学 2020-01-10 Marek Kubis , Zygmunt Vetulani , Mikołaj Wypych , Tomasz Ziętkiewicz

The SdSv challenge Task 2 provided an opportunity to assess efficiency and robustness of modern text-independent speaker verification systems. But it also made it possible to test new approaches, capable of taking into account the main…

声音 · 计算机科学 2024-03-29 Pierre-Michel Bousquet , Mickael Rouvier

Resolution of complex SQL issues persists as a significant bottleneck in real-world database applications. Current Large Language Models (LLMs), while adept at text-to-SQL translation, have not been rigorously evaluated on the more…

Cross-lingual information retrieval (CLIR) addresses the challenge of retrieving relevant documents written in languages different from that of the original query. Research in this area has typically framed the task as monolingual retrieval…

信息检索 · 计算机科学 2025-10-02 Roksana Goworek , Olivia Macmillan-Scott , Eda B. Özyiğit

This paper reports the LEAP submission to the CHiME-6 challenge. The CHiME-6 Automatic Speech Recognition (ASR) challenge Track 1 involved the recognition of speech in noisy and reverberant acoustic conditions in home environments with…

音频与语音处理 · 电气工程与系统科学 2020-05-25 Anirudh Sreeram , Anurenjan Purushothaman , Rohit Kumar , Sriram Ganapathy