English
Related papers

Related papers: UniUD Submission to the EPIC-Kitchens-100 Multi-In…

200 papers

Large Multi-modal Models (LMMs) have significantly advanced a variety of vision-language tasks. The scalability and availability of high-quality training data play a pivotal role in the success of LMMs. In the realm of food, while…

Computer Vision and Pattern Recognition · Computer Science 2024-12-17 Pengkun Jiao , Xinlan Wu , Bin Zhu , Jingjing Chen , Chong-Wah Ngo , Yugang Jiang

Food retrieval is an important task to perform analysis of food-related information, where we are interested in retrieving relevant information about the queried food item such as ingredients, cooking instructions, etc. In this paper, we…

Computer Vision and Pattern Recognition · Computer Science 2021-09-22 Hao Wang , Doyen Sahoo , Chenghao Liu , Ke Shu , Palakorn Achananuparp , Ee-peng Lim , Steven C. H. Hoi

This paper introduces a novel task to evaluate the robust understanding capability of Large Multimodal Models (LMMs), termed $\textbf{Unsolvable Problem Detection (UPD)}$. Multiple-choice question answering (MCQA) is widely used to assess…

Computer Vision and Pattern Recognition · Computer Science 2025-06-10 Atsuyuki Miyai , Jingkang Yang , Jingyang Zhang , Yifei Ming , Qing Yu , Go Irie , Yixuan Li , Hai Li , Ziwei Liu , Kiyoharu Aizawa

Many automated tasks in software maintenance rely on information retrieval techniques to identify specific information within unstructured data. Bug localization is such a typical task, where text in a bug report is analyzed to identify…

Software Engineering · Computer Science 2019-02-08 Anil Koyuncu , Tegawendé F. Bissyandé , Dongsun Kim , Kui Liu , Jacques Klein , Martin Monperrus , Yves Le Traon

This paper introduces the 3rd place solution to the ICCV LargeFineFoodAI Retrieval Competition on Kaggle. Four basic models are independently trained with the weighted sum of ArcFace and Circle loss, then TTA and Ensemble are successively…

Computer Vision and Pattern Recognition · Computer Science 2025-11-03 Yang Zhong , Zhiming Wang , Zhaoyang Li , Jinyu Ma , Xiang Li

This paper reviews the NTIRE 2022 challenge on efficient single image super-resolution with focus on the proposed solutions and results. The task of the challenge was to super-resolve an input image with a magnification factor of $\times$4…

Computer Vision and Pattern Recognition · Computer Science 2022-05-12 Yawei Li , Kai Zhang , Radu Timofte , Luc Van Gool , Fangyuan Kong , Mingxi Li , Songwei Liu , Zongcai Du , Ding Liu , Chenhui Zhou , Jingyi Chen , Qingrui Han , Zheyuan Li , Yingqi Liu , Xiangyu Chen , Haoming Cai , Yu Qiao , Chao Dong , Long Sun , Jinshan Pan , Yi Zhu , Zhikai Zong , Xiaoxiao Liu , Zheng Hui , Tao Yang , Peiran Ren , Xuansong Xie , Xian-Sheng Hua , Yanbo Wang , Xiaozhong Ji , Chuming Lin , Donghao Luo , Ying Tai , Chengjie Wang , Zhizhong Zhang , Yuan Xie , Shen Cheng , Ziwei Luo , Lei Yu , Zhihong Wen , Qi Wu1 , Youwei Li , Haoqiang Fan , Jian Sun , Shuaicheng Liu , Yuanfei Huang , Meiguang Jin , Hua Huang , Jing Liu , Xinjian Zhang , Yan Wang , Lingshun Long , Gen Li , Yuanfan Zhang , Zuowei Cao , Lei Sun , Panaetov Alexander , Yucong Wang , Minjie Cai , Li Wang , Lu Tian , Zheyuan Wang , Hongbing Ma , Jie Liu , Chao Chen , Yidong Cai , Jie Tang , Gangshan Wu , Weiran Wang , Shirui Huang , Honglei Lu , Huan Liu , Keyan Wang , Jun Chen , Shi Chen , Yuchun Miao , Zimo Huang , Lefei Zhang , Mustafa Ayazoğlu , Wei Xiong , Chengyi Xiong , Fei Wang , Hao Li , Ruimian Wen , Zhijing Yang , Wenbin Zou , Weixin Zheng , Tian Ye , Yuncheng Zhang , Xiangzhen Kong , Aditya Arora , Syed Waqas Zamir , Salman Khan , Munawar Hayat , Fahad Shahbaz Khan , Dandan Gaoand Dengwen Zhouand Qian Ning , Jingzhu Tang , Han Huang , Yufei Wang , Zhangheng Peng , Haobo Li , Wenxue Guan , Shenghua Gong , Xin Li , Jun Liu , Wanjun Wang , Dengwen Zhou , Kun Zeng , Hanjiang Lin , Xinyu Chen , Jinsheng Fang

Identification of charged particles in a multilayer detector by the energy loss technique may also be achieved by the use of a neural network. The performance of the network becomes worse when a large fraction of information is missing, for…

Methodology · Statistics 2020-04-14 S. Riggi , D. Riggi , F. Riggi

Information fusion is used widely to improve document classification by the integration of multiple data sources (multimodal) or representations (multiview). However, the field lacks a unified framework, a quantitative synthesis of its…

Computation and Language · Computer Science 2026-05-27 Marcin Michał Mirończuk

This paper describes our approach for the triple scoring task at the WSDM Cup 2017. The task required participants to assign a relevance score for each pair of entities and their types in a knowledge base in order to enhance the ranking…

Computation and Language · Computer Science 2017-04-06 Ikuya Yamada , Motoki Sato , Hiroyuki Shindo

We propose a retrieval-augmented convolutional network and propose to train it with local mixup, a novel variant of the recently proposed mixup algorithm. The proposed hybrid architecture combining a convolutional network and an…

Machine Learning · Computer Science 2018-02-27 Jake Zhao , Kyunghyun Cho

Nonresponse after probability sampling is a universal challenge in survey sampling, often necessitating adjustments to mitigate sampling and selection bias simultaneously. This study explored the removal of bias and effective utilization of…

Methodology · Statistics 2025-11-13 Kosuke Morikawa , Kenji Beppu , Wataru Aida

In this paper, we describe the approach that we employed to address the task of Entity Recognition over Wet Lab Protocols -- a shared task in EMNLP WNUT-2020 Workshop. Our approach is composed of two phases. In the first phase, we…

Computation and Language · Computer Science 2020-12-17 Janvijay Singh , Anshul Wadhawan

This paper addresses the gap between general-purpose text embeddings and the specific demands of item retrieval tasks. We demonstrate the shortcomings of existing models in capturing the nuances necessary for zero-shot performance on item…

Information Retrieval · Computer Science 2024-03-01 Yuxuan Lei , Jianxun Lian , Jing Yao , Mingqi Wu , Defu Lian , Xing Xie

Two questions regarding practitioners' use of patent embeddings arise: (i) Does one fine-tuning recipe suffice for all downstream applications? (ii) Is fine-tuning on one patent landscape sufficient for downstream application on other…

Information Retrieval · Computer Science 2026-05-27 Amirhossein Yousefiramandi , Ciaran Cooney

Unified information extraction (UIE) aims to extract diverse structured information from unstructured text. While large language models (LLMs) have shown promise for UIE, they require significant computational resources and often struggle…

Computation and Language · Computer Science 2025-01-22 Xincheng Liao , Junwen Duan , Yixi Huang , Jianxin Wang

We present MM-Food-100K, a public 100,000-sample multimodal food intelligence dataset with verifiable provenance. It is a curated approximately 10% open subset of an original 1.2 million, quality-accepted corpus of food images annotated for…

Artificial Intelligence · Computer Science 2025-08-15 Yi Dong , Yusuke Muraoka , Scott Shi , Yi Zhang

Automatic emotion recognition is a challenging task. In this paper, we present our effort for the audio-video based sub-challenge of the Emotion Recognition in the Wild (EmotiW) 2018 challenge, which requires participants to assign a single…

Computer Vision and Pattern Recognition · Computer Science 2018-09-18 Zheng Lian , Ya Li , Jianhua Tao , Jian Huang

We participated in the Fifth UNLP shared task on multi-domain document understanding, where systems must answer Ukrainian multiple-choice questions from PDF collections and localize the supporting document and page. We propose a…

Computation and Language · Computer Science 2026-05-12 Anton Bazdyrev , Ivan Bashtovyi , Ivan Havlytskyi , Oleksandr Kharytonov , Artur Khodakovskyi

This report presents the CuriosAI team's submission to the EgoExo4D Proficiency Estimation Challenge at CVPR 2025. We propose two methods for multi-view skill assessment: (1) a multi-task learning framework using Sapiens-2B that jointly…

In this paper, we present our solutions for the Multimodal Sentiment Analysis Challenge (MuSe) 2022, which includes MuSe-Humor, MuSe-Reaction and MuSe-Stress Sub-challenges. The MuSe 2022 focuses on humor detection, emotional reactions and…

Computer Vision and Pattern Recognition · Computer Science 2022-08-15 Jia Li , Ziyang Zhang , Junjie Lang , Yueqi Jiang , Liuwei An , Peng Zou , Yangyang Xu , Sheng Gao , Jie Lin , Chunxiao Fan , Xiao Sun , Meng Wang