English
Related papers

Related papers: Winning Amazon KDD Cup'24

200 papers

This article details our participation (L3iTC) in the FinLLM Challenge Task 2024, focusing on two key areas: Task 1, financial text classification, and Task 2, financial text summarization. To address these challenges, we fine-tuned several…

Computation and Language · Computer Science 2024-08-07 Elvys Linhares Pontes , Carlos-Emiliano González-Gallardo , Mohamed Benjannet , Caryn Qu , Antoine Doucet

This paper presents an overview of the Arabic Natural Language Understanding (ArabicNLU 2024) shared task, focusing on two subtasks: Word Sense Disambiguation (WSD) and Location Mention Disambiguation (LMD). The task aimed to evaluate the…

Computation and Language · Computer Science 2024-07-31 Mohammed Khalilia , Sanad Malaysha , Reem Suwaileh , Mustafa Jarrar , Alaa Aljabari , Tamer Elsayed , Imed Zitouni

Here, we present the outcomes from the second Large Language Model (LLM) Hackathon for Applications in Materials Science and Chemistry, which engaged participants across global hybrid locations, resulting in 34 team submissions. The…

Machine Learning · Computer Science 2025-01-06 Yoel Zimmermann , Adib Bazgir , Zartashia Afzal , Fariha Agbere , Qianxiang Ai , Nawaf Alampara , Alexander Al-Feghali , Mehrad Ansari , Dmytro Antypov , Amro Aswad , Jiaru Bai , Viktoriia Baibakova , Devi Dutta Biswajeet , Erik Bitzek , Joshua D. Bocarsly , Anna Borisova , Andres M Bran , L. Catherine Brinson , Marcel Moran Calderon , Alessandro Canalicchio , Victor Chen , Yuan Chiang , Defne Circi , Benjamin Charmes , Vikrant Chaudhary , Zizhang Chen , Min-Hsueh Chiu , Judith Clymo , Kedar Dabhadkar , Nathan Daelman , Archit Datar , Wibe A. de Jong , Matthew L. Evans , Maryam Ghazizade Fard , Giuseppe Fisicaro , Abhijeet Sadashiv Gangan , Janine George , Jose D. Cojal Gonzalez , Michael Götte , Ankur K. Gupta , Hassan Harb , Pengyu Hong , Abdelrahman Ibrahim , Ahmed Ilyas , Alishba Imran , Kevin Ishimwe , Ramsey Issa , Kevin Maik Jablonka , Colin Jones , Tyler R. Josephson , Greg Juhasz , Sarthak Kapoor , Rongda Kang , Ghazal Khalighinejad , Sartaaj Khan , Sascha Klawohn , Suneel Kuman , Alvin Noe Ladines , Sarom Leang , Magdalena Lederbauer , Sheng-Lun , Liao , Hao Liu , Xuefeng Liu , Stanley Lo , Sandeep Madireddy , Piyush Ranjan Maharana , Shagun Maheshwari , Soroush Mahjoubi , José A. Márquez , Rob Mills , Trupti Mohanty , Bernadette Mohr , Seyed Mohamad Moosavi , Alexander Moßhammer , Amirhossein D. Naghdi , Aakash Naik , Oleksandr Narykov , Hampus Näsström , Xuan Vu Nguyen , Xinyi Ni , Dana O'Connor , Teslim Olayiwola , Federico Ottomano , Aleyna Beste Ozhan , Sebastian Pagel , Chiku Parida , Jaehee Park , Vraj Patel , Elena Patyukova , Martin Hoffmann Petersen , Luis Pinto , José M. Pizarro , Dieter Plessers , Tapashree Pradhan , Utkarsh Pratiush , Charishma Puli , Andrew Qin , Mahyar Rajabi , Francesco Ricci , Elliot Risch , Martiño Ríos-García , Aritra Roy , Tehseen Rug , Hasan M Sayeed , Markus Scheidgen , Mara Schilling-Wilhelmi , Marcel Schloz , Fabian Schöppach , Julia Schumann , Philippe Schwaller , Marcus Schwarting , Samiha Sharlin , Kevin Shen , Jiale Shi , Pradip Si , Jennifer D'Souza , Taylor Sparks , Suraj Sudhakar , Leopold Talirz , Dandan Tang , Olga Taran , Carla Terboven , Mark Tropin , Anastasiia Tsymbal , Katharina Ueltzen , Pablo Andres Unzueta , Archit Vasan , Tirtha Vinchurkar , Trung Vo , Gabriel Vogel , Christoph Völker , Jan Weinreich , Faradawn Yang , Mohd Zaki , Chi Zhang , Sylvester Zhang , Weijie Zhang , Ruijie Zhu , Shang Zhu , Jan Janssen , Calvin Li , Ian Foster , Ben Blaiszik

In this work, we address the challenge of multilingual category relevance judgment in e-commerce search, where traditional ensemble-based systems improve accuracy but at the cost of heavy training, inference, and maintenance complexity. To…

Information Retrieval · Computer Science 2026-01-12 Haotao Xie , Ruilin Chen , Yicheng Wu , Zhan Zhao , Yuanyuan Liu

We present a low-cost retrieval system for the WSDM Cup 2026 multilingual retrieval task, where English queries are used to retrieve relevant documents from a collection of approximately ten million news articles in Chinese, Persian, and…

Information Retrieval · Computer Science 2026-02-20 Chentong Hao , Minmao Wang

Improving the quality of search results can significantly enhance users experience and engagement with search engines. In spite of several recent advancements in the fields of machine learning and data mining, correctly classifying items…

Despite long-standing efforts in accelerating scientific discovery with AI, building AI co-scientists remains challenging due to limited high-quality data for training and evaluation. To tackle this data scarcity issue, we present AutoSDT,…

This work explores a novel data augmentation method based on Large Language Models (LLMs) for predicting item difficulty and response time of retired USMLE Multiple-Choice Questions (MCQs) in the BEA 2024 Shared Task. Our approach is based…

Computation and Language · Computer Science 2024-04-23 Ana-Cristina Rogoz , Radu Tudor Ionescu

Real-world multi-modal problems are rarely solved by a single machine learning model, and often require multi-step computational plans that involve stitching several models. Tool-augmented LLMs hold tremendous promise for automating the…

Computer Vision and Pattern Recognition · Computer Science 2024-09-24 Zixian Ma , Weikai Huang , Jieyu Zhang , Tanmay Gupta , Ranjay Krishna

Despite recent progress in Natural Language Understanding (NLU), the creation of multilingual NLU systems remains a challenge. It is common to have NLU systems limited to a subset of languages due to lack of available data. They also often…

Computation and Language · Computer Science 2022-12-14 Christopher Hench , Charith Peris , Jack FitzGerald , Kay Rottmann

In this report, we introduce Qwen2.5, a comprehensive series of large language models (LLMs) designed to meet diverse needs. Compared to previous iterations, Qwen 2.5 has been significantly improved during both the pre-training and…

In this paper we summarize the results of the Putnam-like benchmark published by Google DeepMind. This dataset consists of 96 original problems in the spirit of the Putnam Competition and 576 solutions generated by LLMs. We analyze the…

Machine Learning · Computer Science 2026-02-03 Bartosz Bieganowski , Daniel Strzelecki , Robert Skiba , Mateusz Topolewski

This paper outlines the LLMs4OL 2024, the first edition of the Large Language Models for Ontology Learning Challenge. LLMs4OL is a community development initiative collocated with the 23rd International Semantic Web Conference (ISWC) to…

Computation and Language · Computer Science 2024-09-17 Hamed Babaei Giglou , Jennifer D'Souza , Sören Auer

The goal of the BabyLM is to stimulate new research connections between cognitive modeling and language model pretraining. We invite contributions in this vein to the BabyLM Workshop, which will also include the 4th iteration of the BabyLM…

This paper introduces two multilingual systems, IKUN and IKUN-C, developed for the general machine translation task in WMT24. IKUN and IKUN-C represent an open system and a constrained system, respectively, built on Llama-3-8b and…

Computation and Language · Computer Science 2024-08-30 Baohao Liao , Christian Herold , Shahram Khadivi , Christof Monz

The exponential growth of scientific literature poses unprecedented challenges for researchers attempting to synthesize knowledge across rapidly evolving fields. We present \textbf{Agentic AutoSurvey}, a multi-agent framework for automated…

Information Retrieval · Computer Science 2025-09-24 Yixin Liu , Yonghui Wu , Denghui Zhang , Lichao Sun

As the number and variety of smart devices increase, users may use myriad devices in their daily lives and the online activities become highly fragmented. Building an accurate user identity becomes a difficult and important problem for…

Social and Information Networks · Computer Science 2016-10-14 Jianxun Lian , Xing Xie

In this paper, the solution of HYU MLLAB KT Team to the Multimodal Algorithmic Reasoning Task: SMART-101 CVPR 2024 Challenge is presented. Beyond conventional visual question-answering problems, the SMART-101 challenge aims to achieve…

Computer Vision and Pattern Recognition · Computer Science 2024-06-11 Jinwoo Ahn , Junhyeok Park , Min-Jun Kim , Kang-Hyeon Kim , So-Yeong Sohn , Yun-Ji Lee , Du-Seong Chang , Yu-Jung Heo , Eun-Sol Kim

This paper provides an overview of the Arabic Sentiment Analysis Challenge organized by King Abdullah University of Science and Technology (KAUST). The task in this challenge is to develop machine learning models to classify a given tweet…

Computation and Language · Computer Science 2021-09-30 Hind Alamro , Manal Alshehri , Basma Alharbi , Zuhair Khayyat , Manal Kalkatawi , Inji Ibrahim Jaber , Xiangliang Zhang

This paper presents an overview of the Volvo Discovery Challenge, held during the ECML-PKDD 2024 conference. The challenge's goal was to predict the failure risk of an anonymized component in Volvo trucks using a newly published dataset.…