English
Related papers

Related papers: [Call for Papers] The 2nd BabyLM Challenge: Sample…

200 papers

A primary challenge in large language model (LLM) development is their onerous pre-training cost. Typically, such pre-training involves optimizing a self-supervised objective (such as next-token prediction) over a large corpus. This paper…

Recently, decomposing complex problems into simple subtasks--a crucial part of human-like natural planning--to solve the given problem has significantly boosted the performance of large language models (LLMs). However, leveraging such…

Computation and Language · Computer Science 2025-07-11 Mihir Parmar , Palash Goyal , Xin Liu , Yiwen Song , Mingyang Ling , Chitta Baral , Hamid Palangi , Tomas Pfister

Large volumes of text data have contributed significantly to the development of large language models (LLMs) in recent years. This data is typically acquired by scraping the internet, leading to pretraining datasets comprised of noisy web…

Computation and Language · Computer Science 2023-09-12 Max Marion , Ahmet Üstün , Luiza Pozzobon , Alex Wang , Marzieh Fadaee , Sara Hooker

This paper describes the results of SemEval 2023 task 7 -- Multi-Evidence Natural Language Inference for Clinical Trial Data (NLI4CT) -- consisting of 2 tasks, a Natural Language Inference (NLI) task, and an evidence selection task on…

Computation and Language · Computer Science 2023-05-12 Maël Jullien , Marco Valentino , Hannah Frost , Paul O'Regan , Donal Landers , André Freitas

LLMs have demonstrated impressive performance in answering medical questions, such as achieving passing scores on medical licensing examinations. However, medical board exams or general clinical questions do not capture the complexity of…

Computation and Language · Computer Science 2026-02-19 Hanjie Chen , Zhouxiang Fang , Yash Singla , Mark Dredze

Recent advances in large language models (LLMs) have made significant progress across multiple biomedical tasks, including biomedical question answering, lay-language summarization of the biomedical literature, and clinical note…

Information Retrieval · Computer Science 2026-03-24 Deepak Gupta , Dina Demner-Fushman , William Hersh , Steven Bedrick , Kirk Roberts

A LightGBM model fed with target word lexical characteristics and features obtained from word frequency lists, psychometric data and bigram association measures has been optimized for the 2021 CMCL Shared Task on Eye-Tracking Data…

Computation and Language · Computer Science 2021-04-28 Yves Bestgen

Recently, Large Language Models (LLM) have demonstrated impressive capability to solve a wide range of tasks. However, despite their success across various tasks, no prior work has investigated their capability in the biomedical domain yet.…

Computation and Language · Computer Science 2024-02-21 Israt Jahan , Md Tahmid Rahman Laskar , Chun Peng , Jimmy Huang

Continual learning (CL) in large language models (LLMs) is an evolving domain that focuses on developing efficient and sustainable training strategies to adapt models to emerging knowledge and achieve robustness in dynamic environments. Our…

Computation and Language · Computer Science 2025-02-13 Çağatay Yıldız , Nishaanth Kanna Ravichandran , Nitin Sharma , Matthias Bethge , Beyza Ermis

Transformer-based Language Models are widely used in Natural Language Processing related tasks. Thanks to their pre-training, they have been successfully adapted to Information Extraction in business documents. However, most pre-training…

Computation and Language · Computer Science 2023-09-12 Thibault Douzon , Stefan Duffner , Christophe Garcia , Jérémy Espinas

In the past few years, the emergence of pre-training models has brought uni-modal fields such as computer vision (CV) and natural language processing (NLP) to a new era. Substantial works have shown they are beneficial for downstream…

Computer Vision and Pattern Recognition · Computer Science 2024-04-19 Feilong Chen , Duzhen Zhang , Minglun Han , Xiuyi Chen , Jing Shi , Shuang Xu , Bo Xu

Conversational multi-doc question answering aims to answer specific questions based on the retrieved documents as well as the contextual conversations. In this paper, we introduce our winning approach for the "Conversational Multi-Doc QA"…

Computation and Language · Computer Science 2024-02-29 Yiming Li , Zhao Zhang

Pretrained language models (PLMs) display impressive performances and have captured the attention of the NLP community. Establishing best practices in pretraining has, therefore, become a major focus of NLP research, especially since…

Computation and Language · Computer Science 2024-10-08 Zihao Li , Shaoxiong Ji , Timothee Mickus , Vincent Segonne , Jörg Tiedemann

The rapid advancement of Large Language Models (LLMs) has introduced new possibilities and challenges in physics education, necessitating rigorous evaluation of their capabilities as both problem solvers and automated assessors. This paper…

Physics Education · Physics 2026-05-25 Jonah R. Donaldson , Aliya Navaz , Konstantinos Doran , Alysta Lim , Mario Campanelli

Early children's developmental trajectories set up a natural goal for sample-efficient pretraining of vision foundation models. We introduce BabyVLM-V2, a developmentally grounded framework for infant-inspired vision-language modeling that…

Large Language Models (LLMs) applied to code-related applications have emerged as a prominent field, attracting significant interest from both academia and industry. However, as new and improved LLMs are developed, existing evaluation…

Software Engineering · Computer Science 2024-06-07 Naman Jain , King Han , Alex Gu , Wen-Ding Li , Fanjia Yan , Tianjun Zhang , Sida Wang , Armando Solar-Lezama , Koushik Sen , Ion Stoica

High-quality textual training data is essential for the success of multimodal data processing tasks, yet outputs from image captioning models like BLIP and GIT often contain errors and anomalies that are difficult to rectify using…

Computation and Language · Computer Science 2025-02-25 Elyas Meguellati , Nardiena Pratama , Shazia Sadiq , Gianluca Demartini

The integration of machine learning (ML) into the physical sciences is reshaping computational paradigms, offering the potential to accelerate demanding simulations such as computational fluid dynamics (CFD). Yet, persistent challenges in…

As language models scale up, it becomes increasingly expensive to verify research ideas because conclusions on small models do not trivially transfer to large ones. A possible solution is to establish a generic system that accurately…

Computation and Language · Computer Science 2024-04-09 Yiqun Yao , Siqi fan , Xiusheng Huang , Xuezhi Fang , Xiang Li , Ziyi Ni , Xin Jiang , Xuying Meng , Peng Han , Shuo Shang , Kang Liu , Aixin Sun , Yequan Wang

Here, we present the outcomes from the second Large Language Model (LLM) Hackathon for Applications in Materials Science and Chemistry, which engaged participants across global hybrid locations, resulting in 34 team submissions. The…

Machine Learning · Computer Science 2025-01-06 Yoel Zimmermann , Adib Bazgir , Zartashia Afzal , Fariha Agbere , Qianxiang Ai , Nawaf Alampara , Alexander Al-Feghali , Mehrad Ansari , Dmytro Antypov , Amro Aswad , Jiaru Bai , Viktoriia Baibakova , Devi Dutta Biswajeet , Erik Bitzek , Joshua D. Bocarsly , Anna Borisova , Andres M Bran , L. Catherine Brinson , Marcel Moran Calderon , Alessandro Canalicchio , Victor Chen , Yuan Chiang , Defne Circi , Benjamin Charmes , Vikrant Chaudhary , Zizhang Chen , Min-Hsueh Chiu , Judith Clymo , Kedar Dabhadkar , Nathan Daelman , Archit Datar , Wibe A. de Jong , Matthew L. Evans , Maryam Ghazizade Fard , Giuseppe Fisicaro , Abhijeet Sadashiv Gangan , Janine George , Jose D. Cojal Gonzalez , Michael Götte , Ankur K. Gupta , Hassan Harb , Pengyu Hong , Abdelrahman Ibrahim , Ahmed Ilyas , Alishba Imran , Kevin Ishimwe , Ramsey Issa , Kevin Maik Jablonka , Colin Jones , Tyler R. Josephson , Greg Juhasz , Sarthak Kapoor , Rongda Kang , Ghazal Khalighinejad , Sartaaj Khan , Sascha Klawohn , Suneel Kuman , Alvin Noe Ladines , Sarom Leang , Magdalena Lederbauer , Sheng-Lun , Liao , Hao Liu , Xuefeng Liu , Stanley Lo , Sandeep Madireddy , Piyush Ranjan Maharana , Shagun Maheshwari , Soroush Mahjoubi , José A. Márquez , Rob Mills , Trupti Mohanty , Bernadette Mohr , Seyed Mohamad Moosavi , Alexander Moßhammer , Amirhossein D. Naghdi , Aakash Naik , Oleksandr Narykov , Hampus Näsström , Xuan Vu Nguyen , Xinyi Ni , Dana O'Connor , Teslim Olayiwola , Federico Ottomano , Aleyna Beste Ozhan , Sebastian Pagel , Chiku Parida , Jaehee Park , Vraj Patel , Elena Patyukova , Martin Hoffmann Petersen , Luis Pinto , José M. Pizarro , Dieter Plessers , Tapashree Pradhan , Utkarsh Pratiush , Charishma Puli , Andrew Qin , Mahyar Rajabi , Francesco Ricci , Elliot Risch , Martiño Ríos-García , Aritra Roy , Tehseen Rug , Hasan M Sayeed , Markus Scheidgen , Mara Schilling-Wilhelmi , Marcel Schloz , Fabian Schöppach , Julia Schumann , Philippe Schwaller , Marcus Schwarting , Samiha Sharlin , Kevin Shen , Jiale Shi , Pradip Si , Jennifer D'Souza , Taylor Sparks , Suraj Sudhakar , Leopold Talirz , Dandan Tang , Olga Taran , Carla Terboven , Mark Tropin , Anastasiia Tsymbal , Katharina Ueltzen , Pablo Andres Unzueta , Archit Vasan , Tirtha Vinchurkar , Trung Vo , Gabriel Vogel , Christoph Völker , Jan Weinreich , Faradawn Yang , Mohd Zaki , Chi Zhang , Sylvester Zhang , Weijie Zhang , Ruijie Zhu , Shang Zhu , Jan Janssen , Calvin Li , Ian Foster , Ben Blaiszik
‹ Prev 1 4 5 6 7 8 10 Next ›