中文
相关论文

相关论文: [Call for Papers] The 2nd BabyLM Challenge: Sample…

200 篇论文

A primary challenge in large language model (LLM) development is their onerous pre-training cost. Typically, such pre-training involves optimizing a self-supervised objective (such as next-token prediction) over a large corpus. This paper…

Recently, decomposing complex problems into simple subtasks--a crucial part of human-like natural planning--to solve the given problem has significantly boosted the performance of large language models (LLMs). However, leveraging such…

计算与语言 · 计算机科学 2025-07-11 Mihir Parmar , Palash Goyal , Xin Liu , Yiwen Song , Mingyang Ling , Chitta Baral , Hamid Palangi , Tomas Pfister

Large volumes of text data have contributed significantly to the development of large language models (LLMs) in recent years. This data is typically acquired by scraping the internet, leading to pretraining datasets comprised of noisy web…

计算与语言 · 计算机科学 2023-09-12 Max Marion , Ahmet Üstün , Luiza Pozzobon , Alex Wang , Marzieh Fadaee , Sara Hooker

This paper describes the results of SemEval 2023 task 7 -- Multi-Evidence Natural Language Inference for Clinical Trial Data (NLI4CT) -- consisting of 2 tasks, a Natural Language Inference (NLI) task, and an evidence selection task on…

计算与语言 · 计算机科学 2023-05-12 Maël Jullien , Marco Valentino , Hannah Frost , Paul O'Regan , Donal Landers , André Freitas

LLMs have demonstrated impressive performance in answering medical questions, such as achieving passing scores on medical licensing examinations. However, medical board exams or general clinical questions do not capture the complexity of…

计算与语言 · 计算机科学 2026-02-19 Hanjie Chen , Zhouxiang Fang , Yash Singla , Mark Dredze

Recent advances in large language models (LLMs) have made significant progress across multiple biomedical tasks, including biomedical question answering, lay-language summarization of the biomedical literature, and clinical note…

信息检索 · 计算机科学 2026-03-24 Deepak Gupta , Dina Demner-Fushman , William Hersh , Steven Bedrick , Kirk Roberts

A LightGBM model fed with target word lexical characteristics and features obtained from word frequency lists, psychometric data and bigram association measures has been optimized for the 2021 CMCL Shared Task on Eye-Tracking Data…

计算与语言 · 计算机科学 2021-04-28 Yves Bestgen

Recently, Large Language Models (LLM) have demonstrated impressive capability to solve a wide range of tasks. However, despite their success across various tasks, no prior work has investigated their capability in the biomedical domain yet.…

计算与语言 · 计算机科学 2024-02-21 Israt Jahan , Md Tahmid Rahman Laskar , Chun Peng , Jimmy Huang

Continual learning (CL) in large language models (LLMs) is an evolving domain that focuses on developing efficient and sustainable training strategies to adapt models to emerging knowledge and achieve robustness in dynamic environments. Our…

计算与语言 · 计算机科学 2025-02-13 Çağatay Yıldız , Nishaanth Kanna Ravichandran , Nitin Sharma , Matthias Bethge , Beyza Ermis

Transformer-based Language Models are widely used in Natural Language Processing related tasks. Thanks to their pre-training, they have been successfully adapted to Information Extraction in business documents. However, most pre-training…

计算与语言 · 计算机科学 2023-09-12 Thibault Douzon , Stefan Duffner , Christophe Garcia , Jérémy Espinas

In the past few years, the emergence of pre-training models has brought uni-modal fields such as computer vision (CV) and natural language processing (NLP) to a new era. Substantial works have shown they are beneficial for downstream…

计算机视觉与模式识别 · 计算机科学 2024-04-19 Feilong Chen , Duzhen Zhang , Minglun Han , Xiuyi Chen , Jing Shi , Shuang Xu , Bo Xu

Conversational multi-doc question answering aims to answer specific questions based on the retrieved documents as well as the contextual conversations. In this paper, we introduce our winning approach for the "Conversational Multi-Doc QA"…

计算与语言 · 计算机科学 2024-02-29 Yiming Li , Zhao Zhang

Pretrained language models (PLMs) display impressive performances and have captured the attention of the NLP community. Establishing best practices in pretraining has, therefore, become a major focus of NLP research, especially since…

计算与语言 · 计算机科学 2024-10-08 Zihao Li , Shaoxiong Ji , Timothee Mickus , Vincent Segonne , Jörg Tiedemann

The rapid advancement of Large Language Models (LLMs) has introduced new possibilities and challenges in physics education, necessitating rigorous evaluation of their capabilities as both problem solvers and automated assessors. This paper…

物理教育 · 物理学 2026-05-25 Jonah R. Donaldson , Aliya Navaz , Konstantinos Doran , Alysta Lim , Mario Campanelli

Early children's developmental trajectories set up a natural goal for sample-efficient pretraining of vision foundation models. We introduce BabyVLM-V2, a developmentally grounded framework for infant-inspired vision-language modeling that…

Large Language Models (LLMs) applied to code-related applications have emerged as a prominent field, attracting significant interest from both academia and industry. However, as new and improved LLMs are developed, existing evaluation…

High-quality textual training data is essential for the success of multimodal data processing tasks, yet outputs from image captioning models like BLIP and GIT often contain errors and anomalies that are difficult to rectify using…

计算与语言 · 计算机科学 2025-02-25 Elyas Meguellati , Nardiena Pratama , Shazia Sadiq , Gianluca Demartini

The integration of machine learning (ML) into the physical sciences is reshaping computational paradigms, offering the potential to accelerate demanding simulations such as computational fluid dynamics (CFD). Yet, persistent challenges in…

As language models scale up, it becomes increasingly expensive to verify research ideas because conclusions on small models do not trivially transfer to large ones. A possible solution is to establish a generic system that accurately…

计算与语言 · 计算机科学 2024-04-09 Yiqun Yao , Siqi fan , Xiusheng Huang , Xuezhi Fang , Xiang Li , Ziyi Ni , Xin Jiang , Xuying Meng , Peng Han , Shuo Shang , Kang Liu , Aixin Sun , Yequan Wang

Here, we present the outcomes from the second Large Language Model (LLM) Hackathon for Applications in Materials Science and Chemistry, which engaged participants across global hybrid locations, resulting in 34 team submissions. The…

机器学习 · 计算机科学 2025-01-06 Yoel Zimmermann , Adib Bazgir , Zartashia Afzal , Fariha Agbere , Qianxiang Ai , Nawaf Alampara , Alexander Al-Feghali , Mehrad Ansari , Dmytro Antypov , Amro Aswad , Jiaru Bai , Viktoriia Baibakova , Devi Dutta Biswajeet , Erik Bitzek , Joshua D. Bocarsly , Anna Borisova , Andres M Bran , L. Catherine Brinson , Marcel Moran Calderon , Alessandro Canalicchio , Victor Chen , Yuan Chiang , Defne Circi , Benjamin Charmes , Vikrant Chaudhary , Zizhang Chen , Min-Hsueh Chiu , Judith Clymo , Kedar Dabhadkar , Nathan Daelman , Archit Datar , Wibe A. de Jong , Matthew L. Evans , Maryam Ghazizade Fard , Giuseppe Fisicaro , Abhijeet Sadashiv Gangan , Janine George , Jose D. Cojal Gonzalez , Michael Götte , Ankur K. Gupta , Hassan Harb , Pengyu Hong , Abdelrahman Ibrahim , Ahmed Ilyas , Alishba Imran , Kevin Ishimwe , Ramsey Issa , Kevin Maik Jablonka , Colin Jones , Tyler R. Josephson , Greg Juhasz , Sarthak Kapoor , Rongda Kang , Ghazal Khalighinejad , Sartaaj Khan , Sascha Klawohn , Suneel Kuman , Alvin Noe Ladines , Sarom Leang , Magdalena Lederbauer , Sheng-Lun , Liao , Hao Liu , Xuefeng Liu , Stanley Lo , Sandeep Madireddy , Piyush Ranjan Maharana , Shagun Maheshwari , Soroush Mahjoubi , José A. Márquez , Rob Mills , Trupti Mohanty , Bernadette Mohr , Seyed Mohamad Moosavi , Alexander Moßhammer , Amirhossein D. Naghdi , Aakash Naik , Oleksandr Narykov , Hampus Näsström , Xuan Vu Nguyen , Xinyi Ni , Dana O'Connor , Teslim Olayiwola , Federico Ottomano , Aleyna Beste Ozhan , Sebastian Pagel , Chiku Parida , Jaehee Park , Vraj Patel , Elena Patyukova , Martin Hoffmann Petersen , Luis Pinto , José M. Pizarro , Dieter Plessers , Tapashree Pradhan , Utkarsh Pratiush , Charishma Puli , Andrew Qin , Mahyar Rajabi , Francesco Ricci , Elliot Risch , Martiño Ríos-García , Aritra Roy , Tehseen Rug , Hasan M Sayeed , Markus Scheidgen , Mara Schilling-Wilhelmi , Marcel Schloz , Fabian Schöppach , Julia Schumann , Philippe Schwaller , Marcus Schwarting , Samiha Sharlin , Kevin Shen , Jiale Shi , Pradip Si , Jennifer D'Souza , Taylor Sparks , Suraj Sudhakar , Leopold Talirz , Dandan Tang , Olga Taran , Carla Terboven , Mark Tropin , Anastasiia Tsymbal , Katharina Ueltzen , Pablo Andres Unzueta , Archit Vasan , Tirtha Vinchurkar , Trung Vo , Gabriel Vogel , Christoph Völker , Jan Weinreich , Faradawn Yang , Mohd Zaki , Chi Zhang , Sylvester Zhang , Weijie Zhang , Ruijie Zhu , Shang Zhu , Jan Janssen , Calvin Li , Ian Foster , Ben Blaiszik