中文
相关论文

相关论文: Functional Abstraction of Knowledge Recall in Larg…

200 篇论文

Large language models (LLMs) have shown remarkable capabilities in solving complex tasks. Recent work has explored decomposing such tasks into subtasks with independent contexts. However, some contextually related subtasks may encounter…

计算与语言 · 计算机科学 2025-04-01 Hongjia Liu , Jinlong Li

This study investigates whether large language models (LLMs) mirror human neurocognition during abstract reasoning. We compared the performance and neural representations of human participants with those of eight open-source LLMs on an…

Language model pre-training has been shown to capture a surprising amount of world knowledge, crucial for NLP tasks such as question answering. However, this knowledge is stored implicitly in the parameters of a neural network, requiring…

计算与语言 · 计算机科学 2020-02-21 Kelvin Guu , Kenton Lee , Zora Tung , Panupong Pasupat , Ming-Wei Chang

People acquire concepts through rich physical and social experiences and use them to understand and navigate the world. In contrast, large language models (LLMs), trained solely through next-token prediction on text, exhibit strikingly…

计算与语言 · 计算机科学 2025-11-11 Ningyu Xu , Qi Zhang , Chao Du , Qiang Luo , Xipeng Qiu , Xuanjing Huang , Menghan Zhang

In-context learning (ICL) has emerged as an effective solution for few-shot learning with large language models (LLMs). However, how LLMs leverage demonstrations to specify a task and learn a corresponding computational function through ICL…

The capabilities of large language models (LLMs) have expanded beyond natural language processing to scientific prediction tasks, including molecular property prediction. However, their effectiveness in in-context learning remains…

Large pre-trained language models (PLMs) have been shown to retain implicit knowledge within their parameters. To enhance this implicit knowledge, we propose Knowledge Injection into Language Models (KILM), a novel approach that injects…

计算与语言 · 计算机科学 2023-02-21 Yan Xu , Mahdi Namazifar , Devamanyu Hazarika , Aishwarya Padmakumar , Yang Liu , Dilek Hakkani-Tür

The statistical study of human memory requires large-scale experiments, involving many stimuli conditions and test subjects. While this approach has proven to be quite fruitful for meaningless material such as random lists of words,…

计算与语言 · 计算机科学 2024-11-26 Antonios Georgiou , Tankut Can , Mikhail Katkov , Misha Tsodyks

Understanding how individuals perceive and recall information in their natural environments is critical to understanding potential failures in perception (e.g., sensory loss) and memory (e.g., dementia). Event segmentation, the process of…

计算与语言 · 计算机科学 2025-10-20 Ryan A. Panela , Alex J. Barnett , Morgan D. Barense , Björn Herrmann

Large multimodal models (LMMs) combine unimodal encoders and large language models (LLMs) to perform multimodal tasks. Despite recent advancements towards the interpretability of these models, understanding internal representations of LMMs…

机器学习 · 计算机科学 2024-12-03 Jayneel Parekh , Pegah Khayatan , Mustafa Shukor , Alasdair Newson , Matthieu Cord

Effective memory management is essential for large language model (LLM) agents handling long-term interactions. Current memory frameworks typically treat agents as passive "recorders" and retrieve information without understanding its…

计算与语言 · 计算机科学 2026-03-03 Xiaohui Zhang , Zequn Sun , Chengyuan Yang , Yaqin Jin , Yazhong Zhang , Wei Hu

Existing Multimodal Large Language Models (MLLMs) process a large number of visual tokens, leading to significant computational costs and inefficiency. Instruction-related visual token compression demonstrates strong task relevance, which…

计算机视觉与模式识别 · 计算机科学 2026-05-05 Lei Lei , Jie Gu , Xiaokang Ma , Chu Tang , Jingmin Chen , Tong Xu

Named entities are fundamental building blocks of knowledge in text, grounding factual information and structuring relationships within language. Despite their importance, it remains unclear how Large Language Models (LLMs) internally…

计算与语言 · 计算机科学 2025-10-13 Victor Morand , Josiane Mothe , Benjamin Piwowarski

Fine-tuning large language models (LLMs) can cause them to lose their general capabilities. However, the intrinsic mechanisms behind such forgetting remain unexplored. In this paper, we begin by examining this phenomenon by focusing on…

人工智能 · 计算机科学 2024-12-02 Gangwei Jiang , Zhaoyi Li , Defu Lian , Ying Wei

Large language models face challenges in long-context question answering, where key evidence of a query may be dispersed across millions of tokens. Existing works equip large language models with a memory buffer that is dynamically updated…

计算与语言 · 计算机科学 2026-03-03 Yaorui Shi , Yuxin Chen , Siyuan Wang , Sihang Li , Hengxing Cai , Qi Gu , Xiang Wang , An Zhang

In this paper, we propose active recap learning (ARL), a framework for enhancing large language model (LLM) in understanding long contexts. ARL enables models to revisit and summarize earlier content through targeted sequence construction…

计算与语言 · 计算机科学 2026-01-21 Chenyu Hui

While large language models (LLMs) excel at factual recall, the real challenge lies in knowledge application. A gap persists between their ability to answer complex questions and their effectiveness in performing tasks that require that…

计算与语言 · 计算机科学 2026-01-21 Siyang Wu , Honglin Bao , Nadav Kunievsky , James A. Evans

Large language models (LLMs) are rapidly changing how researchers in materials science and chemistry discover, organize, and act on scientific knowledge. This paper analyzes a broad set of community-developed LLM applications in an effort…

材料科学 · 物理学 2026-05-06 Aritra Roy , Kevin Shen , Andrew MacBride , Awwal Oladipupo , Mudassra Taskeen , Wojtek Treyde , Ruaa A. E. A. Abakar , Ahmad D. Abbas , Elsayed Abdelfatah , Abbas A. Abdullahi , Seham S. Abyah , Chahd Rahyl Adjmi , Fariha Agbere , Savyasanchi Aggarwal , Muhammad Ahmed , Tasnim Ahmed , Motasem Ajlouni , Mattias Akke , Hussein AlAdwan , Anwaar S. Alazani , Zahra A. Alharbi , Wajd A. Aljulyhi , Mohammed A. AlKubaish , Fatima A. Almahri , Sayed A. Almohri , David Obeh Alobo , Mohammed Alouni , Azizah S. Alqahtani , Omar Alsaigh , Husain Althagafi , Md. Aqib Aman , Lena Ara , Arifin , Ignacio Arretche , Abdulaziz Ashy , Syeda A. Asim , Amro Aswad , Adeel Atta , Sören Auer , Abdullah al Azmi , Toheeb Balogun , Suvo Banik , Viktoriia Baibakova , Shakira A. Baksh , Neus G. Bastús , Christina J. Bayard , Adib Bazgir , Louis Beal , Lejla Biberić , Wahid Billah , Ankita Biswas , Joshua Bocarsly , Montassar T. Bouzidi , Esma B. Boydas , Youssef Briki , Cailin Buchanan , Mauricio Cafiero , Damien Caliste , Yi Cao , Rafael E. Castañeda , Sruthy K. Chandy , Benjamin Charmes , Shayantan Chaudhuri , Yiming Chen , Alexander Chen , Jieneng Chen , Min-Hsueh Chiu , Defne Circi , Cinthya H. Contreras , Yoann Cure , Nathan Daelman , Roshini Dantuluri , Thomas Davy , William Dawson , Leonid Didukh , Rui Ding , Aminu R. Doguwa , Claudia Draxl , Sathya Edamadaka , Oulaya Elargab , Christina Ertural , Matthew L. Evans , Edvin Fako , Hossam Farag , Nur A. Fathurrahman , Merve Fedai , Rodrigo P. Ferreira , Giuseppe Fisicaro , Thomas Frank , Sasi K. Gaddipati , Abhijeet Gangan , Jennifer Garland , James Garrick , Luigi Genovese , Maryam Ghadrdran , Sandip Giri , Maxime Goulet , Jeremy Goumaz , Sara U. Gracia , Jacob Graham , Gabriel Graves , Kevin P. Greenman , Tim Greitemeier , Cameron Gruich , Sophie Gu , Salomé Guilbert , Hans Gundlach , Muriel F. Gusta , Mourad El Haddaoui , Alexander J. Haibel , Anubhab Haldar , Vehaan Handa , Hassan Harb , Nathan D. Harms , Abdullah Al Hasan , Abir Hassan , Qiyao He , Andrés Henao-Aristizábal , Bram Hoex , Sungil Hong , Alexander J. Horvath , Md. Shaib Hossain , Yanqi Huang , Yuqing Huang , Kostiantyn Hubaiev , Donald Intal , Katherine Inzani , Kevin Ishimwe , Tugba Isik , Gopal R. Iyer , Katharina Jager , Jan Janssen , Hyewon Jeong , Michael Jirasek , Tyler R. Josephson , Nisarg Joshi , Yassir Ben Kacem , Remya A. M. Kalapurakal , Rakesh R. Kamath , Sugan Kanagasenthinathan , Dohun Kang , Jason Kantorow , Kübra Kaygisiz , Murat Keceli , Farhana Keya , Muhammad U. Khan , Sartaaj Takrim Khan , Hyungjun Kim , Alexander Kister , Sascha Klawohn , Collin Kovacs , Pranav Krishnan , Maurycy Kryzanowski , Ritesh Kumar , Suman Kumari , Gourav Kumbhojkar , Ryo Kuroki , Shashank Kushwaha , Magdalena Lederbauer , Jaejun Lee , Seunghan Lee , Jeonghwan Lee , Bingcan Li , Calvin Li , Zhanzhao Li , Shi Li , Shicheng Li , Chengyan Liu , Hao Liu , Tung Yan Liu , Yutong Liu , Lucia Vina-Lopez , Chayaphol Lortaraparsert , Andre K. Y. Low , Saffron Luxford , Carlos Madariaga , Rishikesh Magar , Piyush R. Maharana , Rahul Mallela , Shoaib Mahmud , Natesan Mani , Umair Mansoor , Omar B. Mansour , Cassandra Masschelein , Kinga O. Mastej , Ankit Mathanker , Jeffrey Meng , Omran Mezghani , Yidong Ming , Rishav Mitra , Michail Mitsakis , Matthew Miyagishima , Ravikumar Mohan , Naveen R. Mohanraj , Trupti Mohanty , Bernadette Mohr , Francisco A. Molina-Bakhos , Jeremy Monat , Seyed Mohamad Moosavi , Shayan Mousavi , Arman Moussavi , Rubel Mozumber , Muhammad J. Mufti , Diyana Muhammed , Ram Munde , Mrigi Munjal , José A. Márquez , Shankha Nag , Giacomo Nagaro , Juno Nam , Jose M. Napoles-Duarte , Ry Nduma , Xuan-Vu Nguyen , Ebrahim Norouzi , Oluwatosin Ohiro , Ryotaro Okabe , Viejay Ordillo , Shuichiro Ozawa , Sebastian Pagel , Daniel Palmer , Angela Pan , Akash Pandey , Vivek Pandit , Prakul Pandit , Chiku Parida , Jaehee Park , Hyunsoo Park , Hemangi Patel , Shakul Pathak , Taradutt Pattnaik , Elena Patyukova , Noah Paulson , Deepak S. Pendyala , Erick S. Pepek , Martin H. Petersen , Thang D. Pham , Aniket Phutane , Sabila K. Pinky , Étienne Polack , Alison Polasik , Maria Politi , Tim Pongratz , Akhila Ponugoti , Fabio Priante , Thomas Michael Pruyn , Sai S. Puppala , Mohammad A. Qazi , Heike Quosdorf , Gollam Rabby , Mohammad J. Raei , Md. Habibur Rahman , A. B. M. Ashikur Rahman , Subhashree Rajasekaran , Tawfiqur Rakib , Hemanth N. Ramesh , Vrushali Ranadive , Karnamohit Ranka , Bojana Rankovic , Adwaith Ravichandran , Ilija Rašović , Sergei Rigin , Tatem Rios , Varun Rishi , Victor Naden Robinson , Lucas S. Rodrigues , Oswaldo Rodriguez , Mahule Roy , Diptendu Roy , Subhas Roy , Arokia Anto Royan M , Joseph F. Rudzinski , Muhammad Sabih , Subramanyam Sahoo , Srusti Bheem Sain , Thahira Saliya , Vignesh Sampath , Jesus Diaz Sanchez , Arthur S. S. Santos , Muliady Satria , Hasan M. Sayeed , Jörg Schaarschmidt , Philippe Schwaller , Nofit Segal , Abhishec Senthilvel , Sherjeel Shabih , Devanshu Shah , Faezeh Shahmoradi , Samiha Sharlin , Killian Sheriff , Qiuyu Shi , Abubakar D. Shuaibu , Ayesha Siddiqua , M. A. Shadab Siddiqui , Darian Smalley , Benjamin Smith , Taylor D. Sparks , Daniel T. Speckhard , Elena Stojanovska , Akshay Subramanian , Jiwon Sun , Yunkai Sun , Abdul W. Syed , Souvik Ta , Izumi Takahara , Kelly Tallau , Guannan Tang , Ans B. Tariq , Sui X. Tay , Nurlybek Temirbay , Surya P. Tiwari , Febin Tom , Tajah Trapier , Kasidet J. Trerayapiwat , Samanvya Tripathi , Hawra H. Tuhaifa , Mustafa Unal , Mohammad Uzair , Vallabh Vasudevan , Estefania Vazquez , Victor Venturi , Rahul Verma , Ashwini Verma , Alvaro Vazquez-Mayagoitia , Nicholas Wagner , Araki Wakiuchi , Hao Wan , Liaoyaqi Wang , Wolfgang Wenzel , Alexander Wieczorek , Sze H. Wong , Yue Wu , Tong Xie , Andrew Yi , Ziqi Yin , Jodie A. Yuwono , Nahed A. Zaid , Mohd Zaki , Shehtab Zaman , Maimuna U. Zarewa , Mahtab Zehtab , Baosen Zhang , Wenyu Zhang , Melody Zhang , Yangfan Zhang , Yuwen Zhang , Runze Zhang , Zongmin Zhang , Huanhuan Zhao , Yuanlong Bill Zheng , Ramzi Zidani , Xue Zong , Ian Foster , Ben Blaiszik

Catastrophic forgetting (CF) poses a significant challenge in machine learning, where a model forgets previously learned information upon learning new tasks. Despite the advanced capabilities of Large Language Models (LLMs), they continue…

机器学习 · 计算机科学 2025-04-17 Gangwei Jiang , Caigao Jiang , Zhaoyi Li , Siqiao Xue , Jun Zhou , Linqi Song , Defu Lian , Ying Wei

Large language model (LLM) based agents have recently attracted much attention from the research and industry communities. Compared with original LLMs, LLM-based agents are featured in their self-evolving capability, which is the basis for…

人工智能 · 计算机科学 2024-04-23 Zeyu Zhang , Xiaohe Bo , Chen Ma , Rui Li , Xu Chen , Quanyu Dai , Jieming Zhu , Zhenhua Dong , Ji-Rong Wen