English
Related papers

Related papers: LLM Benchmark-User Need Misalignment for Climate C…

200 papers

With the increasing impacts of climate change, there is a growing demand for accessible tools that can provide reliable future climate information to support planning, finance, and other decision-making applications. Large language models…

Machine Learning · Computer Science 2024-11-22 Yang Wang , Hassan A. Karimi

Large language models (LLMs) exhibit superior performance on various natural language tasks, but they are susceptible to issues stemming from outdated data and domain-specific limitations. In order to address these challenges, researchers…

Computation and Language · Computer Science 2024-10-24 Zhangyin Feng , Weitao Ma , Weijiang Yu , Lei Huang , Haotian Wang , Qianglong Chen , Weihua Peng , Xiaocheng Feng , Bing Qin , Ting liu

Large language models (LLMs) are rapidly changing how researchers in materials science and chemistry discover, organize, and act on scientific knowledge. This paper analyzes a broad set of community-developed LLM applications in an effort…

Materials Science · Physics 2026-05-06 Aritra Roy , Kevin Shen , Andrew MacBride , Awwal Oladipupo , Mudassra Taskeen , Wojtek Treyde , Ruaa A. E. A. Abakar , Ahmad D. Abbas , Elsayed Abdelfatah , Abbas A. Abdullahi , Seham S. Abyah , Chahd Rahyl Adjmi , Fariha Agbere , Savyasanchi Aggarwal , Muhammad Ahmed , Tasnim Ahmed , Motasem Ajlouni , Mattias Akke , Hussein AlAdwan , Anwaar S. Alazani , Zahra A. Alharbi , Wajd A. Aljulyhi , Mohammed A. AlKubaish , Fatima A. Almahri , Sayed A. Almohri , David Obeh Alobo , Mohammed Alouni , Azizah S. Alqahtani , Omar Alsaigh , Husain Althagafi , Md. Aqib Aman , Lena Ara , Arifin , Ignacio Arretche , Abdulaziz Ashy , Syeda A. Asim , Amro Aswad , Adeel Atta , Sören Auer , Abdullah al Azmi , Toheeb Balogun , Suvo Banik , Viktoriia Baibakova , Shakira A. Baksh , Neus G. Bastús , Christina J. Bayard , Adib Bazgir , Louis Beal , Lejla Biberić , Wahid Billah , Ankita Biswas , Joshua Bocarsly , Montassar T. Bouzidi , Esma B. Boydas , Youssef Briki , Cailin Buchanan , Mauricio Cafiero , Damien Caliste , Yi Cao , Rafael E. Castañeda , Sruthy K. Chandy , Benjamin Charmes , Shayantan Chaudhuri , Yiming Chen , Alexander Chen , Jieneng Chen , Min-Hsueh Chiu , Defne Circi , Cinthya H. Contreras , Yoann Cure , Nathan Daelman , Roshini Dantuluri , Thomas Davy , William Dawson , Leonid Didukh , Rui Ding , Aminu R. Doguwa , Claudia Draxl , Sathya Edamadaka , Oulaya Elargab , Christina Ertural , Matthew L. Evans , Edvin Fako , Hossam Farag , Nur A. Fathurrahman , Merve Fedai , Rodrigo P. Ferreira , Giuseppe Fisicaro , Thomas Frank , Sasi K. Gaddipati , Abhijeet Gangan , Jennifer Garland , James Garrick , Luigi Genovese , Maryam Ghadrdran , Sandip Giri , Maxime Goulet , Jeremy Goumaz , Sara U. Gracia , Jacob Graham , Gabriel Graves , Kevin P. Greenman , Tim Greitemeier , Cameron Gruich , Sophie Gu , Salomé Guilbert , Hans Gundlach , Muriel F. Gusta , Mourad El Haddaoui , Alexander J. Haibel , Anubhab Haldar , Vehaan Handa , Hassan Harb , Nathan D. Harms , Abdullah Al Hasan , Abir Hassan , Qiyao He , Andrés Henao-Aristizábal , Bram Hoex , Sungil Hong , Alexander J. Horvath , Md. Shaib Hossain , Yanqi Huang , Yuqing Huang , Kostiantyn Hubaiev , Donald Intal , Katherine Inzani , Kevin Ishimwe , Tugba Isik , Gopal R. Iyer , Katharina Jager , Jan Janssen , Hyewon Jeong , Michael Jirasek , Tyler R. Josephson , Nisarg Joshi , Yassir Ben Kacem , Remya A. M. Kalapurakal , Rakesh R. Kamath , Sugan Kanagasenthinathan , Dohun Kang , Jason Kantorow , Kübra Kaygisiz , Murat Keceli , Farhana Keya , Muhammad U. Khan , Sartaaj Takrim Khan , Hyungjun Kim , Alexander Kister , Sascha Klawohn , Collin Kovacs , Pranav Krishnan , Maurycy Kryzanowski , Ritesh Kumar , Suman Kumari , Gourav Kumbhojkar , Ryo Kuroki , Shashank Kushwaha , Magdalena Lederbauer , Jaejun Lee , Seunghan Lee , Jeonghwan Lee , Bingcan Li , Calvin Li , Zhanzhao Li , Shi Li , Shicheng Li , Chengyan Liu , Hao Liu , Tung Yan Liu , Yutong Liu , Lucia Vina-Lopez , Chayaphol Lortaraparsert , Andre K. Y. Low , Saffron Luxford , Carlos Madariaga , Rishikesh Magar , Piyush R. Maharana , Rahul Mallela , Shoaib Mahmud , Natesan Mani , Umair Mansoor , Omar B. Mansour , Cassandra Masschelein , Kinga O. Mastej , Ankit Mathanker , Jeffrey Meng , Omran Mezghani , Yidong Ming , Rishav Mitra , Michail Mitsakis , Matthew Miyagishima , Ravikumar Mohan , Naveen R. Mohanraj , Trupti Mohanty , Bernadette Mohr , Francisco A. Molina-Bakhos , Jeremy Monat , Seyed Mohamad Moosavi , Shayan Mousavi , Arman Moussavi , Rubel Mozumber , Muhammad J. Mufti , Diyana Muhammed , Ram Munde , Mrigi Munjal , José A. Márquez , Shankha Nag , Giacomo Nagaro , Juno Nam , Jose M. Napoles-Duarte , Ry Nduma , Xuan-Vu Nguyen , Ebrahim Norouzi , Oluwatosin Ohiro , Ryotaro Okabe , Viejay Ordillo , Shuichiro Ozawa , Sebastian Pagel , Daniel Palmer , Angela Pan , Akash Pandey , Vivek Pandit , Prakul Pandit , Chiku Parida , Jaehee Park , Hyunsoo Park , Hemangi Patel , Shakul Pathak , Taradutt Pattnaik , Elena Patyukova , Noah Paulson , Deepak S. Pendyala , Erick S. Pepek , Martin H. Petersen , Thang D. Pham , Aniket Phutane , Sabila K. Pinky , Étienne Polack , Alison Polasik , Maria Politi , Tim Pongratz , Akhila Ponugoti , Fabio Priante , Thomas Michael Pruyn , Sai S. Puppala , Mohammad A. Qazi , Heike Quosdorf , Gollam Rabby , Mohammad J. Raei , Md. Habibur Rahman , A. B. M. Ashikur Rahman , Subhashree Rajasekaran , Tawfiqur Rakib , Hemanth N. Ramesh , Vrushali Ranadive , Karnamohit Ranka , Bojana Rankovic , Adwaith Ravichandran , Ilija Rašović , Sergei Rigin , Tatem Rios , Varun Rishi , Victor Naden Robinson , Lucas S. Rodrigues , Oswaldo Rodriguez , Mahule Roy , Diptendu Roy , Subhas Roy , Arokia Anto Royan M , Joseph F. Rudzinski , Muhammad Sabih , Subramanyam Sahoo , Srusti Bheem Sain , Thahira Saliya , Vignesh Sampath , Jesus Diaz Sanchez , Arthur S. S. Santos , Muliady Satria , Hasan M. Sayeed , Jörg Schaarschmidt , Philippe Schwaller , Nofit Segal , Abhishec Senthilvel , Sherjeel Shabih , Devanshu Shah , Faezeh Shahmoradi , Samiha Sharlin , Killian Sheriff , Qiuyu Shi , Abubakar D. Shuaibu , Ayesha Siddiqua , M. A. Shadab Siddiqui , Darian Smalley , Benjamin Smith , Taylor D. Sparks , Daniel T. Speckhard , Elena Stojanovska , Akshay Subramanian , Jiwon Sun , Yunkai Sun , Abdul W. Syed , Souvik Ta , Izumi Takahara , Kelly Tallau , Guannan Tang , Ans B. Tariq , Sui X. Tay , Nurlybek Temirbay , Surya P. Tiwari , Febin Tom , Tajah Trapier , Kasidet J. Trerayapiwat , Samanvya Tripathi , Hawra H. Tuhaifa , Mustafa Unal , Mohammad Uzair , Vallabh Vasudevan , Estefania Vazquez , Victor Venturi , Rahul Verma , Ashwini Verma , Alvaro Vazquez-Mayagoitia , Nicholas Wagner , Araki Wakiuchi , Hao Wan , Liaoyaqi Wang , Wolfgang Wenzel , Alexander Wieczorek , Sze H. Wong , Yue Wu , Tong Xie , Andrew Yi , Ziqi Yin , Jodie A. Yuwono , Nahed A. Zaid , Mohd Zaki , Shehtab Zaman , Maimuna U. Zarewa , Mahtab Zehtab , Baosen Zhang , Wenyu Zhang , Melody Zhang , Yangfan Zhang , Yuwen Zhang , Runze Zhang , Zongmin Zhang , Huanhuan Zhao , Yuanlong Bill Zheng , Ramzi Zidani , Xue Zong , Ian Foster , Ben Blaiszik

This paper delves into the dynamic landscape of artificial intelligence, specifically focusing on the burgeoning prominence of large language models (LLMs). We underscore the pivotal role of Reinforcement Learning from Human Feedback (RLHF)…

Computers and Society · Computer Science 2024-03-18 Dana Alsagheer , Rabimba Karanjai , Nour Diallo , Weidong Shi , Yang Lu , Suha Beydoun , Qiaoning Zhang

Large Language Models (LLMs) have introduced a paradigm shift in interaction with AI technology, enabling knowledge workers to complete tasks by specifying their desired outcome in natural language. LLMs have the potential to increase…

Human-Computer Interaction · Computer Science 2025-03-24 Michelle Brachman , Amina El-Ashry , Casey Dugan , Werner Geyer

Large language models (LLMs), a recent advance in deep learning and machine intelligence, have manifested astonishing capacities, now considered among the most promising for artificial general intelligence. With human-like capabilities,…

Artificial Intelligence · Computer Science 2025-09-19 Zhilun Zhou , Jing Yi Wang , Nicholas Sukiennik , Chen Gao , Fengli Xu , Yong Li , James Evans

Behavioral simulation is increasingly used to anticipate responses to interventions. Large language models (LLMs) enable researchers to specify population characteristics and intervention context in natural language, but it remains unclear…

Computers and Society · Computer Science 2026-04-14 Zonghan Li , Feng Ji

Large language models (LLMs) are increasingly used to simulate human behavior in experimental settings, but they systematically diverge from human decisions in complex decision-making environments, where participants must anticipate others'…

Artificial Intelligence · Computer Science 2026-01-06 Letian Kong , Qianran , Jin , Renyu Zhang

Climate disinformation has become a major challenge in today digital world, especially with the rise of misleading images and videos shared widely on social media. These false claims are often convincing and difficult to detect, which can…

Artificial Intelligence · Computer Science 2026-01-23 Marzieh Adeli Shamsabad , Hamed Ghodrati

Large language models (LLMs) are increasingly used to predict human behavior. We propose a measure for evaluating how much knowledge a pretrained LLM brings to such a prediction: its equivalent sample size, defined as the amount of…

Econometrics · Economics 2026-01-21 Wayne Gao , Sukjin Han , Annie Liang

This paper introduces the Word Synchronization Challenge, a novel benchmark to evaluate large language models (LLMs) in Human-Computer Interaction (HCI). This benchmark uses a dynamic game-like framework to test LLMs ability to mimic human…

Human-Computer Interaction · Computer Science 2026-01-15 Tanguy Cazalets , Joni Dambre

The recent breakthroughs in the research on Large Language Models (LLMs) have triggered a transformation across several research domains. Notably, the integration of LLMs has greatly enhanced performance in robot Task And Motion Planning…

Robotics · Computer Science 2024-06-12 Yuchen Liu , Luigi Palmieri , Sebastian Koch , Ilche Georgievski , Marco Aiello

A long-standing challenge in developing accurate recommendation models is simulating user behavior, mainly due to the complex and stochastic nature of user interactions. Towards this, one promising line of work has been the use of Large…

Information Retrieval · Computer Science 2025-09-15 Himanshu Thakur , Eshani Agrawal , Smruthi Mukund

An essential problem in artificial intelligence is whether LLMs can simulate human cognition or merely imitate surface-level behaviors, while existing datasets suffer from either synthetic reasoning traces or population-level aggregation,…

Computation and Language · Computer Science 2026-03-31 Yuxuan Gu , Lunjun Liu , Xiaocheng Feng , Kun Zhu , Weihong Zhong , Lei Huang , Bing Qin

Large language models (LLMs) are increasingly used for academic expert recommendation. Existing audits typically evaluate model outputs in isolation, largely ignoring end-user inference-time interventions. As a result, it remains unclear…

Information Retrieval · Computer Science 2026-02-10 Lisette Espin-Noboa , Gonzalo Gabriel Mendez

Human-produced emissions are growing at an alarming rate, causing already observable changes in the climate and environment in general. Each year global carbon dioxide emissions hit a new record, and it is reported that 0.5% of total US…

Computers and Society · Computer Science 2024-08-06 Aida Usmanova , Junbo Huang , Debayan Banerjee , Ricardo Usbeck

Accurately simulating human opinion dynamics is crucial for understanding a variety of societal phenomena, including polarization and the spread of misinformation. However, the agent-based models (ABMs) commonly used for such simulations…

This paper examines the challenges associated with achieving life-long superalignment in AI systems, particularly large language models (LLMs). Superalignment is a theoretical framework that aspires to ensure that superintelligent AI…

Computers and Society · Computer Science 2024-03-25 Gokul Puthumanaillam , Manav Vora , Pranay Thangeda , Melkior Ornik

Large Language Models (LLMs) have emerged as personalized assistants for users across a wide range of tasks -- from offering writing support to delivering tailored recommendations or consultations. Over time, the interaction history between…

Computation and Language · Computer Science 2025-10-28 Bowen Jiang , Zhuoqun Hao , Young-Min Cho , Bryan Li , Yuan Yuan , Sihao Chen , Lyle Ungar , Camillo J. Taylor , Dan Roth

In recent years, knowledge graphs have been integrated into recommender systems as item-side auxiliary information, enhancing recommendation accuracy. However, constructing and integrating structural user-side knowledge remains a…

Information Retrieval · Computer Science 2024-12-19 Zheng Hu , Zhe Li , Ziyun Jiao , Satoshi Nakagawa , Jiawen Deng , Shimin Cai , Tao Zhou , Fuji Ren