English
Related papers

Related papers: How Well Do LLMs Understand Drug Mechanisms? A Kno…

200 papers

Retrosynthesis, the process of breaking down a target molecule into simpler precursors through a series of valid reactions, stands at the core of organic chemistry and drug development. Although recent machine learning (ML) research has…

Artificial Intelligence · Computer Science 2026-05-12 Haorui Wang , Jeff Guo , Lingkai Kong , Rampi Ramprasad , Philippe Schwaller , Yuanqi Du , Chao Zhang

Clinical document classification is essential for converting unstructured medical texts into standardised ICD-10 diagnoses, yet it faces challenges due to complex medical language, privacy constraints, and limited annotated datasets. Large…

Computation and Language · Computer Science 2026-02-03 Akram Mustafa , Usman Naseem , Mostafa Rahimi Azghadi

Large language models (LLMs) are increasingly used in domains where causal reasoning matters, yet it remains unclear whether their judgments reflect normative causal computation, human-like shortcuts, or brittle pattern matching. We…

Artificial Intelligence · Computer Science 2026-03-16 Hanna M. Dettki , Charley M. Wu , Bob Rehder

Recent advancements in Large Language Models (LLMs) have showcased their ability to perform complex reasoning tasks, but their effectiveness in planning remains underexplored. In this study, we evaluate the planning capabilities of OpenAI's…

Artificial Intelligence · Computer Science 2024-10-15 Kevin Wang , Junbo Li , Neel P. Bhatt , Yihan Xi , Qiang Liu , Ufuk Topcu , Zhangyang Wang

Large language models (LLMs) are rapidly changing how researchers in materials science and chemistry discover, organize, and act on scientific knowledge. This paper analyzes a broad set of community-developed LLM applications in an effort…

Materials Science · Physics 2026-05-06 Aritra Roy , Kevin Shen , Andrew MacBride , Awwal Oladipupo , Mudassra Taskeen , Wojtek Treyde , Ruaa A. E. A. Abakar , Ahmad D. Abbas , Elsayed Abdelfatah , Abbas A. Abdullahi , Seham S. Abyah , Chahd Rahyl Adjmi , Fariha Agbere , Savyasanchi Aggarwal , Muhammad Ahmed , Tasnim Ahmed , Motasem Ajlouni , Mattias Akke , Hussein AlAdwan , Anwaar S. Alazani , Zahra A. Alharbi , Wajd A. Aljulyhi , Mohammed A. AlKubaish , Fatima A. Almahri , Sayed A. Almohri , David Obeh Alobo , Mohammed Alouni , Azizah S. Alqahtani , Omar Alsaigh , Husain Althagafi , Md. Aqib Aman , Lena Ara , Arifin , Ignacio Arretche , Abdulaziz Ashy , Syeda A. Asim , Amro Aswad , Adeel Atta , Sören Auer , Abdullah al Azmi , Toheeb Balogun , Suvo Banik , Viktoriia Baibakova , Shakira A. Baksh , Neus G. Bastús , Christina J. Bayard , Adib Bazgir , Louis Beal , Lejla Biberić , Wahid Billah , Ankita Biswas , Joshua Bocarsly , Montassar T. Bouzidi , Esma B. Boydas , Youssef Briki , Cailin Buchanan , Mauricio Cafiero , Damien Caliste , Yi Cao , Rafael E. Castañeda , Sruthy K. Chandy , Benjamin Charmes , Shayantan Chaudhuri , Yiming Chen , Alexander Chen , Jieneng Chen , Min-Hsueh Chiu , Defne Circi , Cinthya H. Contreras , Yoann Cure , Nathan Daelman , Roshini Dantuluri , Thomas Davy , William Dawson , Leonid Didukh , Rui Ding , Aminu R. Doguwa , Claudia Draxl , Sathya Edamadaka , Oulaya Elargab , Christina Ertural , Matthew L. Evans , Edvin Fako , Hossam Farag , Nur A. Fathurrahman , Merve Fedai , Rodrigo P. Ferreira , Giuseppe Fisicaro , Thomas Frank , Sasi K. Gaddipati , Abhijeet Gangan , Jennifer Garland , James Garrick , Luigi Genovese , Maryam Ghadrdran , Sandip Giri , Maxime Goulet , Jeremy Goumaz , Sara U. Gracia , Jacob Graham , Gabriel Graves , Kevin P. Greenman , Tim Greitemeier , Cameron Gruich , Sophie Gu , Salomé Guilbert , Hans Gundlach , Muriel F. Gusta , Mourad El Haddaoui , Alexander J. Haibel , Anubhab Haldar , Vehaan Handa , Hassan Harb , Nathan D. Harms , Abdullah Al Hasan , Abir Hassan , Qiyao He , Andrés Henao-Aristizábal , Bram Hoex , Sungil Hong , Alexander J. Horvath , Md. Shaib Hossain , Yanqi Huang , Yuqing Huang , Kostiantyn Hubaiev , Donald Intal , Katherine Inzani , Kevin Ishimwe , Tugba Isik , Gopal R. Iyer , Katharina Jager , Jan Janssen , Hyewon Jeong , Michael Jirasek , Tyler R. Josephson , Nisarg Joshi , Yassir Ben Kacem , Remya A. M. Kalapurakal , Rakesh R. Kamath , Sugan Kanagasenthinathan , Dohun Kang , Jason Kantorow , Kübra Kaygisiz , Murat Keceli , Farhana Keya , Muhammad U. Khan , Sartaaj Takrim Khan , Hyungjun Kim , Alexander Kister , Sascha Klawohn , Collin Kovacs , Pranav Krishnan , Maurycy Kryzanowski , Ritesh Kumar , Suman Kumari , Gourav Kumbhojkar , Ryo Kuroki , Shashank Kushwaha , Magdalena Lederbauer , Jaejun Lee , Seunghan Lee , Jeonghwan Lee , Bingcan Li , Calvin Li , Zhanzhao Li , Shi Li , Shicheng Li , Chengyan Liu , Hao Liu , Tung Yan Liu , Yutong Liu , Lucia Vina-Lopez , Chayaphol Lortaraparsert , Andre K. Y. Low , Saffron Luxford , Carlos Madariaga , Rishikesh Magar , Piyush R. Maharana , Rahul Mallela , Shoaib Mahmud , Natesan Mani , Umair Mansoor , Omar B. Mansour , Cassandra Masschelein , Kinga O. Mastej , Ankit Mathanker , Jeffrey Meng , Omran Mezghani , Yidong Ming , Rishav Mitra , Michail Mitsakis , Matthew Miyagishima , Ravikumar Mohan , Naveen R. Mohanraj , Trupti Mohanty , Bernadette Mohr , Francisco A. Molina-Bakhos , Jeremy Monat , Seyed Mohamad Moosavi , Shayan Mousavi , Arman Moussavi , Rubel Mozumber , Muhammad J. Mufti , Diyana Muhammed , Ram Munde , Mrigi Munjal , José A. Márquez , Shankha Nag , Giacomo Nagaro , Juno Nam , Jose M. Napoles-Duarte , Ry Nduma , Xuan-Vu Nguyen , Ebrahim Norouzi , Oluwatosin Ohiro , Ryotaro Okabe , Viejay Ordillo , Shuichiro Ozawa , Sebastian Pagel , Daniel Palmer , Angela Pan , Akash Pandey , Vivek Pandit , Prakul Pandit , Chiku Parida , Jaehee Park , Hyunsoo Park , Hemangi Patel , Shakul Pathak , Taradutt Pattnaik , Elena Patyukova , Noah Paulson , Deepak S. Pendyala , Erick S. Pepek , Martin H. Petersen , Thang D. Pham , Aniket Phutane , Sabila K. Pinky , Étienne Polack , Alison Polasik , Maria Politi , Tim Pongratz , Akhila Ponugoti , Fabio Priante , Thomas Michael Pruyn , Sai S. Puppala , Mohammad A. Qazi , Heike Quosdorf , Gollam Rabby , Mohammad J. Raei , Md. Habibur Rahman , A. B. M. Ashikur Rahman , Subhashree Rajasekaran , Tawfiqur Rakib , Hemanth N. Ramesh , Vrushali Ranadive , Karnamohit Ranka , Bojana Rankovic , Adwaith Ravichandran , Ilija Rašović , Sergei Rigin , Tatem Rios , Varun Rishi , Victor Naden Robinson , Lucas S. Rodrigues , Oswaldo Rodriguez , Mahule Roy , Diptendu Roy , Subhas Roy , Arokia Anto Royan M , Joseph F. Rudzinski , Muhammad Sabih , Subramanyam Sahoo , Srusti Bheem Sain , Thahira Saliya , Vignesh Sampath , Jesus Diaz Sanchez , Arthur S. S. Santos , Muliady Satria , Hasan M. Sayeed , Jörg Schaarschmidt , Philippe Schwaller , Nofit Segal , Abhishec Senthilvel , Sherjeel Shabih , Devanshu Shah , Faezeh Shahmoradi , Samiha Sharlin , Killian Sheriff , Qiuyu Shi , Abubakar D. Shuaibu , Ayesha Siddiqua , M. A. Shadab Siddiqui , Darian Smalley , Benjamin Smith , Taylor D. Sparks , Daniel T. Speckhard , Elena Stojanovska , Akshay Subramanian , Jiwon Sun , Yunkai Sun , Abdul W. Syed , Souvik Ta , Izumi Takahara , Kelly Tallau , Guannan Tang , Ans B. Tariq , Sui X. Tay , Nurlybek Temirbay , Surya P. Tiwari , Febin Tom , Tajah Trapier , Kasidet J. Trerayapiwat , Samanvya Tripathi , Hawra H. Tuhaifa , Mustafa Unal , Mohammad Uzair , Vallabh Vasudevan , Estefania Vazquez , Victor Venturi , Rahul Verma , Ashwini Verma , Alvaro Vazquez-Mayagoitia , Nicholas Wagner , Araki Wakiuchi , Hao Wan , Liaoyaqi Wang , Wolfgang Wenzel , Alexander Wieczorek , Sze H. Wong , Yue Wu , Tong Xie , Andrew Yi , Ziqi Yin , Jodie A. Yuwono , Nahed A. Zaid , Mohd Zaki , Shehtab Zaman , Maimuna U. Zarewa , Mahtab Zehtab , Baosen Zhang , Wenyu Zhang , Melody Zhang , Yangfan Zhang , Yuwen Zhang , Runze Zhang , Zongmin Zhang , Huanhuan Zhao , Yuanlong Bill Zheng , Ramzi Zidani , Xue Zong , Ian Foster , Ben Blaiszik

Large language models (LLMs) have shown remarkable reasoning capabilities given chain-of-thought prompts (examples with intermediate reasoning steps). Existing benchmarks measure reasoning ability indirectly, by evaluating accuracy on…

Computation and Language · Computer Science 2023-03-03 Abulhair Saparov , He He

Recent advances in test-time scaling of large language models (LLMs), exemplified by DeepSeek-R1 and OpenAI's o1, show that extending the chain of thought during inference can significantly improve general reasoning performance. However,…

Computation and Language · Computer Science 2025-11-11 Yinghao Hu , Yaoyao Yu , Leilei Gan , Bin Wei , Kun Kuang , Fei Wu

Predicting cancer treatment outcomes requires models that are both accurate and interpretable, particularly in the presence of heterogeneous clinical data. While large language models (LLMs) have shown strong performance in biomedical NLP,…

Computation and Language · Computer Science 2025-10-21 Raghu Vamshi Hemadri , Geetha Krishna Guruju , Kristi Topollai , Anna Ewa Choromanska

Large Language Models (LLMs) have emerged as highly capable systems and are increasingly being integrated into various uses. However, the rapid pace of their deployment has outpaced a comprehensive understanding of their internal mechanisms…

Computation and Language · Computer Science 2025-10-27 Gabriele Prato , Jerry Huang , Prasanna Parthasarathi , Shagun Sodhani , Sarath Chandar

Large language models (LLMs) have demonstrated remarkable proficiency in understanding and generating responses to complex queries through large-scale pre-training. However, the efficacy of these models in memorizing and reasoning among…

Computation and Language · Computer Science 2024-02-23 Qiyuan He , Yizhong Wang , Wenya Wang

Multimodal large language models (MLLMs) are increasingly deployed in open-ended, real-world environments where inputs are messy, underspecified, and not always trustworthy. Unlike curated benchmarks, these settings frequently involve…

Artificial Intelligence · Computer Science 2025-08-26 Qianqi Yan , Hongquan Li , Shan Jiang , Yang Zhao , Xinze Guan , Ching-Chen Kuo , Xin Eric Wang

Large Language Models (LLMs) have demonstrated potential in predicting mental health outcomes from online text, yet traditional classification methods often lack interpretability and robustness. This study evaluates structured reasoning…

Computation and Language · Computer Science 2026-01-09 Avinash Patil , Amardeep Kour Gedhu

Large language models (LLMs) are increasingly being used in a zero-shot fashion to assess mental health conditions, yet we have limited knowledge on what factors affect their accuracy. In this study, we utilize a clinical dataset of natural…

Enabling Large Language Models (LLMs) to handle a wider range of complex tasks (e.g., coding, math) has drawn great attention from many researchers. As LLMs continue to evolve, merely increasing the number of model parameters yields…

The deployment of Large Language Models (LLMs) in mental health counseling faces the dual challenges of hallucinations and lack of empathy. While the former may be mitigated by RAG (retrieval-augmented generation) by anchoring answers in…

Computation and Language · Computer Science 2026-01-06 Md Abdullah Al Kafi , Raka Moni , Sumit Kumar Banshal

Recent progress in Large Language Models (LLMs) has drawn attention to their potential for accelerating drug discovery. However, a central problem remains: translating theoretical ideas into robust implementations in the highly specialized…

Machine Learning · Computer Science 2025-03-06 Sizhe Liu , Yizhou Lu , Siyu Chen , Xiyang Hu , Jieyu Zhao , Yingzhou Lu , Yue Zhao

Large reasoning models (LRMs) like OpenAI o1 and DeepSeek R1 have demonstrated impressive performance on complex reasoning tasks like mathematics and programming with long Chain-of-Thought (CoT) reasoning sequences (slow-thinking), compared…

Artificial Intelligence · Computer Science 2025-07-15 Jason Zhu , Hongyu Li

To what extent can a neural network systematically reason over symbolic facts? Evidence suggests that large pre-trained language models (LMs) acquire some reasoning capacity, but this ability is difficult to control. Recently, it has been…

Computation and Language · Computer Science 2020-11-17 Alon Talmor , Oyvind Tafjord , Peter Clark , Yoav Goldberg , Jonathan Berant

Analogical reasoning -- the capacity to identify and map structural relationships between different domains -- is fundamental to human cognition and learning. Recent studies have shown that large language models (LLMs) can sometimes match…

Computation and Language · Computer Science 2025-11-21 Sam Musker , Alex Duchnowski , Raphaël Millière , Ellie Pavlick

Recent Large Reasoning Models (LRMs), such as DeepSeek-R1 and OpenAI o1, have demonstrated strong performance gains by scaling up the length of Chain-of-Thought (CoT) reasoning during inference. However, a growing concern lies in their…

‹ Prev 1 3 4 5 6 7 10 Next ›