English
Related papers

Related papers: PolyReal: A Benchmark for Real-World Polymer Scien…

200 papers

The integration of Multimodal Large Language Models (MLLMs) into chemistry promises to revolutionize scientific discovery, yet their ability to comprehend the dense, graphical language of reactions within authentic literature remains…

Computer Vision and Pattern Recognition · Computer Science 2026-01-29 Hanzheng Li , Xi Fang , Yixuan Li , Chaozheng Huang , Junjie Wang , Xi Wang , Hongzhe Bai , Bojun Hao , Shenyu Lin , Huiqi Liang , Linfeng Zhang , Guolin Ke

Despite impressive advances in large language models (LLMs), existing benchmarks often focus on single-turn or single-step tasks, failing to capture the kind of iterative reasoning required in real-world settings. To address this…

Computation and Language · Computer Science 2025-11-26 Yiran Zhang , Mo Wang , Xiaoyang Li , Kaixuan Ren , Chencheng Zhu , Usman Naseem

Multimodal large language models (MLLMs) hold significant potential in medical applications, including disease diagnosis and clinical decision-making. However, these tasks require highly accurate, context-sensitive, and professionally…

Computation and Language · Computer Science 2025-09-01 Meidan Ding , Jipeng Zhang , Wenxuan Wang , Cheng-Yi Li , Wei-Chieh Fang , Hsin-Yu Wu , Haiqin Zhong , Wenting Chen , Linlin Shen

Large language models (LLMs) have shown remarkable ability in various language tasks, especially with their emergent in-context learning capability. Extending LLMs to incorporate visual inputs, large vision-language models (LVLMs) have…

Machine Learning · Computer Science 2025-10-13 Aneesh Komanduri , Karuna Bhaila , Xintao Wu

Innovation is a key driving force of human civilization. As the body of knowledge has grown considerably, bridging knowledge across different disciplines, where significant innovation often emerges, has become increasingly challenging. The…

Computation and Language · Computer Science 2026-05-07 Yuanhao Shen , Daniel Xavier de Sousa , Ricardo Marçal , Hongyu Guo , Xiaodan Zhu

Large language models (LLMs) have shown potential in assisting scientific research, yet their ability to discover high-quality research hypotheses remains unexamined due to the lack of a dedicated benchmark. To address this gap, we…

Computation and Language · Computer Science 2026-04-21 Yujie Liu , Zonglin Yang , Tong Xie , Jinjie Ni , Ben Gao , Yuqiang Li , Shixiang Tang , Wanli Ouyang , Erik Cambria , Dongzhan Zhou

With the rapid growth of academic publications, peer review has become an essential yet time-consuming responsibility within the research community. Large Language Models (LLMs) have increasingly been adopted to assist in the generation of…

Computation and Language · Computer Science 2025-10-09 Xian Gao , Jiacheng Ruan , Zongyun Zhang , Jingsheng Gao , Ting Liu , Yuzhuo Fu

Recent advances in multimodal large language models (MLLMs) have substantially expanded the capabilities of multimodal retrieval, enabling systems to align and retrieve information across visual and textual modalities. Yet, existing…

Computer Vision and Pattern Recognition · Computer Science 2026-03-03 Xuan Lu , Kangle Li , Haohang Huang , Rui Meng , Wenjun Zeng , Xiaoyu Shen

Multimodal large language models (MLLMs) demonstrate considerable potential in clinical diagnostics, a domain that inherently requires synthesizing complex visual and textual data alongside consulting authoritative medical literature.…

Computation and Language · Computer Science 2026-03-23 Yannian Gu , Zhongzhen Huang , Linjie Mu , Xizhuo Zhang , Shaoting Zhang , Xiaofan Zhang

While existing benchmarks probe the reasoning abilities of large language models (LLMs) across diverse domains, they predominantly assess passive reasoning, providing models with all the information needed to reach a solution. By contrast,…

Machine Learning · Computer Science 2025-06-11 Zhanke Zhou , Xiao Feng , Zhaocheng Zhu , Jiangchao Yao , Sanmi Koyejo , Bo Han

Large language models (LLMs) have made impressive progress in chemistry applications. However, the community lacks an LLM specifically designed for chemistry. The main challenges are two-fold: firstly, most chemical data and scientific…

We introduce MLRC-Bench, a benchmark designed to quantify how effectively language agents can tackle challenging Machine Learning (ML) Research Competitions, with a focus on open research problems that demand novel methodologies. Unlike…

Large language models (LLMs) have become a disruptive force in the industry, introducing unprecedented capabilities in natural language processing, logical reasoning and so on. However, the challenges of knowledge updates and hallucination…

Machine Learning · Computer Science 2025-04-22 Chunjing Gan , Dan Yang , Binbin Hu , Ziqi Liu , Yue Shen , Zhiqiang Zhang , Jian Wang , Jun Zhou

Visual reasoning is central to human cognition, enabling individuals to interpret and abstractly understand their environment. Although recent Multimodal Large Language Models (MLLMs) have demonstrated impressive performance across language…

Computer Vision and Pattern Recognition · Computer Science 2025-03-17 Jing Bi , Junjia Guo , Susan Liang , Guangyu Sun , Luchuan Song , Yunlong Tang , Jinxi He , Jiarui Wu , Ali Vosoughi , Chen Chen , Chenliang Xu

Large Language Models (LLMs) have shown impressive performance in domains such as mathematics and programming, yet their capabilities in physics remain underexplored and poorly understood. Physics poses unique challenges that demand not…

Can large language models predict physical and mechanical polymer properties simply by reading unstructured scientific prose? Polymer performance is rarely determined by chemical structure alone; identical nominal polymers can exhibit…

Machine Learning · Computer Science 2026-05-12 Yuchu Liu , Rui Zhu , Jingwei Xiong , Haixu Tang

Large language models (LLMs) have become increasingly pivotal across various domains, especially in handling complex data types. This includes structured data processing, as exemplified by ChartQA and ChatGPT-Ada, and multimodal…

Developing large-scale foundational datasets is a critical milestone in advancing artificial intelligence (AI)-driven scientific innovation. However, unlike AI-mature fields such as natural language processing, materials science,…

Chemical Physics · Physics 2025-11-18 Ryo Yoshida , Yoshihiro Hayashi , Hidemine Furuya , Ryohei Hosoya , Kazuyoshi Kaneko , Hiroki Sugisawa , Yu Kaneko , Aiko Takahashi , Yoh Noguchi , Shun Nanjo , Keiko Shinoda , Tomu Hamakawa , Mitsuru Ohno , Takuya Kitamura , Misaki Yonekawa , Stephen Wu , Masato Ohnishi , Chang Liu , Teruki Tsurimoto , Arifin , Araki Wakiuchi , Kohei Noda , Junko Morikawa , Teruaki Hayakawa , Junichiro Shiomi , Masanobu Naito , Kazuya Shiratori , Tomoki Nagai , Norio Tomotsu , Hiroto Inoue , Ryuichi Sakashita , Masashi Ishii , Isao Kuwajima , Kenji Furuichi , Norihiko Hiroi , Yuki Takemoto , Takahiro Ohkuma , Keita Yamamoto , Naoya Kowatari , Masato Suzuki , Naoya Matsumoto , Seiryu Umetani , Hisaki Ikebata , Yasuyuki Shudo , Mayu Nagao , Shinya Kamada , Kazunori Kamio , Taichi Shomura , Kensaku Nakamura , Yudai Iwamizu , Atsutoshi Abe , Koki Yoshitomi , Yuki Horie , Katsuhiko Koike , Koichi Iwakabe , Shinya Gima , Kota Usui , Gikyo Usuki , Takuro Tsutsumi , Keitaro Matsuoka , Kazuki Sada , Masahiro Kitabata , Takuma Kikutsuji , Akitaka Kamauchi , Yusuke Iijima , Tsubasa Suzuki , Takenori Goda , Yuki Takabayashi , Kazuko Imai , Yuji Mochizuki , Hideo Doi , Koji Okuwaki , Hiroya Nitta , Taku Ozawa , Hitoshi Kamijima , Toshiaki Shintani , Takuma Mitamura , Massimiliano Zamengo , Yuitsu Sugami , Seiji Akiyama , Yoshinari Murakami , Atsushi Betto , Naoya Matsuo , Satoru Kagao , Tetsuya Kobayashi , Norie Matsubara , Shosei Kubo , Yuki Ishiyama , Yuri Ichioka , Mamoru Usami , Satoru Yoshizaki , Seigo Mizutani , Yosuke Hanawa , Shogo Kunieda , Mitsuru Yambe , Takeru Nakamura , Hiromori Murashima , Kenji Takahashi , Naoki Wada , Masahiro Kawano , Yosuke Harada , Takehiro Fujita , Erina Fujita , Ryoji Himeno , Hiori Kino , Kenji Fukumizu

While Multimodal Large Language Models (MLLMs) have achieved impressive performance on semantic tasks, their spatial intelligence--crucial for robust and grounded AI systems--remains underdeveloped. Existing benchmarks fall short of…

Computer Vision and Pattern Recognition · Computer Science 2025-12-30 Mingrui Wu , Zhaozhi Wang , Fangjinhua Wang , Jiaolong Yang , Marc Pollefeys , Tong Zhang