English
Related papers

Related papers: Quantifying Self-diagnostic Atomic Knowledge in Ch…

200 papers

Foundation models (FMs) are transforming computational pathology by offering new ways to analyze histopathology images. However, FMs typically require weeks of training on large databases, making their creation a resource-intensive process.…

Image and Video Processing · Electrical Eng. & Systems 2026-01-27 Till Nicke , Daniela Schacherer , Jan Raphael Schäfer , Natalia Artysh , Antje Prasse , André Homeyer , Andrea Schenk , Henning Höfener , Johannes Lotz

Current open-source training pipelines for Chinese medical language models predominantly emphasize optimizing training methodologies to enhance the performance of large language models (LLMs), yet lack comprehensive exploration into…

Machine Learning · Computer Science 2025-09-03 Wei Huang , Anda Cheng , Zhao Zhang , Yinggui Wang

Molecular property prediction integrates quantum chemistry, cheminformatics, and deep learning to connect molecular structure with physicochemical and biological behavior. This survey traces four complementary paradigms, including Quantum,…

Synthetic data has become essential for training foundation models, yet benchmark contamination threatens evaluation integrity. Although existing detection methods identify token-level overlap, they fail to detect semantic-level…

Machine Learning · Computer Science 2025-11-25 Sushant Mehta

Automated segmentation is a fundamental medical image analysis task, which enjoys significant advances due to the advent of deep learning. While foundation models have been useful in natural language processing and some vision tasks for…

Computer Vision and Pattern Recognition · Computer Science 2025-05-12 Hanxue Gu , Haoyu Dong , Jichen Yang , Maciej A. Mazurowski

Recent advancements in large language models have significantly accelerated their adoption in healthcare applications, including AI-powered medical consultations, diagnostic report assistance, and medical search tools. However, medical…

Clinical deployment of automated brain MRI analysis faces a fundamental challenge: clinical data is heterogeneous and noisy, and high-quality labels are prohibitively costly to obtain. Self-supervised learning (SSL) can address this by…

Computer Vision and Pattern Recognition · Computer Science 2026-05-25 Asbjørn Munk , Stefano Cerri , Vardan Nersesjan , Christian Hedeager Krag , Jakob Ambsdorf , Pablo Rocamora García , Julia Machnio , Peirong Liu , Suhyun Ahn , Nasrin Akbari , Yasmina Al Khalil , Kimberly Amador , Sina Amirrajab , Tal Arbel , Meritxell Bach Cuadra , Ujjwal Baid , Bhakti Baheti , Jaume Banus , Kamil Barbierik , Christoph Brune , Yansong Bu , Baptiste Callard , Yuhan Chen , Cornelius Crijnen , Corentin Dancette , Peter Drotar , Prasad Dutande , Nils D. Forkert , Saurabh Garg , Jakub Gazda , Matej Gazda , Benoît Gérin , Partha Ghosh , Weikang Gong , Pedro M. Gordaliza , Sam Hashemi , Tobias Heimann , Fucang Jia , Jiexin Jiang , Emily Kaczmarek , Chris Kang , Seung Kwan Kang , Mohammad Khazaei , Julien Khlaut , Petros Koutsouvelis , Jae Sung Lee , Yuchong Li , Mengye Lyu , Mingchen Ma , Anant Madabhushi , Klaus H. Maier-Hein , Pierre Manceron , Andrés Martínez Mora , Moona Mazher , Felix Meister , Nataliia Molchanova , Steven A. Niederer , Leonard Nürnberg , Jinah Park , Abdul Qayyum , Jonas Richiardi , Antoine Saporta , Branislav Setlak , Ning Shen , Justin Szeto , Constantin Ulrich , Puru Vaish , Vibujithan Vigneshwaran , Leroy Volmer , Zihao Wang , Siqi Wei , Anthony Winder , Jelmer M. Wolterink , Maxence Wynen , Chang Yang , Si Young Yie , Mostafa Mehdipour Ghazi , Akshay Pai , Espen Jimenez Solem , Sebastian Nørgaard Llambias , Mikael Boesen , Michael Eriksen Benros , Juan Eugenio Iglesias , Mads Nielsen

Quantization-aware training (QAT) and Knowledge Distillation (KD) are combined to achieve competitive performance in creating low-bit deep learning models. However, existing works applying KD to QAT require tedious hyper-parameter tuning to…

Machine Learning · Computer Science 2024-03-19 Kaiqi Zhao , Ming Zhao

Materials with bespoke properties have long been identified by computational searches, and their experimental realisation is now coming within reach through autonomous laboratories. Scattering experiments are central to verifying the atomic…

Closed-source frontier labs do not disclose parameter counts, and the standard alternative -- inference economics -- carries $2\times$+ uncertainty from hardware, batching, and serving-stack assumptions external to the model. We exploit a…

Machine Learning · Computer Science 2026-04-29 Bojie Li

Traditional Chinese Medicine (TCM) is a holistic medical system with millennia of accumulated clinical experience, playing a vital role in global healthcare-particularly across East Asia. However, the implicit reasoning, diverse textual…

Computation and Language · Computer Science 2025-06-03 Shufeng Kong , Xingru Yang , Yuanyuan Wei , Zijie Wang , Hao Tang , Jiuqi Qin , Shuting Lan , Yingheng Wang , Junwen Bai , Zhuangbin Chen , Zibin Zheng , Caihua Liu , Hao Liang

The rapid advancement of Chinese LLMs underscores the need for vertical-domain evaluations to ensure reliable applications. However, existing benchmarks often lack domain coverage and provide limited insights into the Chinese working…

Computation and Language · Computer Science 2025-09-04 Mengze Hong , Wailing Ng , Chen Jason Zhang , Di Jiang

Recent advances in reasoning-enhanced Large Language Models such as OpenAI-o1/3 and DeepSeek-R1 have significantly improved performance on complex tasks. However, the quality and transparency of their internal reasoning processes remain…

Computation and Language · Computer Science 2025-06-04 Juncheng Wu , Sheng Liu , Haoqin Tu , Hang Yu , Xiaoke Huang , James Zou , Cihang Xie , Yuyin Zhou

Atomistic simulations of matter, especially those that leverage first-principles (ab initio) electronic structure theory, provide a microscopic view of the world, underpinning much of our understanding of chemistry and materials science.…

Chemical Physics · Physics 2025-09-08 Ilyes Batatia , Philipp Benner , Yuan Chiang , Alin M. Elena , Dávid P. Kovács , Janosh Riebesell , Xavier R. Advincula , Mark Asta , Matthew Avaylon , William J. Baldwin , Fabian Berger , Noam Bernstein , Arghya Bhowmik , Filippo Bigi , Samuel M. Blau , Vlad Cărare , Michele Ceriotti , Sanggyu Chong , James P. Darby , Sandip De , Flaviano Della Pia , Volker L. Deringer , Rokas Elijošius , Zakariya El-Machachi , Fabio Falcioni , Edvin Fako , Andrea C. Ferrari , John L. A. Gardner , Mikolaj J. Gawkowski , Annalena Genreith-Schriever , Janine George , Rhys E. A. Goodall , Jonas Grandel , Clare P. Grey , Petr Grigorev , Shuang Han , Will Handley , Hendrik H. Heenen , Kersti Hermansson , Christian Holm , Cheuk Hin Ho , Stephan Hofmann , Jad Jaafar , Konstantin S. Jakob , Hyunwook Jung , Venkat Kapil , Aaron D. Kaplan , Nima Karimitari , James R. Kermode , Panagiotis Kourtis , Namu Kroupa , Jolla Kullgren , Matthew C. Kuner , Domantas Kuryla , Guoda Liepuoniute , Chen Lin , Johannes T. Margraf , Ioan-Bogdan Magdău , Angelos Michaelides , J. Harry Moore , Aakash A. Naik , Samuel P. Niblett , Sam Walton Norwood , Niamh O'Neill , Christoph Ortner , Kristin A. Persson , Karsten Reuter , Andrew S. Rosen , Louise A. M. Rosset , Lars L. Schaaf , Christoph Schran , Benjamin X. Shi , Eric Sivonxay , Tamás K. Stenczel , Viktor Svahn , Christopher Sutton , Thomas D. Swinburne , Jules Tilly , Cas van der Oord , Santiago Vargas , Eszter Varga-Umbrich , Tejs Vegge , Martin Vondrák , Yangshuai Wang , William C. Witt , Thomas Wolf , Fabian Zills , Gábor Csányi

Although data-driven methods usually have noticeable performance on disease diagnosis and treatment, they are suspected of leakage of privacy due to collecting data for model training. Recently, federated learning provides a secure and…

Artificial Intelligence · Computer Science 2023-06-27 Yawei Zhao , Qinghe Liu , Xinwang Liu , Kunlun He

LLMs are increasingly used as ``digital consumers'' to simulate public opinion, pre-test marketing decisions, and anticipate audience response. However, existing evaluations rarely ask whether a model can reconstruct the concrete reaction…

Computation and Language · Computer Science 2026-05-19 Tianyu Wang , Jiajun Li , Jianghao Lin

The evaluation and improvement of medical large language models (LLMs) are critical for their real-world deployment, particularly in ensuring accuracy, safety, and ethical alignment. Existing frameworks inadequately dissect domain-specific…

Computation and Language · Computer Science 2025-03-11 Luyi Jiang , Jiayuan Chen , Lu Lu , Xinwei Peng , Lihao Liu , Junjun He , Jie Xu

The clinical diagnosis of most mental disorders primarily relies on the conversations between psychiatrist and patient. The creation of such diagnostic conversation datasets is promising to boost the AI mental healthcare community. However,…

Computation and Language · Computer Science 2024-12-30 Congchi Yin , Feng Li , Shu Zhang , Zike Wang , Jun Shao , Piji Li , Jianhua Chen , Xun Jiang

The integration of artificial intelligence (AI) in medical diagnostics represents a significant advancement in managing upper gastrointestinal (GI) cancer, a major cause of global cancer mortality. Specifically for gastric cancer (GC),…

Predicting transcriptional responses to novel drugs provides a unique opportunity to accelerate biomedical research and advance drug discovery efforts. However, the inherent complexity and high dimensionality of cellular responses, combined…