English
Related papers

Related papers: A Guideline-Aware AI Agent for Zero-Shot Target Vo…

200 papers

Large Language Model (LLM) agents deployed for real-world tasks face a fundamental dilemma: user requests are underspecified, yet agents must decide whether to act on incomplete information or interrupt users for clarification. Existing…

Computation and Language · Computer Science 2026-01-13 Yijiang River Dong , Tiancheng Hu , Zheng Hui , Caiqi Zhang , Ivan Vulić , Andreea Bobu , Nigel Collier

Radiation therapy is the mainstay treatment for cervical cancer, and its ultimate goal is to ensure the planning target volume (PTV) reaches the prescribed dose while reducing dose deposition of organs-at-risk (OARs) as much as possible. To…

Computer Vision and Pattern Recognition · Computer Science 2024-08-27 Lu Wen , Wenxia Yin , Zhenghao Feng , Xi Wu , Deng Xiong , Yan Wang

Autonomous agents utilizing Large Language Models (LLMs) have demonstrated remarkable capabilities in isolated medical tasks like diagnosis and image analysis, but struggle with integrated clinical workflows that connect diagnostic…

Artificial Intelligence · Computer Science 2025-10-14 Hongjie Zheng , Zesheng Shi , Ping Yi

FDG PET/CT imaging is a resource intensive examination critical for managing malignant disease and is particularly important for longitudinal assessment during therapy. Approaches to automate longtudinal analysis present many challenges…

Image and Video Processing · Electrical Eng. & Systems 2021-08-05 Anirudh Joshi , Sabri Eyuboglu , Shih-Cheng Huang , Jared Dunnmon , Arjun Soin , Guido Davidzon , Akshay Chaudhari , Matthew P Lungren

Agent evaluation requires assessing complex multi-step behaviors involving tool use and intermediate reasoning, making it costly and expertise-intensive. A natural question arises: can frontier coding assistants reliably automate this…

This paper introduces BioAgent Bench, a benchmark dataset and an evaluation suite designed for measuring the performance and robustness of AI agents in common bioinformatics tasks. The benchmark contains curated end-to-end tasks (e.g.,…

Artificial Intelligence · Computer Science 2026-05-08 Dionizije Fa , Marko Culjak , Bruno Pandza , Mateo Cupic

Adopting Convolutional Neural Networks (CNNs) in the daily routine of primary diagnosis requires not only near-perfect precision, but also a sufficient degree of generalization to data acquisition shifts and transparency. Existing CNN…

Computer Vision and Pattern Recognition · Computer Science 2023-06-22 Mara Graziani , Sebastian Otalora , Stephane Marchand-Maillet , Henning Muller , Vincent Andrearczyk

Brain tumor segmentation is a challenging problem in medical image analysis. The endpoint is to generate the salient masks that accurately identify brain tumor regions in an fMRI screening. In this paper, we propose a novel attention gate…

Image and Video Processing · Electrical Eng. & Systems 2021-07-08 Tim Cvetko

Custom Storyboard Generation (CSG) aims to produce high-quality, multi-character consistent storytelling. Current approaches based on static diffusion models, whether used in a one-shot manner or within multi-agent frameworks, face three…

Computer Vision and Pattern Recognition · Computer Science 2026-02-25 Hailong Yan , Shice Liu , Tao Wang , Xiangtao Zhang , Yijie Zhong , Jinwei Chen , Le Zhang , Bo Li

Amodal completion, generating invisible parts of occluded objects, is vital for applications like image editing and AR. Prior methods face challenges with data needs, generalization, or error accumulation in progressive pipelines. We…

Computer Vision and Pattern Recognition · Computer Science 2025-09-23 Hongxing Fan , Lipeng Wang , Haohua Chen , Zehuan Huang , Jiangtao Wu , Lu Sheng

The selection of appropriate medical imaging procedures is a critical and complex clinical decision, guided by extensive evidence-based standards such as the ACR Appropriateness Criteria (ACR-AC). However, the underutilization of these…

Quantitative Methods · Quantitative Biology 2025-10-07 Satrio Pambudi , Filippo Menolascina

We evaluated the temporal performance of a deep learning (DL) based artificial intelligence (AI) model for auto segmentation in prostate radiotherapy, seeking to correlate its efficacy with changes in clinical landscapes. Our study involved…

Image and Video Processing · Electrical Eng. & Systems 2023-11-17 Biling Wang , Michael Dohopolski , Ti Bai , Junjie Wu , Raquibul Hannan , Neil Desai , Aurelie Garant , Daniel Yang , Dan Nguyen , Mu-Han Lin , Robert Timmerman , Xinlei Wang , Steve Jiang

In order to optimize the radiotherapy delivery for cancer treatment, especially when dealing with complex treatments such as Total Marrow and Lymph Node Irradiation (TMLI), the accurate contouring of the Planning Target Volume (PTV) is…

Computer Vision and Pattern Recognition · Computer Science 2024-02-12 Ricardo Coimbra Brioso , Damiano Dei , Nicola Lambri , Daniele Loiacono , Pietro Mancosu , Marta Scorsetti

Convolutional neural networks (CNNs) have been used quite successfully for semantic segmentation of brain tumors. However, current CNNs and attention mechanisms are stochastic in nature and neglect the morphological indicators used by…

Image and Video Processing · Electrical Eng. & Systems 2021-07-12 Andrea Liew , Chun Cheng Lee , Boon Leong Lan , Maxine Tan

Treatment planning is currently a patient specific, time-consuming, and resource demanding task in radiotherapy. Dose-volume histogram (DVH) prediction plays a critical role in automating this process. The geometric relationship between…

Artificial Intelligence · Computer Science 2024-02-13 Zehao Dong , Yixin Chen , Hiram Gay , Yao Hao , Geoffrey D. Hugo , Pamela Samson , Tianyu Zhao

Reliable identification of anatomical body regions is a prerequisite for many automated medical imaging workflows, yet existing solutions remain heavily dependent on unreliable DICOM metadata. Current solutions mainly use supervised…

Computer Vision and Pattern Recognition · Computer Science 2026-02-10 Farnaz Khun Jush , Grit Werner , Mark Klemens , Matthias Lenga

Large Audio-Language Models (LALMs) perform well on audio understanding tasks but lack multistep reasoning and tool-calling found in recent Large Language Models (LLMs). This paper presents AudioToolAgent, a framework that coordinates…

Sound · Computer Science 2026-02-16 Gijs Wijngaard , Elia Formisano , Michel Dumontier , Jenia Jitsev

Building Large Language Model agents that expand their capabilities by interacting with external tools represents a new frontier in AI research and applications. In this paper, we introduce InfoAgent, a deep research agent powered by an…

Chest computed tomography (CT) is central to the detection and management of thoracic disease, yet the growing scale and complexity of volumetric imaging increasingly exceed what can be addressed by scan-level prediction alone. Clinically…

Computer Vision and Pattern Recognition · Computer Science 2026-04-28 Xuguang Bai , Mingxuan Liu , Tongxi Song , Yifei Chen , Hongjia Yang , Kasidit Anmahapong , Zihan Li , Ying Zhou , Qiyuan Tian

We present QuadAgent, a training-free agent system for agile quadrotor flight guided by vision-language inputs. Unlike prior end-to-end or serial agent approaches, QuadAgent decouples high-level reasoning from low-level control using an…

Robotics · Computer Science 2026-04-06 Ao Zhuang , Feng Yu , Tianbao Zhang , Linzuo Zhang , Danping Zou