中文
相关论文

相关论文: Reinforcement learning and Bayesian data assimilat…

200 篇论文

Data augmentation (DA) is a crucial technique for enhancing the sample efficiency of visual reinforcement learning (RL) algorithms. Notably, employing simple observation transformations alone can yield outstanding performance without extra…

机器学习 · 计算机科学 2023-10-30 Guozheng Ma , Linrui Zhang , Haoyu Wang , Lu Li , Zilin Wang , Zhen Wang , Li Shen , Xueqian Wang , Dacheng Tao

Background: Phase I trials desire to identify the maximum tolerated dose (MTD) early and proceed quickly to an expansion cohort or phase II trial for efficacy. We propose an early completion method based on multiple dosages to accelerate…

定量方法 · 定量生物学 2021-10-04 Masahiro Kojima

Motivated by the size of cell line drug sensitivity data, researchers have been developing machine learning (ML) models for predicting drug response to advance cancer treatment. As drug sensitivity studies continue generating data, a common…

Medical imaging has revolutionized diagnosis, yet unnecessary procedures are rising, exposing patients to radiation and stress, limiting equitable access, and straining healthcare systems. The American College of Radiology Appropriateness…

定量方法 · 定量生物学 2025-10-08 Anni Tziakouri , Filippo Menolascina

We develop a portfolio allocation framework that leverages deep learning techniques to address challenges arising from high-dimensional, non-stationary, and low-signal-to-noise market information. Our approach includes a dynamic embedding…

投资组合管理 · 定量金融 2025-01-31 Jinghai He , Cheng Hua , Chunyang Zhou , Zeyu Zheng

Many real-world applications require an agent to make robust and deliberate decisions with multimodal information (e.g., robots with multi-sensory inputs). However, it is very challenging to train the agent via reinforcement learning (RL)…

机器学习 · 计算机科学 2023-02-21 Jinming Ma , Feng Wu , Yingfeng Chen , Xianpeng Ji , Yu Ding

Mutual Information (MI) is a crucial measure for capturing dependencies between variables, but exact computation is challenging in high dimensions with intractable likelihoods, impacting accuracy and robustness. One idea is to use an…

机器学习 · 统计学 2025-03-13 Forough Fazeliasl , Michael Minyi Zhang , Bei Jiang , Linglong Kong

Many multi-genic systemic diseases such as neurological disorders, inflammatory diseases, and the majority of cancers do not have effective treatments yet. Reinforcement learning powered systems pharmacology is a potentially effective…

生物大分子 · 定量生物学 2022-02-25 Ryan K. Tan , Yang Liu , Lei Xie

Dose escalation radiotherapy allows increased control of prostate cancer (PCa) but requires segmentation of dominant index lesions (DIL), motivating the development of automated methods for fast, accurate, and consistent segmentation of PCa…

图像与视频处理 · 电气工程与系统科学 2023-03-08 Josiah Simeth , Jue Jiang , Anton Nosov , Andreas Wibmer , Michael Zelefsky , Neelam Tyagi , Harini Veeraraghavan

Accurate parameter estimation in electrochemical battery models is essential for monitoring and assessing the performance of lithium-ion batteries (LiBs). This paper presents a novel approach that combines deep reinforcement learning (DRL)…

系统与控制 · 电气工程与系统科学 2025-06-25 Mehmet Fatih Ozkan , Samuel Filgueira da Silva , Faissal El Idrissi , Prashanth Ramesh , Marcello Canova

Autonomous driving involves multiple, often conflicting objectives such as safety, efficiency, and comfort. In reinforcement learning (RL), these objectives are typically combined through weighted summation, which collapses their relative…

机器人学 · 计算机科学 2026-03-24 Ahmed Abouelazm , Jonas Michel , Daniel Bogdoll , Philip Schörner , J. Marius Zöllner

Dynamic Treatment Regimes (DTRs) provide a systematic approach for making sequential treatment decisions that adapt to individual patient characteristics, particularly in clinical contexts where survival outcomes are of interest.…

机器学习 · 计算机科学 2025-03-11 Animesh Kumar Paul , Russell Greiner

Aims: Combinations of treatments can offer additional benefit over the treatments individually. However, trials of these combinations are lower priority than the development of novel therapies, which can restrict funding, timelines and…

Randomized discontinuation design (RDD) is an enrichment strategy commonly used to address limitations of traditional placebo-controlled trials, particularly the ethical concern of prolonged placebo exposure. RDD consists of two phases: an…

统计方法学 · 统计学 2025-06-03 Ayon Mukherjee , Oleksandr Sverdlov , Ngoc-Thuy Ha , Yu Deng

Performance of deep learning segmentation models is significantly challenged in its transferability across different medical imaging domains, particularly when aiming to adapt these models to a target domain with insufficient annotated data…

图像与视频处理 · 电气工程与系统科学 2024-06-27 Arnaud Judge , Thierry Judge , Nicolas Duchateau , Roman A. Sandler , Joseph Z. Sokol , Olivier Bernard , Pierre-Marc Jodoin

Phase I early-phase clinical studies aim at investigating the safety and the underlying dose-toxicity relationship of a drug or combination. While little may still be known about the compound's properties, it is crucial to consider…

统计方法学 · 统计学 2022-09-13 Christian Röver , Moreno Ursino , Tim Friede , Sarah Zohar

Estimating personalized treatment effects from high-dimensional observational data is essential in situations where experimental designs are infeasible, unethical, or expensive. Existing approaches rely on fitting deep models on outcomes…

机器学习 · 计算机科学 2022-02-02 Andrew Jesson , Panagiotis Tigas , Joost van Amersfoort , Andreas Kirsch , Uri Shalit , Yarin Gal

Prostate cancer diagnosis through MR imaging have currently relied on radiologists' interpretation, whilst modern AI-based methods have been developed to detect clinically significant cancers independent of radiologists. In this study, we…

图像与视频处理 · 电气工程与系统科学 2026-01-09 Xiangcen Wu , Yipei Wang , Qianye Yang , Natasha Thorley , Shonit Punwani , Veeru Kasivisvanathan , Ester Bonmati , Yipeng Hu

Rheumatoid arthritis (RA) is an autoimmune condition caused when patients' immune system mistakenly targets their own tissue. Machine learning (ML) has the potential to identify patterns in patient electronic health records (EHR) to…

定量方法 · 定量生物学 2022-10-25 Shengjia Chen , Nikunj Gupta , Woodward B. Galbraith , Valay Shah , Jacopo Cirrone

This study introduces the P5 model - a foundational method that utilizes reinforcement learning (RL) to augment control, effectiveness, and scalability in molecular dynamics simulations (MD). Our innovative strategy optimizes the sampling…

机器学习 · 计算机科学 2023-07-25 Paloma Gonzalez-Rojas , Andrew Emmel , Luis Martinez , Neil Malur , Gregory Rutledge