English
Related papers

Related papers: MedGemma 1.5 Technical Report

200 papers

Medical artificial general intelligence (AGI) is an emerging field that aims to develop systems specifically designed for medical applications that possess the ability to understand, learn, and apply knowledge across a wide range of tasks…

Artificial Intelligence · Computer Science 2023-06-21 Juexiao Zhou , Xiuying Chen , Xin Gao

SAM 3D Body (3DB) achieves state-of-the-art accuracy in monocular 3D human mesh recovery, yet its inference latency of several seconds per image precludes real-time application. We present Fast SAM 3D Body, a training-free acceleration…

Computer Vision and Pattern Recognition · Computer Science 2026-03-17 Timing Yang , Sicheng He , Hongyi Jing , Jiawei Yang , Zhijian Liu , Chuhang Zou , Yue Wang

Biomedical multimodal assistants have the potential to unify radiology, pathology, and clinical-text reasoning, yet a critical deployment gap remains: top-performing systems are either closed-source or computationally prohibitive,…

Computation and Language · Computer Science 2026-03-03 Kai Zhang , Zhengqing Yuan , Cheng Peng , Songlin Zhao , Mengxian Lyu , Ziyi Chen , Yanfang Ye , Wei Liu , Ying Zhang , Kaleb E Smith , Lifang He , Lichao Sun , Yonghui Wu

Driven by the large foundation models, the development of artificial intelligence has witnessed tremendous progress lately, leading to a surge of general interest from the public. In this study, we aim to assess the performance of OpenAI's…

Computer Vision and Pattern Recognition · Computer Science 2023-12-05 Chaoyi Wu , Jiayu Lei , Qiaoyu Zheng , Weike Zhao , Weixiong Lin , Xiaoman Zhang , Xiao Zhou , Ziheng Zhao , Ya Zhang , Yanfeng Wang , Weidi Xie

High computation costs and latency of large language models such as GPT-4 have limited their deployment in clinical settings. Small language models (SLMs) offer a cost-effective alternative, but their limited capacity requires biomedical…

Chest X-ray images are commonly used for predicting acute and chronic cardiopulmonary conditions, but efforts to integrate them with structured clinical data face challenges due to incomplete electronic health records (EHR). This paper…

Computer Vision and Pattern Recognition · Computer Science 2025-04-17 Mai A. Shaaban , Adnan Khan , Mohammad Yaqub

The rapid advancement of generative AI in medical imaging has introduced both significant opportunities and serious challenges, especially the risk that fake medical images could undermine healthcare systems. These synthetic images pose…

Computer Vision and Pattern Recognition · Computer Science 2025-09-22 Shuaibo Li , Zhaohu Xing , Hongqiu Wang , Pengfei Hao , Xingyu Li , Zekai Liu , Lei Zhu

Recent advances in interactive 3D segmentation from 2D images have demonstrated impressive performance. However, current models typically require extensive scene-specific training to accurately reconstruct and segment objects, which limits…

Computer Vision and Pattern Recognition · Computer Science 2025-03-18 Yansong Guo , Jie Hu , Yansong Qu , Liujuan Cao

Medical Visual Question Answering (Med-VQA) holds significant potential for clinical decision support, yet existing efforts primarily focus on 2D imaging with limited task diversity. This paper presents 3D-RAD, a large-scale dataset…

Computer Vision and Pattern Recognition · Computer Science 2025-10-28 Xiaotang Gai , Jiaxiang Liu , Yichen Li , Zijie Meng , Jian Wu , Zuozhu Liu

Intracranial aneurysms pose a significant clinical risk yet are difficult to detect, delineate and model due to limited annotated 3D data. We propose a cross-domain feature-transfer approach that leverages the latent geometric embeddings…

Computer Vision and Pattern Recognition · Computer Science 2025-09-04 Clément Hervé , Paul Garnier , Jonathan Viquerat , Elie Hachem

We introduce the world's first clinical terminology for the Chinese healthcare community, namely MedCT, accompanied by a clinical foundation model MedBERT and an entity linking model MedLink. The MedCT system enables standardized and…

Computation and Language · Computer Science 2025-04-11 Ye Chen , Dongdong Huang , Haoyun Xu , Cong Fu , Lin Sheng , Qingli Zhou , Yuqiang Shen , Kai Wang

Rapid advances in medical imaging technology underscore the critical need for precise and automated image quality assessment (IQA) to ensure diagnostic accuracy. Existing medical IQA methods, however, struggle to generalize across diverse…

Computer Vision and Pattern Recognition · Computer Science 2025-07-28 Siyi Xun , Yue Sun , Jingkun Chen , Zitong Yu , Tong Tong , Xiaohong Liu , Mingxiang Wu , Tao Tan

Over the past few years, the rapid development of deep learning technologies for computer vision has significantly improved the performance of medical image segmentation (MedISeg). However, the diverse implementation strategies of various…

Computer Vision and Pattern Recognition · Computer Science 2023-05-09 Dong Zhang , Yi Lin , Hao Chen , Zhuotao Tian , Xin Yang , Jinhui Tang , Kwang Ting Cheng

The advancement of artificial intelligence (AI) for organ segmentation and tumor detection is propelled by the growing availability of computed tomography (CT) datasets with detailed, per-voxel annotations. However, these AI models often…

Image and Video Processing · Electrical Eng. & Systems 2024-05-29 Jie Liu , Yixiao Zhang , Kang Wang , Mehmet Can Yavuz , Xiaoxi Chen , Yixuan Yuan , Haoliang Li , Yang Yang , Alan Yuille , Yucheng Tang , Zongwei Zhou

We present MeFEm, a vision model based on a modified Joint Embedding Predictive Architecture (JEPA) for biometric and medical analysis from facial images. Key modifications include an axial stripe masking strategy to focus learning on…

Computer Vision and Pattern Recognition · Computer Science 2026-02-17 Yury Borets , Stepan Botman

Accurate sarcopenia diagnosis via ultrasound remains challenging due to subtle imaging cues, limited labeled data, and the absence of clinical context in most models. We propose MedVQA-TREE, a multimodal framework that integrates a…

Image and Video Processing · Electrical Eng. & Systems 2025-08-28 Pardis Moradbeiki , Nasser Ghadiri , Sayed Jalal Zahabi , Uffe Kock Wiil , Kristoffer Kittelmann Brockhattingen , Ali Ebrahimi

Leveraging the Segment Anything Model (SAM) for medical image segmentation remains challenging due to its limited adaptability across diverse medical domains. Although fine-tuned variants, such as MedSAM, improve performance in scenarios…

Computer Vision and Pattern Recognition · Computer Science 2026-01-06 Jianghao Wu , Yicheng Wu , Yutong Xie , Wenjia Bai , You Zhang , Feilong Tang , Yulong Li , Imran Razzak , Daniel F Schmidt , Yasmeen George

A global shortage of radiologists has been exacerbated by the significant volume of chest X-ray workloads, particularly in primary care. Although multimodal large language models show promise, existing evaluations predominantly rely on…

We present MM1.5, a new family of multimodal large language models (MLLMs) designed to enhance capabilities in text-rich image understanding, visual referring and grounding, and multi-image reasoning. Building upon the MM1 architecture,…

Despite advances in physics-based 3D motion synthesis, current methods face key limitations: reliance on pre-reconstructed 3D Gaussian Splatting (3DGS) built from dense multi-view images with time-consuming per-scene optimization; physics…

Computer Vision and Pattern Recognition · Computer Science 2026-03-18 Chunji Lv , Zequn Chen , Donglin Di , Weinan Zhang , Hao Li , Wei Chen , Yinjie Lei , Changsheng Li
‹ Prev 1 8 9 10 Next ›