中文
相关论文

相关论文: Multi-Interactive-Modality based Modeling for Myop…

200 篇论文

In lens cataract, the clouding change in lens leads to a decline of transparency of part of the lens. There are three types of senile cataract: cortical cataract, nuclear cataract, and posterior/anterior sub-capsular cataract. The most…

组织与器官 · 定量生物学 2018-02-06 Jicun Wang-Michelitsch , Thomas M. Michelitsch

What makes generalization hard for imitation learning in visual robotic manipulation? This question is difficult to approach at face value, but the environment from the perspective of a robot can often be decomposed into enumerable factors…

机器人学 · 计算机科学 2023-07-10 Annie Xie , Lisa Lee , Ted Xiao , Chelsea Finn

Multimodal large language models (MLLMs) often suffer from perceptual impairments under extended reasoning modes, particularly in visual question answering (VQA) tasks. We identify attention dispersion as the underlying cause: during…

计算机视觉与模式识别 · 计算机科学 2026-05-21 Ruiying Peng , Xueyu Wu , Jing Lei , Lu Hou , Yuanzheng Ma , Xiaohui Li

Vision-Language Models (VLMs) have been shown to be blind, often underutilizing their visual inputs even on tasks that require visual reasoning. In this work, we demonstrate that VLMs are selectively blind. They modulate the amount of…

计算机视觉与模式识别 · 计算机科学 2026-03-23 Wan-Cyuan Fan , Jiayun Luo , Declan Kutscher , Leonid Sigal , Ritwik Gupta

Myopic macular degeneration is the most common complication of myopia and the primary cause of vision loss in individuals with pathological myopia. Early detection and prompt treatment are crucial in preventing vision impairment due to…

图像与视频处理 · 电气工程与系统科学 2024-01-09 Yihao Li , Philippe Zhang , Yubo Tan , Jing Zhang , Zhihan Wang , Weili Jiang , Pierre-Henri Conze , Mathieu Lamard , Gwenolé Quellec , Mostafa El Habib Daho

For human children as well as machine learning systems, a key challenge in learning a word is linking the word to the visual phenomena it describes. We explore this aspect of word learning by using the performance of computer vision systems…

计算与语言 · 计算机科学 2023-09-12 Sunayana Rane , Mira L. Nencheva , Zeyu Wang , Casey Lew-Williams , Olga Russakovsky , Thomas L. Griffiths

We have modified the sexual Penna model by introducing the fluctuating environment and fluctuations representing physiological functions of individuals. Additionally, we have introduced the mother care corresponding to the protection…

种群与进化 · 定量生物学 2008-11-04 Przemyslaw Biecek , Katarzyna Bonkowska , Stanislaw Cebrat

Our main goal is to discover the main factors influencing students' academic trajectory and students' academic evolution within such environment. Our results indicate strong correlation in this virtual learning environment between student…

数据库 · 计算机科学 2016-12-06 Eid Aldikanji , Khalil Ajami

Learning disorders are neurological conditions that affect the brain's ability to interconnect communication areas. Dyslexic students experience problems with reading, memorizing, and exposing concepts; however the magnitude of these can be…

We consider learning from labeled data collected across multiple environments, where the data distribution may vary across these environments. This problem is commonly approached from a causal perspective, seeking invariant representations…

机器学习 · 统计学 2026-04-30 Yuli Slavutsky , David M. Blei

The performance of computer vision models are susceptible to unexpected changes in input images caused by sensor errors or extreme imaging environments, known as common corruptions (e.g. noise, blur, illumination changes). These corruptions…

计算机视觉与模式识别 · 计算机科学 2024-09-17 Shunxin Wang , Raymond Veldhuis , Christoph Brune , Nicola Strisciuglio

Vision-language models (VLMs) can respond to queries about images in many languages. However, beyond language, culture affects how we see things. For example, individuals from Western cultures focus more on the central figure in an image…

计算与语言 · 计算机科学 2025-03-04 Amith Ananthram , Elias Stengel-Eskin , Mohit Bansal , Kathleen McKeown

We investigate strong lensing by non-singular finite isothermal ellipsoids taking into account the influence of the matter along the line of sight and in the close lens vicinity. We compare three descriptions of light propagation: the full…

宇宙学与河外天体物理 · 物理学 2015-06-04 M. Jaroszynski , Z. Kostrzewa-Rutkowska

Appearance-based gaze estimation (AGE) has achieved remarkable performance in constrained settings, yet we reveal a significant generalization gap where existing AGE models often fail in practical, unconstrained scenarios, particularly…

计算机视觉与模式识别 · 计算机科学 2026-03-31 Zhenhao Li , Zheng Liu , Seunghyun Lee , Amin Fadaeinejad , Yuanhao Yu

It is a challenging task for visually impaired people to perceive their surrounding environment due to the complexity of the natural scenes. Their personal and social activities are thus highly limited. This paper introduces a Large…

计算机视觉与模式识别 · 计算机科学 2025-04-28 Zezhou Chen , Zhaoxiang Liu , Kai Wang , Kohou Wang , Shiguo Lian

Multimodal large language models (MLLMs) frequently suffer from object hallucinations, yet the visual perceptual mechanism underlying this failure remains poorly understood. In this work, we reveal that hallucinations are strongly…

计算机视觉与模式识别 · 计算机科学 2026-05-26 Quanjiang Li , Zhiming Liu , Wei Luo , Tingjin Luo , Chenping Hou

Early detection of eye diseases like glaucoma, macular degeneration, and diabetic retinopathy is crucial for preventing vision loss. While artificial intelligence (AI) foundation models hold significant promise for addressing these…

计算机视觉与模式识别 · 计算机科学 2024-09-12 Danli Shi , Weiyi Zhang , Jiancheng Yang , Siyu Huang , Xiaolan Chen , Mayinuer Yusufu , Kai Jin , Shan Lin , Shunming Liu , Qing Zhang , Mingguang He

While mainstream vision-language models (VLMs) have advanced rapidly in understanding image level information, they still lack the ability to focus on specific areas designated by humans. Rather, they typically rely on large volumes of…

计算机视觉与模式识别 · 计算机科学 2025-02-13 Kangyu Zhu , Ziyuan Qin , Huahui Yi , Zekun Jiang , Qicheng Lao , Shaoting Zhang , Kang Li

Current popular Large Vision-Language Models (LVLMs) are suffering from Hallucinations on Object Attributes (HoOA), leading to incorrect determination of fine-grained attributes in the input images. Leveraging significant advancements in 3D…

计算机视觉与模式识别 · 计算机科学 2025-01-20 Zhijie Tan , Yuzhi Li , Shengwei Meng , Xiang Yuan , Weiping Li , Tong Mo , Bingce Wang , Xu Chu

In this study we provide the analysis of eye movement behavior elicited by low-level feature distinctiveness with a dataset of synthetically-generated image patterns. Design of visual stimuli was inspired by the ones used in previous…

计算机视觉与模式识别 · 计算机科学 2018-11-19 David Berga , Xosé Ramón Fdez-Vidal , Xavier Otazu , Víctor Leborán , Xosé M. Pardo