English
Related papers

Related papers: Multi-Interactive-Modality based Modeling for Myop…

200 papers

Myopia in human and animals is caused by the axial elongation of the eye and is closely linked to the thinning of the sclera which supports the eye tissue. This thinning has been correlated with the overproduction of matrix…

Systems and Control · Computer Science 2013-09-05 Qin Shu , Diana Catalina Ardila , Ricardo G. Sanfelice , Jonathan P. Vande Geest

Human memory exhibits significant vulnerability in cognitive tasks and daily life. Comparisons between visual working memory and new perceptual input (e.g., during cognitive tasks) can lead to unintended memory distortions. Previous studies…

Neurons and Cognition · Quantitative Biology 2025-07-31 Yuang Cao , Jiachen Zou , Chen Wei , Quanying Liu

Multimodal Large Language Models (MLLMs) have emerged as a central focus in both industry and academia, but often suffer from biases introduced by visual and language priors, which can lead to multimodal hallucination. These biases arise…

Computer Vision and Pattern Recognition · Computer Science 2025-02-19 Guanyu Zhou , Yibo Yan , Xin Zou , Kun Wang , Aiwei Liu , Xuming Hu

Detecting mind wandering is crucial in online education, and it occurs 30% of the time, as it directly impacts learners' retention, comprehension, and overall success in self-directed learning environments. Integrating automated detection…

Over 140 million people worldwide and over 45 million people in the United States wear contact lenses; it is estimated that 12%-27.4% contact lens users stop wearing them due to discomfort. Contact lens mechanical interactions with the…

Numerical Analysis · Mathematics 2025-12-11 Lucia Carichino , Kara L. Maki , David S. Ross , Riley K. Supple , Evan Rysdam

Analyzing air pollution data is challenging as there are various analysis focuses from different aspects: feature (what), space (where), and time (when). As in most geospatial analysis problems, besides high-dimensional features, the…

Machine Learning · Computer Science 2022-02-14 Yun-Hsin Kuo , Takanori Fujiwara , Charles C. -K. Chou , Chun-houh Chen , Kwan-Liu Ma

Vision-language models (VLMs) have exhibited remarkable generalization capabilities, and prompt learning for VLMs has attracted great attention for the ability to adapt pre-trained VLMs to specific downstream tasks. However, existing…

Machine Learning · Computer Science 2025-01-15 Song-Lin Lv , Yu-Yang Chen , Zhi Zhou , Ming Yang , Lan-Zhe Guo

Multimodal large language models (MLLMs) achieve strong performance on vision-language tasks, yet their visual processing is opaque. Most black-box evaluations measure task accuracy, but reveal little about underlying mechanisms. Drawing on…

Computer Vision and Pattern Recognition · Computer Science 2025-10-23 John Burden , Jonathan Prunty , Ben Slater , Matthieu Tehenan , Greg Davis , Lucy Cheke

The evolution of Large Vision-Language Models (LVLMs) has progressed from single to multi-image reasoning. Despite this advancement, our findings indicate that LVLMs struggle to robustly utilize information across multiple images, with…

Computer Vision and Pattern Recognition · Computer Science 2025-03-19 Xinyu Tian , Shu Zou , Zhaoyuan Yang , Jing Zhang

Obesity is currently affecting very large portions of the global population. Effective prevention and treatment starts at the early age and requires objective knowledge of population-level behavior on the region/neighborhood scale. To this…

We explore social perception of human faces in CLIP, a widely used open-source vision-language model. To this end, we compare the similarity in CLIP embeddings between different textual prompts and a set of face images. Our textual prompts…

Computer Vision and Pattern Recognition · Computer Science 2025-09-16 Carina I. Hausladen , Manuel Knott , Colin F. Camerer , Pietro Perona

Age Related Macular Degeneration(AMD) has been one of the most leading causes of permanent vision impairment in ophthalmology. Though treatments, such as anti VEGF drugs or photodynamic therapies, were developed to slow down the…

Image and Video Processing · Electrical Eng. & Systems 2025-12-29 Daeyoung Kim

Keeping the health condition at the juvenile population is the concern of both the medical staff and the physical education teachers. Considering the steady tendency of the growth of the number of the youth with excessive weight…

Computers and Society · Computer Science 2009-03-04 Tiberiu Marius Karnyanszky , Corina Musuroi , Carla Amira Karnyanszky

We review evidence supporting the role of early life programming in the susceptibility for adult neurodegenerative diseases while highlighting questions and proposing avenues for future research to advance our understanding of this…

Tissues and Organs · Quantitative Biology 2019-11-11 Paula Desplats , Ashley M. Gutierrez , Marta C. Antonelli , Martin G. Frasch

Visual representation learning is ubiquitous in various real-world applications, including visual comprehension, video understanding, multi-modal analysis, human-computer interaction, and urban computing. Due to the emergence of huge…

Computer Vision and Pattern Recognition · Computer Science 2023-03-23 Yang Liu , Yushen Wei , Hong Yan , Guanbin Li , Liang Lin

Large Vision-Language Models (LVLMs) demonstrate remarkable capabilities in multimodal tasks, but visual object hallucination remains a persistent issue. It refers to scenarios where models generate inaccurate visual object-related…

Computer Vision and Pattern Recognition · Computer Science 2025-05-06 Liqiang Jing , Guiming Hardy Chen , Ehsan Aghazadeh , Xin Eric Wang , Xinya Du

Large language models (LLMs) have become increasingly useful computational models of human language processing, but it remains unclear whether vision-language learning makes text representations more human-like during natural reading. Here,…

Computation and Language · Computer Science 2026-05-28 Jinzhou Wu , Zhengwu Ma , Jixing Li , Baoping Tang , Zitong Lu

Contrastive vision-language models (VLMs), like CLIP, have gained popularity for their versatile applicability to various downstream tasks. Despite their successes in some tasks, like zero-shot object recognition, they perform surprisingly…

Computer Vision and Pattern Recognition · Computer Science 2025-04-17 Simon Schrodi , David T. Hoffmann , Max Argus , Volker Fischer , Thomas Brox

Large vision language models (LVLMs) often suffer from object hallucination, producing objects not present in the given images. While current benchmarks for object hallucination primarily concentrate on the presence of a single object class…

Computer Vision and Pattern Recognition · Computer Science 2024-11-04 Xuweiyi Chen , Ziqiao Ma , Xuejun Zhang , Sihan Xu , Shengyi Qian , Jianing Yang , David F. Fouhey , Joyce Chai
‹ Prev 1 3 4 5 6 7 10 Next ›