中文
相关论文

相关论文: Visible Light-Based Human Visual System Conceptual…

200 篇论文

Large vision-language models (LVLMs) have achieved impressive results in various vision-language tasks. However, despite showing promising performance, LVLMs suffer from hallucinations caused by language bias, leading to diminished focus on…

计算机视觉与模式识别 · 计算机科学 2026-05-29 Haozhe Zhao , Shuzheng Si , Liang Chen , Yichi Zhang , Maosong Sun , Mingjia Zhang , Baobao Chang

Current benchmarks for assessing vision-language models (VLMs) often focus on their perception or problem-solving capabilities and neglect other critical aspects such as fairness, multilinguality, or toxicity. Furthermore, they differ in…

计算机视觉与模式识别 · 计算机科学 2024-10-25 Tony Lee , Haoqin Tu , Chi Heem Wong , Wenhao Zheng , Yiyang Zhou , Yifan Mai , Josselin Somerville Roberts , Michihiro Yasunaga , Huaxiu Yao , Cihang Xie , Percy Liang

Visible Light Communication (VLC) is emerging as a means to network computing devices that ameliorates many hurdles of radio-frequency (RF) communications, for example, the limited available spectrum. Enabling VLC in wearable computing,…

网络与互联网体系结构 · 计算机科学 2018-03-26 Ambuj Varshney , Luca Mottola , Thiemo Voigt

Visual hallucination (VH) means that a multi-modal LLM (MLLM) imagines incorrect details about an image in visual question answering. Existing studies find VH instances only in existing image datasets, which results in biased understanding…

计算机视觉与模式识别 · 计算机科学 2024-06-18 Wen Huang , Hongbin Liu , Minxin Guo , Neil Zhenqiang Gong

This paper presents an integrated approach to Visual SLAM, merging online sequential photometric calibration within a Hybrid direct-indirect visual SLAM (H-SLAM). Photometric calibration helps normalize pixel intensity values under…

机器人学 · 计算机科学 2024-09-26 Nicolas Abboud , Malak Sayour , Imad H. Elhajj , John Zelek , Daniel Asmar

Various robots, rovers, drones, and other agents of mass-produced products are expected to encounter scenes where they intersect and collaborate in the near future. In such multi-agent systems, individual identification and communication…

多智能体系统 · 计算机科学 2024-02-16 Haruyuki Nakagawa , Yoshitaka Miyatani , Asako Kanezaki

Vision-language models (VLMs) excel in semantic tasks but falter at a core human capability: detecting hidden content in optical illusions or AI-generated images through perceptual adjustments like zooming. We introduce HC-Bench, a…

计算与语言 · 计算机科学 2025-10-16 Sifan Li , Yujun Cai , Yiwei Wang

There exists an intrinsic relationship between the spectral sensitivity of the Human Visual System (HVS) and colour perception; these intertwined phenomena are often overlooked in perceptual compression research. In general, most previously…

计算机视觉与模式识别 · 计算机科学 2022-01-25 Lee Prangnell , Victor Sanchez

Color representation is essential in computer vision and human-computer interaction. There are multiple color models available. The choice of a suitable color model is critical for various applications. This paper presents a review of color…

计算机视觉与模式识别 · 计算机科学 2026-03-13 Muragul Muratbekova , Nuray Toganas , Ayan Igali , Maksat Shagyrov , Elnara Kadyrgali , Adilet Yerkin , Pakizar Shamoi

Video see-through (VST) technology aims to seamlessly blend virtual and physical worlds by reconstructing reality through cameras. While manufacturers promise perceptual fidelity, it remains unclear how close these systems are to…

人机交互 · 计算机科学 2026-01-07 Jialin Wang , Songming Ping , Kemu Xu , Yue Li , Hai-Ning Liang

Vision sensors are versatile and can capture a wide range of visual cues, such as color, texture, shape, and depth. This versatility, along with the relatively inexpensive availability of machine vision cameras, played an important role in…

计算机视觉与模式识别 · 计算机科学 2024-04-18 Muhammad Z. Alam , Zeeshan Kaleem , Sousso Kelouwani

Leveraging large-scale Text-to-Image (TTI) models have become a common technique for generating exemplar or training dataset in the fields of image synthesis, video editing, 3D reconstruction. However, semantic structural visual…

计算机视觉与模式识别 · 计算机科学 2026-03-09 Bumsoo Kim , Wonseop Shin , Kyuchul Lee , Yonghoon Jung , Sanghyun Seo

Visible light communication (VLC) builds upon the dual use of existing lighting infrastructure for wireless data transmission. VLC has recently gained interest as cost-effective, secure, and energy-efficient wireless access technology…

信号处理 · 电气工程与系统科学 2018-11-15 Saad Al-Ahmadi , Omar Maraqa , Murat Uysal , Sadiq M. Sait

In perceptual image coding applications, the main objective is to decrease, as much as possible, Bits Per Pixel (BPP) while avoiding noticeable distortions in the reconstructed image. In this paper, we propose a novel perceptual image…

图像与视频处理 · 电气工程与系统科学 2021-02-10 Lee Prangnell , Victor Sanchez

Color constancy aims to keep object colors consistent under varying illumination. Cross-camera generalization in color constancy remains challenging because learning-based models often overfit to the color response characteristics of the…

计算机视觉与模式识别 · 计算机科学 2026-05-20 Shuwei Li , Lei Tan , Robby T. Tan

Vision-language models (VLMs) have demonstrated impressive performance by effectively integrating visual and textual information to solve complex tasks. However, it is not clear how these models reason over the visual and textual data…

人工智能 · 计算机科学 2025-04-15 Pouya Pezeshkpour , Moin Aminnaseri , Estevam Hruschka

Vision-language modeling (VLM) aims to bridge the information gap between images and natural language. Under the new paradigm of first pre-training on massive image-text pairs and then fine-tuning on task-specific data, VLM in the remote…

计算机视觉与模式识别 · 计算机科学 2025-06-11 Xingxing Weng , Chao Pang , Gui-Song Xia

Vision-language models (VLMs), such as CLIP and ALIGN, are generally trained on datasets consisting of image-caption pairs obtained from the web. However, real-world multimodal datasets, such as healthcare data, are significantly more…

计算机视觉与模式识别 · 计算机科学 2023-08-23 Maya Varma , Jean-Benoit Delbrouck , Sarah Hooper , Akshay Chaudhari , Curtis Langlotz

Light plays an important role in human well-being. However, most computer vision tasks treat pixels without considering their relationship to physical luminance. To address this shortcoming, we introduce the Laval Photometric Indoor HDR…

计算机视觉与模式识别 · 计算机科学 2023-10-16 Christophe Bolduc , Justine Giroux , Marc Hébert , Claude Demers , Jean-François Lalonde

Image segmentation has been a very active research topic in image analysis area. Currently, most of the image segmentation algorithms are designed based on the idea that images are partitioned into a set of regions preserving homogeneous…

计算机视觉与模式识别 · 计算机科学 2015-03-17 Yu Su , Margaret H. Dunham