Related papers: Hebbian control of fixations in a dyslexic reader
Multimodal large language models (MLLMs) have become a key interface for visual reasoning and grounded question answering, yet they remain vulnerable to visual hallucinations, where generated responses contradict image content or mention…
The flickering of the light of fluorescent lamps (FL), whose basic frequency, 100 Hz, is close to that of the micro-saccadic (the eye-muscles' tremor component) eye movement, is a severe problem for autists (autistic humans), a problem for…
Eye movements during fixation of a stationary target prevent the adaptation of the photoreceptors to continuous illumination and inhibit fading of the image. These random, involuntary, small, movements are restricted at long time scales so…
Hallucinations in Large Vision-Language Models (LVLMs) significantly undermine their reliability, motivating researchers to explore the causes of hallucination. However, most studies primarily focus on the language aspect rather than the…
In reading tasks drift can move fixations from one word to another or even another line, invalidating the eye tracking recording. Manual correction is time-consuming and subjective, while automated correction is fast yet limited in…
Large Vision-Language Models (LVLMs) are an extension of Large Language Models (LLMs) that facilitate processing both image and text inputs, expanding AI capabilities. However, LVLMs struggle with object hallucinations due to their reliance…
Vibration is an efficient way of conveying information from a device to its user, and it is increasingly used for wrist or finger-worn devices such as smart rings. Unexpected vibrations or sounds from the environment may disrupt the…
Highlighted text in the Internet (i.e. Hypertext) is predominantly blue and underlined. The percept of these hypertext characteristics were heavily questioned by applied research and empirical tests resulted in inconclusive results. The…
Time delays between lensed multiple images have been known to provide an interesting probe of the Hubble constant, but such application is often limited by degeneracies with the shape of lens potentials. We propose a new statistical…
To adapt their behaviour in changing environments, cells sense concentrations by binding external ligands to their receptors. However, incorrect ligands may bind nonspecifically to receptors, and when their concentration is large, this…
Eye movements during reading offer insights into both the reader's cognitive processes and the characteristics of the text that is being read. Hence, the analysis of scanpaths in reading have attracted increasing attention across fields,…
Continual learning is the problem of sequentially learning new tasks or knowledge while protecting previously acquired knowledge. However, catastrophic forgetting poses a grand challenge for neural networks performing such learning process.…
The study of light lensed by cosmic matter has yielded much information about astrophysical questions. Observations are explained using geometrical optics following a ray-based description of light. After deflection the lensed light…
Large Vision Language Models (LVLMs) have demonstrated remarkable capabilities in understanding and describing visual content, achieving state-of-the-art performance across various vision-language tasks. However, these models often generate…
Large Vision-Language Models (LVLMs) bridge the gap between visual and linguistic modalities, demonstrating strong potential across a variety of domains. However, despite significant progress, LVLMs still suffer from severe hallucination…
Recent advancements in large vision-language models (LVLM) have significantly enhanced their ability to comprehend visual inputs alongside natural language. However, a major challenge in their real-world application is hallucination, where…
Recent advancements in multimodal large language models (MLLMs) have significantly improved performance in visual question answering. However, they often suffer from hallucinations. In this work, hallucinations are categorized into two main…
Despite the rapid progress of multimodal large language models (MLLMs), they have largely overlooked the importance of visual processing. In a simple yet revealing experiment, we interestingly find that language-only models, when provided…
The importance of an element in a visual stimulus is commonly associated with the fixations during a free-viewing task. We argue that fixations are not always correlated with attention or awareness of visual objects. We suggest to filter…
Much has been learned about plasticity of biological synapses from empirical studies. Hebbian plasticity is driven by correlated activity of presynaptic and postsynaptic neurons. Synapses that converge onto the same neuron often behave as…