中文
相关论文

相关论文: Logarithmic mathematical morphology: a new framewo…

200 篇论文

Aesthetic image cropping aims to enhance the aesthetic quality of an image by improving its composition through spatial cropping. Previous methods often rely on saliency prediction or retrieval augmentation, ignoring the task's core…

计算机视觉与模式识别 · 计算机科学 2026-05-14 Zhitong Dong , Chao Li , Jie Yu , Hao Chen

This paper describes a novel perspective on the foundations of mathematics: how mathematics may be seen to be largely about 'information compression via the matching and unification of patterns' (ICMUP). ICMUP is itself a novel approach to…

人工智能 · 计算机科学 2018-10-10 J Gerard Wolff

The free-form deformation model can represent a wide range of non-rigid deformations by manipulating a control point lattice over the image. However, due to a large number of parameters, it is challenging to fit the free-form deformation…

计算机视觉与模式识别 · 计算机科学 2022-06-10 Takumi Nakane , Haoran Xie , Chao Zhang

Interferometric closure invariants encode calibration-independent details of an object's morphology. Excepting simple cases, a direct backward transformation from closure invariants to morphologies is not well established. We demonstrate…

天体物理仪器与方法 · 物理学 2024-08-27 Nithyanandan Thyagarajan , Lucas Hoefs , O. Ivy Wong

We considers how a particular kind of graph corresponds to multiplicative intuitionistic linear logic formula. The main feature of the graphical notation is that it absorbs certain symmetries between conjunction and implication. We look at…

计算机科学中的逻辑 · 计算机科学 2022-08-08 Lucas Dixon

This paper aims to establish the theoretical foundation for shift inclusion in mathematical morphology. In this paper, we prove that the morphological opening and closing concerning structuring elements of shift inclusion property would…

离散数学 · 计算机科学 2020-12-25 Chuan-Shen Hu , Yu-Min Chung

Multimodal Large Language Models (MLLMs) still struggle with fine-grained visual understanding, where answers often depend on small but decisive evidence in the full image. We observe a regional-to-global perception gap: the same MLLM…

计算机视觉与模式识别 · 计算机科学 2026-05-28 Qianhao Yuan , Jie Lou , Xing Yu , Hongyu Lin , Le Sun , Xianpei Han , Yaojie Lu

Morphological reconstruction (MR) is often employed by seeded image segmentation algorithms such as watershed transform and power watershed as it is able to filter seeds (regional minima) to reduce over-segmentation. However, MR might…

计算机视觉与模式识别 · 计算机科学 2019-10-02 Tao Lei , Xiaohong Jia , Tongliang Liu , Shigang Liu , Hongying Meng , Asoke K. Nandi

Text-guided image generation enables the creation of visual content from textual descriptions. However, certain visual concepts cannot be effectively conveyed through language alone. This has sparked a renewed interest in utilizing the CLIP…

计算机视觉与模式识别 · 计算机科学 2024-06-04 Elad Richardson , Yuval Alaluf , Ali Mahdavi-Amiri , Daniel Cohen-Or

Digital memcomputing machines (DMMs) are a class of computational machines designed to solve combinatorial optimization problems. A practical realization of DMMs can be accomplished via electrical circuits of highly non-linear,…

新兴技术 · 计算机科学 2019-10-02 Massimiliano Di Ventra , Igor V. Ovchinnikov

In rapidly evolving field of vision-language models (VLMs), contrastive language-image pre-training (CLIP) has made significant strides, becoming foundation for various downstream tasks. However, relying on one-to-one (image, text)…

计算机视觉与模式识别 · 计算机科学 2024-12-03 Haicheng Wang , Chen Ju , Weixiong Lin , Shuai Xiao , Mengting Chen , Yixuan Huang , Chang Liu , Mingshuai Yao , Jinsong Lan , Ying Chen , Qingwen Liu , Yanfeng Wang

Multimodal LLMs can accurately perceive numerical content across modalities yet fail to perform exact multi-digit multiplication when the identical underlying arithmetic problem is presented as numerals, number words, images, or in audio…

计算与语言 · 计算机科学 2026-04-21 Samuel G. Balter , Ethan Jerzak , Connor T. Jerzak

Contrastive Language-Image Pre-training (CLIP) has achieved widely applications in various computer vision tasks, e.g., text-to-image generation, Image-Text retrieval and Image captioning. However, CLIP suffers from high memory and…

计算机视觉与模式识别 · 计算机科学 2026-02-06 Kangjie Zhang , Wenxuan Huang , Xin Zhou , Boxiang Zhou , Dejia Song , Yuan Xie , Baochang Zhang , Lizhuang Ma , Nemo Chen , Xu Tang , Yao Hu , Shaohui Lin

Identifying when different images are of the same object despite changes caused by imaging technologies, or processes such as growth, has many applications in fields such as computer vision and biological image analysis. One approach to…

计算机视觉与模式识别 · 计算机科学 2016-08-12 Stephen Marsland , Robert McLachlan

Feature-based object matching is a fundamental problem for many applications in computer vision, such as object recognition, 3D reconstruction, tracking, and motion segmentation. In this work, we consider simultaneously matching object…

计算机视觉与模式识别 · 计算机科学 2015-04-01 Kui Jia , Tsung-Han Chan , Zinan Zeng , Shenghua Gao , Gang Wang , Tianzhu Zhang , Yi Ma

Large language models (LLMs) have enabled the creation of multi-modal LLMs that exhibit strong comprehension of visual data such as images and videos. However, these models usually rely on extensive visual tokens from visual encoders,…

计算机视觉与模式识别 · 计算机科学 2025-07-30 Yiwu Zhong , Zhuoming Liu , Yin Li , Liwei Wang

Optical flow is the pattern of apparent motion of objects in a scene. The computation of optical flow is a critical component in numerous computer vision tasks such as object detection, visual object tracking, and activity recognition.…

信号处理 · 电气工程与系统科学 2024-01-15 Muhammad Wasim Nawaz , Abdesselam Bouzerdoum , Muhammad Mahboob Ur Rahman , Ghulam Abbas , Faizan Rashid

Ising machines are emerging as a powerful physical alternative to digital processors for solving combinatorial optimization problems. Among them, spatial photonic Ising machines (SPIMs) offer compact, room-temperature hardware with…

Light projection is a powerful technique to edit appearances of objects in the real world. Based on pixel-wise modification of light transport, previous techniques have successfully modified static surface properties such as surface color,…

图形学 · 计算机科学 2016-03-15 Takahiro Kawabe , Taiki Fukiage , Masataka Sawayama , Shin'ya Nishida

Metric learning seeks to embed images of objects suchthat class-defined relations are captured by the embeddingspace. However, variability in images is not just due to different depicted object classes, but also depends on other latent…

计算机视觉与模式识别 · 计算机科学 2019-09-26 Karsten Roth , Biagio Brattoli , Björn Ommer