中文
相关论文

相关论文: High Performance Software in Multidimensional Redu…

200 篇论文

Comparing two images in terms of Commonalities and Differences (CaD) is a fundamental human capability that forms the basis of advanced visual reasoning and interpretation. It is essential for the generation of detailed and contextually…

计算机视觉与模式识别 · 计算机科学 2024-06-14 Wei Lin , Muhammad Jehanzeb Mirza , Sivan Doveh , Rogerio Feris , Raja Giryes , Sepp Hochreiter , Leonid Karlinsky

Optimized for pixel fidelity metrics, images compressed by existing image codec are facing systematic challenges when used for visual analysis tasks, especially under low-bitrate coding. This paper proposes a visual analysis-motivated…

图像与视频处理 · 电气工程与系统科学 2021-04-22 Zhimeng Huang , Chuanmin Jia , Shanshe Wang , Siwei Ma

Vector Symbolic Architecture (VSA) is emerging in machine learning due to its efficiency, but they are hindered by issues of hyperdimensionality and accuracy. As a promising mitigation, the Low-Dimensional Computing (LDC) method…

机器学习 · 计算机科学 2025-03-18 Shijin Duan , Yejia Liu , Gaowen Liu , Ramana Rao Kompella , Shaolei Ren , Xiaolin Xu

Linking between two data sources is a basic building block in numerous computer vision problems. In this paper, we set to answer a fundamental cognitive question: are prior correspondences necessary for linking between different domains?…

计算机视觉与模式识别 · 计算机科学 2018-04-03 Yedid Hoshen , Lior Wolf

We investigated whether a combination of k-space undersampling and variable density averaging enhances image quality for low-SNR MRI acquisitions. We implemented 3D Cartesian k-space prospective undersampling with a variable number of…

Despite the advances of deep learning in specific tasks using images, the principled assessment of image fidelity and similarity is still a critical ability to develop. As it has been shown that Mean Squared Error (MSE) is insufficient for…

图像与视频处理 · 电气工程与系统科学 2019-08-27 Benyamin Ghojogh , Fakhri Karray , Mark Crowley

To improve Multimodal Large Language Models' (MLLMs) ability to process images and complex instructions, researchers predominantly curate large-scale visual instruction tuning datasets, which are either sourced from existing vision tasks or…

计算与语言 · 计算机科学 2025-02-28 Zhenyu Liu , Yunxin Li , Baotian Hu , Wenhan Luo , Yaowei Wang , Min Zhang

The screen content images (SCIs) usually comprise various content types with sharp edges, in which the artifacts or distortions can be well sensed by the vanilla structure similarity measurement in a full reference manner. Nonetheless,…

计算机视觉与模式识别 · 计算机科学 2020-08-13 Chenglizhao Chen , Hongmeng Zhao , Huan Yang , Chong Peng , Teng Yu

The success of algorithms in the analysis of high-dimensional data is often attributed to the manifold hypothesis, which supposes that this data lie on or near a manifold of much lower dimension. It is often useful to determine or estimate…

机器学习 · 统计学 2024-09-10 Anna C. Gilbert , Kevin O'Neill

Recent advances in photoacoustic (PA) imaging have enabled detailed images of microvascular structure and quantitative measurement of blood oxygenation or perfusion. Standard reconstruction methods for PA imaging are based on solving an…

信号处理 · 电气工程与系统科学 2020-04-17 MinWoo Kim , Geng-Shi Jeng , Ivan Pelivanov , Matthew O'Donnell

The dominant image-to-image translation methods are based on fully convolutional networks, which extract and translate an image's features and then reconstruct the image. However, they have unacceptable computational costs when working with…

计算机视觉与模式识别 · 计算机科学 2022-07-12 Yuda Song , Hui Qian , Xin Du

Canonical Correlation Analysis (CCA) is a widely used spectral technique for finding correlation structures in multi-view datasets. In this paper, we tackle the problem of large scale CCA, where classical algorithms, usually requiring…

机器学习 · 统计学 2015-06-29 Zhuang Ma , Yichao Lu , Dean Foster

Principal component analysis (PCA) is one of the most popular dimension reduction techniques in statistics and is especially powerful when a multivariate distribution is concentrated near a lower-dimensional subspace. Multivariate extreme…

统计方法学 · 统计学 2025-07-15 Felix Reinbott , Anja Janßen

Vision-Language Models (VLMs) leverage aligned visual encoders to transform images into visual tokens, allowing them to be processed similarly to text by the backbone large language model (LLM). This unified input paradigm enables VLMs to…

计算机视觉与模式识别 · 计算机科学 2025-03-18 Bangzheng Li , Fei Wang , Wenxuan Zhou , Nan Xu , Ben Zhou , Sheng Zhang , Hoifung Poon , Muhao Chen

Developing interpretable machine learning models has become an increasingly important issue. One way in which data scientists have been able to develop interpretable models has been to use dimension reduction techniques. In this paper, we…

机器学习 · 计算机科学 2023-03-23 Sean H. Merritt , Alexander P. Christensen

Convolutional sparse coding (CSC) is an important building block of many computer vision applications ranging from image and video compression to deep learning. We present two contributions to the state of the art in CSC. First, we…

计算机视觉与模式识别 · 计算机科学 2018-04-10 Lama Affara , Bernard Ghanem , Peter Wonka

This project intends to study the image representation based on attention mechanism and multimodal data. By adding multiple pattern layers to the attribute model, the semantic and hidden layers of image content are integrated. The word…

计算与语言 · 计算机科学 2024-06-14 Dan Sun , Yaxin Liang , Yining Yang , Yuhan Ma , Qishi Zhan , Erdi Gao

When handling complicated text images (e.g., irregular structures, low resolution, heavy occlusion, and uneven illumination), existing supervised text recognition methods are data-hungry. Although these methods employ large-scale synthetic…

计算机视觉与模式识别 · 计算机科学 2023-08-21 Tongkun Guan , Wei Shen , Xue Yang , Qi Feng , Zekun Jiang , Xiaokang Yang

The usage of digital content (photos and videos) in a variety of applications has increased due to the popularity of multimedia devices. These uses include advertising campaigns, educational resources, and social networking platforms. There…

计算机视觉与模式识别 · 计算机科学 2025-02-11 Muhammad Turab

In this work, we present a new Vector Space Model (VSM) of speech utterances for the task of spoken dialect identification. Generally, DID systems are built using two sets of features that are extracted from speech utterances; acoustic and…

计算与语言 · 计算机科学 2016-09-20 Sameer Khurana , Ahmed Ali , Steve Renals
‹ 上一页 1 8 9 10 下一页 ›