中文
相关论文

相关论文: TDMD: A Database for Dynamic Color Mesh Subjective…

200 篇论文

Convolutional neural networks (CNNs) have shown great effectiveness in medical image segmentation. However, they may be limited in modeling large inter-subject variations in organ shapes and sizes and exploiting global long-range contextual…

图像与视频处理 · 电气工程与系统科学 2024-10-04 Jin Yang , Daniel S. Marcus , Aristeidis Sotiras

Due to limitations in data quality, some essential visual tasks are difficult to perform independently. Introducing previously unavailable information to transfer informative dark knowledge has been a common way to solve such hard tasks.…

计算机视觉与模式识别 · 计算机科学 2023-06-29 Lingyu Si , Hongwei Dong , Wenwen Qiang , Junzhi Yu , Wenlong Zhai , Changwen Zheng , Fanjiang Xu , Fuchun Sun

Accurate and efficient plasma models are essential to understand and control experimental devices. Existing magnetohydrodynamic or kinetic models are nonlinear, computationally intensive, and can be difficult to interpret, while often only…

等离子体物理 · 物理学 2020-03-04 Alan A. Kaptanoglu , Kyle D. Morgan , Chris J. Hansen , Steven L. Brunton

Reconstruction-based anomaly detection models achieve their purpose by suppressing the generalization ability for anomaly. However, diverse normal patterns are consequently not well reconstructed as well. Although some efforts have been…

计算机视觉与模式识别 · 计算机科学 2023-03-10 Wenrui Liu , Hong Chang , Bingpeng Ma , Shiguang Shan , Xilin Chen

Vision-language models (VLMs) are increasingly used in settings where sensitivity to low-level image degradations matters, including content moderation, image restoration, and quality monitoring. Yet their ability to recognize distortion…

计算机视觉与模式识别 · 计算机科学 2026-04-23 Divyanshu Goyal , Akhil Eppa , Vanya Bannihatti Kumar

In lossy image compression, the objective is to achieve minimal signal distortion while compressing images to a specified bit rate. The increasing demand for visual analysis applications, particularly in classification tasks, has emphasized…

多媒体 · 计算机科学 2024-05-07 Yuefeng Zhang

Temporal volume images with 3D+t (4D) information are often used in medical imaging to statistically analyze temporal dynamics or capture disease progression. Although deep-learning-based generative models for natural images have been…

图像与视频处理 · 电气工程与系统科学 2022-06-28 Boah Kim , Jong Chul Ye

In this study, we investigated the stability of dynamic mode decomposition (DMD) algorithms to noisy data. To achieve a stable DMD algorithm, we applied the truncated total least squares (T-TLS) regression and optimal truncation level…

机器学习 · 计算机科学 2021-11-08 Yuya Ohmichi , Yosuke Sugioka , Kazuyuki Nakakita

We propose a perceptual video quality assessment (PVQA) metric for distorted videos by analyzing the power spectral density (PSD) of a group of pictures. This is an estimation approach that relies on the changes in video dynamic calculated…

计算机视觉与模式识别 · 计算机科学 2018-12-14 Mohammed A. Aabed , Gukyeong Kwon , Ghassan AlRegib

Recent advancements in Large Vision-Language Models (LVLMs) have demonstrated impressive capabilities across various multimodal tasks. They continue, however, to struggle with trivial scenarios such as reading values from Digital…

计算机视觉与模式识别 · 计算机科学 2025-09-01 João Valente , Atabak Dehban , Rodrigo Ventura

Diffusion Models (DMs) have demonstrated state-of-the-art performance in content generation without requiring adversarial training. These models are trained using a two-step process. First, a forward - diffusion - process gradually adds…

计算机视觉与模式识别 · 计算机科学 2024-03-13 Anwaar Ulhaq , Naveed Akhtar

Abstract. The advancement of deep learning has coincided with the proliferation of both models and available data. The surge in dataset sizes and the subsequent surge in computational requirements have led to the development of the Dataset…

计算机视觉与模式识别 · 计算机科学 2024-09-24 Jun-Yeong Moon , Jung Uk Kim , Gyeong-Moon Park

Differential Dynamic Microscopy (DDM) analyzes traditional real-space microscope images to extract information on sample dynamics in a way akin to light scattering, by decomposing each image in a sequence into Fourier modes, and evaluating…

软凝聚态物质 · 物理学 2017-11-10 Fabio Giavazzi , Paolo Edera , Peter J. Lu , Roberto Cerbino

Deep neural network (DNN) model compression for efficient on-device inference is becoming increasingly important to reduce memory requirements and keep user data on-device. To this end, we propose a novel differentiable k-means clustering…

机器学习 · 计算机科学 2022-02-22 Minsik Cho , Keivan A. Vahid , Saurabh Adya , Mohammad Rastegari

Deep neural networks based methods have been proved to achieve outstanding performance on object detection and classification tasks. Despite significant performance improvement, due to the deep structures, they still require prohibitive…

计算机视觉与模式识别 · 计算机科学 2020-01-08 Mohammad Farhadi , Yezhou Yang

Perspective distortion (PD) causes unprecedented changes in shape, size, orientation, angles, and other spatial relationships of visual concepts in images. Precisely estimating camera intrinsic and extrinsic parameters is a challenging task…

计算机视觉与模式识别 · 计算机科学 2024-07-16 Prakash Chandra Chhipa , Meenakshi Subhash Chippa , Kanjar De , Rajkumar Saini , Marcus Liwicki , Mubarak Shah

Scientific research and engineering practice often require the modeling and decomposition of nonlinear systems. The Dynamic Mode Decomposition (DMD) is a novel Koopman-based technique that effectively dissects high-dimensional nonlinear…

Over nearly two decades, Differential Dynamic Microscopy (DDM) has become a standard technique for extracting dynamic correlation functions from time-lapse microscopy data, with applications spanning colloidal suspensions, polymer…

软凝聚态物质 · 物理学 2025-11-11 Enrico Lattuada , Fabian Krautgasser , Maxime Lavaud , Fabio Giavazzi , Roberto Cerbino

The performance of machine learning models depends heavily on training data. The scarcity of large-scale, well-annotated datasets poses significant challenges in creating robust models. To address this, synthetic data generated through…

计算机视觉与模式识别 · 计算机科学 2025-10-09 Ayush Zenith , Arnold Zumbrun , Neel Raut , Jing Lin

3D Morphable Models (3DMMs) are statistical models that represent facial texture and shape variations using a set of linear bases and more particular Principal Component Analysis (PCA). 3DMMs were used as statistical priors for…

计算机视觉与模式识别 · 计算机科学 2019-04-09 Yuxiang Zhou , Jiankang Deng , Irene Kotsia , Stefanos Zafeiriou