中文
相关论文

相关论文: Delving Deep into Pixel Alignment Feature for Accu…

200 篇论文

Due to the effective performance of multi-scale feature fusion, Path Aggregation FPN (PAFPN) is widely employed in YOLO detectors. However, it cannot efficiently and adaptively integrate high-level semantic information with low-level…

计算机视觉与模式识别 · 计算机科学 2024-07-08 Zhiqiang Yang , Qiu Guan , Keer Zhao , Jianmin Yang , Xinli Xu , Haixia Long , Ying Tang

Existing Human NeRF methods for reconstructing 3D humans typically rely on multiple 2D images from multi-view cameras or monocular videos captured from fixed camera views. However, in real-world scenarios, human images are often captured…

计算机视觉与模式识别 · 计算机科学 2023-08-17 Shoukang Hu , Fangzhou Hong , Liang Pan , Haiyi Mei , Lei Yang , Ziwei Liu

We present "Humans and Structure from Motion" (HSfM), a method for jointly reconstructing multiple human meshes, scene point clouds, and camera parameters in a metric world coordinate system from a sparse set of uncalibrated multi-view…

计算机视觉与模式识别 · 计算机科学 2025-05-22 Lea Müller , Hongsuk Choi , Anthony Zhang , Brent Yi , Jitendra Malik , Angjoo Kanazawa

Pixel-aligned implicit models, such as PIFu, PIFuHD, and ICON, are used for single-view clothed human reconstruction. These models need to be trained using a sampling training scheme. Existing sampling training schemes either fail to…

计算机视觉与模式识别 · 计算机科学 2024-11-12 Kennard Yanting Chan , Fayao Liu , Guosheng Lin , Chuan Sheng Foo , Weisi Lin

Due to the effective multi-scale feature fusion capabilities of the Path Aggregation FPN (PAFPN), it has become a widely adopted component in YOLO-based detectors. However, PAFPN struggles to integrate high-level semantic cues with…

计算机视觉与模式识别 · 计算机科学 2025-02-27 Zhiqiang Yang , Qiu Guan , Zhongwen Yu , Xinli Xu , Haixia Long , Sheng Lian , Haigen Hu , Ying Tang

Infrared and visible image fusion (IVIF) integrates complementary modalities to enhance scene perception. Current methods predominantly focus on optimizing handcrafted losses and objective metrics, often resulting in fusion outcomes that do…

计算机视觉与模式识别 · 计算机科学 2026-03-05 Jinyuan Liu , Xingyuan Li , Qingyun Mei , Haoyuan Xu , Zhiying Jiang , Long Ma , Risheng Liu , Xin Fan

Multi-exposure image fusion is a method for producing an image with a wide dynamic range by fusing multiple images taken under various exposure values. In this paper, we discuss color distortion included in fused images, and propose a novel…

图像与视频处理 · 电气工程与系统科学 2019-07-29 Artit Visavakitcharoen , Yuma Kinoshita , Hitoshi Kiya

While the problem of estimating shapes and diffuse reflectances of human faces from images has been extensively studied, there is relatively less work done on recovering the specular albedo. This paper presents a lightweight solution for…

计算机视觉与模式识别 · 计算机科学 2019-11-28 Guoxian Song , Jianmin Zheng , Jianfei Cai , Tat-Jen Cham

Human reconstruction and synthesis from monocular RGB videos is a challenging problem due to clothing, occlusion, texture discontinuities and sharpness, and framespecific pose changes. Many methods employ deferred rendering, NeRFs and…

计算机视觉与模式识别 · 计算机科学 2023-03-16 Rohit Jena , Pratik Chaudhari , James Gee , Ganesh Iyer , Siddharth Choudhary , Brandon M. Smith

Multimodal medical image fusion (MMIF) aims to integrate images from different modalities to produce a comprehensive image that enhances medical diagnosis by accurately depicting organ structures, tissue textures, and metabolic information.…

计算机视觉与模式识别 · 计算机科学 2025-12-02 Tao Luo , Weihua Xu

We introduce a unified single and multi-view neural implicit 3D reconstruction framework VPFusion. VPFusion attains high-quality reconstruction using both - 3D feature volume to capture 3D-structure-aware context, and pixel-aligned image…

计算机视觉与模式识别 · 计算机科学 2022-07-19 Jisan Mahmud , Jan-Michael Frahm

Recent random-forest (RF)-based image super-resolution approaches inherit some properties from dictionary-learning-based algorithms, but the effectiveness of the properties in RF is overlooked in the literature. In this paper, we present a…

计算机视觉与模式识别 · 计算机科学 2017-12-15 Hailiang Li , Kin-Man Lam , Miaohui Wang

We propose a novel Enhanced Feature Aggregation and Selection network (EFASNet) for multi-person 2D human pose estimation. Due to enhanced feature representation, our method can well handle crowded, cluttered and occluded scenes. More…

计算机视觉与模式识别 · 计算机科学 2020-03-24 Xixia Xu , Qi Zou , Xue Lin

While NeRF has shown great success for neural reconstruction and rendering, its limited MLP capacity and long per-scene optimization times make it challenging to model large-scale indoor scenes. In contrast, classical 3D reconstruction…

计算机视觉与模式识别 · 计算机科学 2022-03-23 Xiaoshuai Zhang , Sai Bi , Kalyan Sunkavalli , Hao Su , Zexiang Xu

Multi-exposure image fusion (MEF) is an important area in computer vision and has attracted increasing interests in recent years. Apart from conventional algorithms, deep learning techniques have also been applied to multi-exposure image…

计算机视觉与模式识别 · 计算机科学 2020-07-31 Xingchen Zhang

Generalizable rendering of an animatable human avatar from sparse inputs relies on data priors and inductive biases extracted from training on large data to avoid scene-specific optimization and to enable fast reconstruction. This raises…

计算机视觉与模式识别 · 计算机科学 2025-02-14 Jing Wen , Alexander G. Schwing , Shenlong Wang

Dynamic multi-person mesh recovery has broad applications in sports broadcasting, virtual reality, and video games. However, current multi-view frameworks rely on a time-consuming camera calibration procedure. In this work, we focus on…

计算机视觉与模式识别 · 计算机科学 2024-12-30 Buzhen Huang , Jingyi Ju , Yuan Shu , Yangang Wang

Achieving high-quality High Dynamic Range (HDR) imaging on resource-constrained edge devices is a critical challenge in computer vision, as its performance directly impacts downstream tasks such as intelligent surveillance and autonomous…

计算机视觉与模式识别 · 计算机科学 2025-09-25 Yu-Shen Huang , Tzu-Han Chen , Cheng-Yen Hsiao , Shaou-Gang Miaou

Image super-resolution reconstruction achieves better results than traditional methods with the help of the powerful nonlinear representation ability of convolution neural network. However, some existing algorithms also have some problems,…

计算机视觉与模式识别 · 计算机科学 2022-05-30 Yuxi Cai , Huicheng Lai

High resolution images can be acquired using a non-regular sampling sensor which consists of an underlying low resolution sensor that is covered with a non-regular sampling mask. The reconstructed high resolution image is then obtained…

图像与视频处理 · 电气工程与系统科学 2022-04-08 Markus Jonscher , Karina Jaskolka , Jürgen Seiler , André Kaup