English
Related papers

Related papers: 3One2: One-step Regression Plus One-step Diffusion…

200 papers

Video face restoration aims to enhance degraded face videos into high-quality results with realistic facial details, stable identity, and temporal coherence. Recent diffusion-based methods have brought strong generative priors to…

Computer Vision and Pattern Recognition · Computer Science 2026-04-17 Zheng Chen , Bowen Chai , Rongjun Gao , Mingtao Nie , Xi Li , Bingnan Duan , Jianping Fang , Xiaohong Liu , Linghe Kong , Yulun Zhang

Coded aperture snapshot spectral imaging (CASSI) is a technique used to reconstruct three-dimensional hyperspectral images (HSIs) from one or several two-dimensional projection measurements. However, fewer projection measurements or more…

Image and Video Processing · Electrical Eng. & Systems 2024-10-28 Qile Zhao , Xianhong Zhao , Xu Ma , Xudong Chen , Gonzalo R. Arce

Video snapshot compressive imaging (SCI) enables the reconstruction of dynamic scenes from a single snapshot measurement. Recently, NeRF-based methods have shown promising reconstruction performance. However, such methods typically adopt…

Computer Vision and Pattern Recognition · Computer Science 2026-05-01 Yubo Dong , Danhua Liu , Anqi Li , Zhenyuan Lin

Image super-resolution is a process to enhance image resolution. It is widely used in medical imaging, satellite imaging, target recognition, etc. In this paper, we conduct continuous modeling and assume that the unknown image intensity…

Computer Vision and Pattern Recognition · Computer Science 2015-03-13 Liang-Jian Deng , Weihong Guo , Ting-Zhu Huang

Video Captioning (VC) is a challenging multi-modal task since it requires describing the scene in language by understanding various and complex videos. For machines, the traditional VC follows the…

Computer Vision and Pattern Recognition · Computer Science 2024-01-11 Jianqiao Sun , Yudi Su , Hao Zhang , Ziheng Cheng , Zequn Zeng , Zhengjue Wang , Bo Chen , Xin Yuan

Recent breakthroughs in video diffusion models have significantly accelerated the development of video editing techniques. However, existing methods often rely on inpainting video frames based on masked input, which requires extracting the…

Computer Vision and Pattern Recognition · Computer Science 2026-05-15 Zheng Hui , Yunlong Bai

The ability of snapshot compressive imaging (SCI) systems to efficiently capture high-dimensional (HD) data has led to an inverse problem, which consists of recovering the HD signal from the compressed and noisy measurement. While…

Image and Video Processing · Electrical Eng. & Systems 2023-03-01 Yaping Zhao , Siming Zheng , Xin Yuan

Coded aperture snapshot spectral imaging (CASSI) is a promising technique to capture the three-dimensional hyperspectral image (HSI) using a single coded two-dimensional (2D) measurement, in which algorithms are used to perform the inverse…

Image and Video Processing · Electrical Eng. & Systems 2021-01-01 Wei He , Naoto Yokoya , Xin Yuan

Digital cameras consume ~0.1 microjoule per pixel to capture and encode video, resulting in a power usage of ~20W for a 4K sensor operating at 30 fps. Imagining gigapixel cameras operating at 100-1000 fps, the current processing model is…

Computer Vision and Pattern Recognition · Computer Science 2025-09-11 Miao Cao , Siming Zheng , Lishun Wang , Ziyang Chen , David Brady , Xin Yuan

Inversion by Direct Iteration (InDI) is a new formulation for supervised image restoration that avoids the so-called "regression to the mean" effect and produces more realistic and detailed images than existing regression-based methods. It…

Image and Video Processing · Electrical Eng. & Systems 2024-02-05 Mauricio Delbracio , Peyman Milanfar

It is well known that many open-released foundational diffusion models have difficulty in generating images that substantially depart from average brightness, despite such images being present in the training data. This is due to an…

Computer Vision and Pattern Recognition · Computer Science 2023-11-28 Minghui Hu , Jianbin Zheng , Chuanxia Zheng , Chaoyue Wang , Dacheng Tao , Tat-Jen Cham

Optimizing a deep neural network is a fundamental task in computer vision, yet direct training methods often suffer from over-fitting. Teacher-student optimization aims at providing complementary cues from a model trained previously, but…

Computer Vision and Pattern Recognition · Computer Science 2018-12-04 Chenglin Yang , Lingxi Xie , Chi Su , Alan L. Yuille

The rapid progress in artificial intelligence-generated content (AIGC), especially with diffusion models, has significantly advanced development of high-quality video generation. However, current video diffusion models exhibit demanding…

Computer Vision and Pattern Recognition · Computer Science 2024-11-05 Zheng Zhan , Yushu Wu , Yifan Gong , Zichong Meng , Zhenglun Kong , Changdi Yang , Geng Yuan , Pu Zhao , Wei Niu , Yanzhi Wang

Coded aperture snapshot spectral imaging (CASSI) makes it possible to recover 3D hyperspectral data from a single 2D image. However, the reconstruction problem is severely underdetermined and efforts to improve the compression ratio…

Optics · Physics 2022-08-03 Jiri Hlubucek , Jakub Lukes , Jan Vaclavik , Karel Zidek

Diffusion models have significantly advanced the fields of image, audio, and video generation, but they depend on an iterative sampling process that causes slow generation. To overcome this limitation, we propose consistency models, a new…

Machine Learning · Computer Science 2023-06-01 Yang Song , Prafulla Dhariwal , Mark Chen , Ilya Sutskever

Snapshot compressive imaging (SCI) captures high-dimensional data efficiently by compressing it into two-dimensional observations and reconstructing high-dimensional data from two-dimensional observations with various algorithms. The…

Image and Video Processing · Electrical Eng. & Systems 2025-03-06 Takashi Matsuda , Ryo Hayakawa , Youji Iiguni

The usually reported pixel resolution of single pixel imaging (SPI) varies between $32 \times 32$ and $256 \times 256$ pixels falling far below imaging standards with classical methods. Low resolution results from the trade-off between the…

Optics · Physics 2022-06-22 Rafał Stojek , Anna Pastuszczak , Piotr Wróbel , Rafał Kotyński

Recent advances in Video Foundation Models (VFMs) have revolutionized human-centric video synthesis, yet fine-grained and independent editing of subjects and scenes remains a critical challenge. Recent attempts to incorporate richer…

Computer Vision and Pattern Recognition · Computer Science 2026-04-02 Fengyuan Yang , Luying Huang , Jiazhi Guan , Quanwei Yang , Dongwei Pan , Jianglin Fu , Haocheng Feng , Wei He , Kaisiyuan Wang , Hang Zhou , Angela Yao

Few-shot action recognition aims to enable models to quickly learn new action categories from limited labeled samples, addressing the challenge of data scarcity in real-world applications. Current research primarily addresses three core…

Computer Vision and Pattern Recognition · Computer Science 2025-12-23 Xiaoyang Li , Mingming Lu , Ruiqi Wang , Hao Li , Zewei Le

Video capture is limited by the trade-off between spatial and temporal resolution: when capturing videos of high temporal resolution, the spatial resolution decreases due to bandwidth limitations in the capture system. Achieving both high…

Graphics · Computer Science 2018-06-14 Ana Serrano , Elena Garces , Diego Gutierrez , Belen Masia
‹ Prev 1 4 5 6 7 8 10 Next ›