中文
相关论文

相关论文: UltraZoom: Generating Gigapixel Images from Regula…

200 篇论文

Diffusion models for image generation have been a subject of increasing interest due to their ability to generate diverse, high-quality images. Image generation has immense potential in medical imaging because open-source medical images are…

计算机视觉与模式识别 · 计算机科学 2025-02-13 Benoit Freiche , Anthony El-Khoury , Ali Nasiri-Sarvi , Mahdi S. Hosseini , Damien Garcia , Adrian Basarab , Mathieu Boily , Hassan Rivaz

In this work, we present a geometry-based grasping algorithm that is capable of efficiently generating both top and side grasps for unknown objects, using a single view RGB-D camera, and of selecting the most promising one. We demonstrate…

机器人学 · 计算机科学 2019-07-19 Brice Denoun , Beatriz Leon , Claudio Zito , Rustam Stolkin , Lorenzo Jamone , Miles Hansard

High resolution reconstruction of complicated objects from incomplete and noisy data can be achieved by solving modulation equations iteratively under physical constraints. This direct demodulation method is a powerful technique for dealing…

天体物理学 · 物理学 2009-11-10 Ti-Pei Li , Mei Wu

This paper presents UltraEdit, a large-scale (approximately 4 million editing samples), automatically generated dataset for instruction-based image editing. Our key idea is to address the drawbacks in existing image editing datasets like…

计算机视觉与模式识别 · 计算机科学 2024-12-20 Haozhe Zhao , Xiaojian Ma , Liang Chen , Shuzheng Si , Rujie Wu , Kaikai An , Peiyu Yu , Minjia Zhang , Qing Li , Baobao Chang

This paper proposes a technique for the unsupervised detection and tracking of arbitrary objects in videos. It is intended to reduce the need for detection and localization methods tailored to specific object types and serve as a general…

机器学习 · 统计学 2012-10-12 Willie Neiswanger , Frank Wood

Despite recent advancements in neural 3D reconstruction, the dependence on dense multi-view captures restricts their broader applicability. Additionally, 3D scene generation is vital for advancing embodied AI and world models, which depend…

计算机视觉与模式识别 · 计算机科学 2025-11-19 Yuxin Zhang , Ziyu Lu , Hongbo Duan , Keyu Fan , Pengting Luo , Peiyu Zhuang , Mengyu Yang , Houde Liu

Currently, personalized image generation methods mostly require considerable time to finetune and often overfit the concept resulting in generated images that are similar to custom concepts but difficult to edit by prompts. We propose an…

计算机视觉与模式识别 · 计算机科学 2024-03-19 Yuxuan Zhang , Yiren Song , Jinpeng Yu , Han Pan , Zhongliang Jing

Current cameras are capable of recording high resolution video. While viewing on a mobile device, a user can manually zoom into this high resolution video to get more detailed view of objects and activities. However, manual zooming is not…

多媒体 · 计算机科学 2019-09-27 Mukesh Saini , Benjamin Guthier , Hao Kuang , Dwarikanath Mahapatra , Abdulmotaleb El Saddik

Diffusion models have recently achieved significant success in various image manipulation tasks, including image super-resolution and perceptual quality enhancement. Pretrained text-to-image models, such as Stable Diffusion, have exhibited…

计算机视觉与模式识别 · 计算机科学 2025-10-16 Sanchar Palit , Subhasis Chaudhuri , Biplab Banerjee

Generating new images with desired properties (e.g. new view/poses) from source images has been enthusiastically pursued recently, due to its wide range of potential applications. One way to ensure high-quality generation is to use multiple…

计算机视觉与模式识别 · 计算机科学 2022-02-03 Jiawei Lu , He Wang , Tianjia Shao , Yin Yang , Kun Zhou

In this paper, we propose an intuitive method to recover background from multiple images. The implementation consists of three stages: model initialization, model update, and background output. We consider the pixels whose values change…

计算机视觉与模式识别 · 计算机科学 2019-11-05 Lei Gao , Yixing Huang , Andreas Maier

Generating visible-like face images from thermal images is essential to perform manual and automatic cross-spectrum face recognition. We successfully propose a solution based on cascaded refinement network that, unlike previous works,…

计算机视觉与模式识别 · 计算机科学 2019-10-22 Naser Damer , Fadi Boutros , Khawla Mallat , Florian Kirchbuchner , Jean-Luc Dugelay , Arjan Kuijper

In recent years, image blending has gained popularity for its ability to create visually stunning content. However, the current image blending algorithms mainly have the following problems: manually creating image blending masks requires a…

计算机视觉与模式识别 · 计算机科学 2023-11-30 Haochen Xue , Mingyu Jin , Chong Zhang , Yuxuan Huang , Qian Weng , Xiaobo Jin

Existing methods for image synthesis utilized a style encoder based on stacks of convolutions and pooling layers to generate style codes from input images. However, the encoded vectors do not necessarily contain local information of the…

计算机视觉与模式识别 · 计算机科学 2021-12-20 Jonghyun Kim , Gen Li , Cheolkon Jung , Joongkyu Kim

Multispectral image fusion is a computer vision process that is essential to remote sensing. For applications such as dehazing and object detection, there is a need to offer solutions that can perform in real-time on any type of scene.…

计算机视觉与模式识别 · 计算机科学 2023-05-09 Nati Ofir , Jean-Christophe Nebel

Recent deep image-to-image translation techniques allow fast generation of face images from freehand sketches. However, existing solutions tend to overfit to sketches, thus requiring professional sketches or even edge maps as input. To…

图形学 · 计算机科学 2020-06-09 Shu-Yu Chen , Wanchao Su , Lin Gao , Shihong Xia , Hongbo Fu

Text to Image Synthesis refers to the process of automatic generation of a photo-realistic image starting from a given text and is revolutionizing many real-world applications. In order to perform such process it is necessary to exploit…

We address the challenge of creating 3D assets for household articulated objects from a single image. Prior work on articulated object creation either requires multi-view multi-state input, or only allows coarse control over the generation…

计算机视觉与模式识别 · 计算机科学 2025-03-21 Jiayi Liu , Denys Iliash , Angel X. Chang , Manolis Savva , Ali Mahdavi-Amiri

We address the ambiguities in the super-resolution problem under translation. We demonstrate that combinations of low-resolution images at different scales can be used to make the super-resolution problem well posed. Such differences in…

计算机视觉与模式识别 · 计算机科学 2026-04-24 Daniel Fu , Gabby Litterio , Pedro Felzenszwalb , Rashid Zia

Super-resolution algorithms often struggle with images from surveillance environments due to adverse conditions such as unknown degradation, variations in pose, irregular illumination, and occlusions. However, acquiring multiple images,…

计算机视觉与模式识别 · 计算机科学 2024-10-22 Marcelo dos Santos , Rayson Laroca , Rafael O. Ribeiro , João C. Neves , David Menotti