中文
相关论文

相关论文: BokehFlow: Depth-Free Controllable Bokeh Rendering…

200 篇论文

Bokeh effect is an optical phenomenon that offers a pleasant visual experience, typically generated by high-end cameras with wide aperture lenses. The task of bokeh effect transformation aims to produce a desired effect in one set of lenses…

计算机视觉与模式识别 · 计算机科学 2023-06-08 Zhihao Yang , Wenyi Lian , Siyuan Lai

As mobile cameras with compact optics are unable to produce a strong bokeh effect, lots of interest is now devoted to deep learning-based solutions for this task. In this Mobile AI challenge, the target was to develop an efficient…

Current gaze input methods for VR headsets predominantly utilize the gaze ray as a pointing cursor, often neglecting depth information in it. This study introduces FocusFlow, a novel gaze interaction technique that integrates focal depth…

人机交互 · 计算机科学 2023-08-15 Chenyang Zhang , Tiansu Chen , Rohan Nedungadi , Eric Shaffer , Elahe Soltanaghai

Image composition targets at synthesizing a realistic composite image from a pair of foreground and background images. Recently, generative composition methods are built on large pretrained diffusion models to generate composite images,…

计算机视觉与模式识别 · 计算机科学 2023-08-22 Bo Zhang , Yuxuan Duan , Jun Lan , Yan Hong , Huijia Zhu , Weiqiang Wang , Li Niu

A shallow depth-of-field image keeps the subject in focus, and the foreground and background contexts blurred. This effect requires much larger lens apertures than those of smartphone cameras. Conventional methods acquire RGB-D images and…

计算机视觉与模式识别 · 计算机科学 2022-07-15 Meng-Lin Wu , Venkata Ravi Kiran Dayana , Hau Hwang

We propose a novel training-free image generation algorithm that precisely controls the occlusion relationships between objects in an image. Existing image generation methods typically rely on prompts to influence occlusion, which often…

计算机视觉与模式识别 · 计算机科学 2025-08-12 Xiaohang Zhan , Dingming Liu

Dense 3D facial motion capture from only monocular in-the-wild pairs of RGB images is a highly challenging problem with numerous applications, ranging from facial expression recognition to facial reenactment. In this work, we propose…

计算机视觉与模式识别 · 计算机科学 2020-05-18 Mohammad Rami Koujan , Anastasios Roussos , Stefanos Zafeiriou

Depth in the real world is rarely singular. Transmissive materials create layered ambiguities that confound conventional perception systems. Existing models remain passive; conventional approaches typically estimate static depth maps…

计算机视觉与模式识别 · 计算机科学 2026-03-26 Junhong Min , Jimin Kim , Minwook Kim , Cheol-Hui Min , Youngpil Jeon , Minyong Choi

In this paper, we tackle the problem of monocular bokeh synthesis, where we attempt to render a shallow depth of field image from a single all-in-focus image. Unlike in DSLR cameras, this effect can not be captured directly in mobile…

计算机视觉与模式识别 · 计算机科学 2022-08-29 Brian Lee , Fei Lei , Huaijin Chen , Alexis Baudron

Automatic black-and-white image sequence colorization while preserving character and object identity (ID) is a complex task with significant market demand, such as in cartoon or comic series colorization. Despite advancements in visual…

计算机视觉与模式识别 · 计算机科学 2025-05-06 Junhao Zhuang , Xuan Ju , Zhaoyang Zhang , Yong Liu , Shiyi Zhang , Chun Yuan , Ying Shan

Performing facial expression transfer under one-shot setting has been increasing in popularity among research community with a focus on precise control of expressions. Existing techniques showcase compelling results in perceiving…

计算机视觉与模式识别 · 计算机科学 2024-04-24 Siddharth Nijhawan , Takuya Yashima , Tamaki Kojima

Recent advances in flow-based generative models have enabled training-free, text-guided image editing by inverting an image into its latent noise and regenerating it under a new target conditional guidance. However, existing methods…

计算机视觉与模式识别 · 计算机科学 2026-04-03 Thinh Dao , Zhen Wang , Kien T. Pham , Long Chen

Optical flow estimation is a crucial subfield of computer vision, serving as a foundation for video tasks. However, the real-world robustness is limited by animated synthetic datasets for training. This introduces domain gaps when applied…

计算机视觉与模式识别 · 计算机科学 2025-06-10 Yingping Liang , Ying Fu , Yutao Hu , Wenqi Shao , Jiaming Liu , Debing Zhang

This paper introduces innovative solutions to enhance spatial controllability in diffusion models reliant on text queries. We first introduce vision guidance as a foundational spatial cue within the perturbed distribution. This…

计算机视觉与模式识别 · 计算机科学 2024-12-02 Zipeng Qi , Guoxi Huang , Chenyang Liu , Fei Ye

Coarse-to-fine autoregressive modeling has recently shown strong promise for visuomotor policy learning, combining the inference efficiency of autoregressive methods with the global trajectory coherence of diffusion-based policies. However,…

机器人学 · 计算机科学 2026-03-31 Daichi Yashima , Koki Seno , Shuhei Kurita , Yusuke Oda , Komei Sugiura

Training-free image editing has attracted increasing attention for its efficiency and independence from training data. However, existing approaches predominantly rely on inversion-reconstruction trajectories, which impose an inherent…

计算机视觉与模式识别 · 计算机科学 2026-02-03 Menglin Han , Zhangkai Ni

Recent advancements in video generation models have significantly improved their ability to follow text prompts. However, the customization of dynamic visual effects, defined as temporally evolving and appearance-driven visual phenomena…

计算机视觉与模式识别 · 计算机科学 2026-03-24 Rui Zhao , Mike Zheng Shou

The Bokeh Effect is one of the most desirable effects in photography for rendering artistic and aesthetic photos. Usually, it requires a DSLR camera with different aperture and shutter settings and certain photography skills to generate…

计算机视觉与模式识别 · 计算机科学 2021-05-18 Saikat Dutta , Sourya Dipta Das , Nisarg A. Shah , Anil Kumar Tiwari

A photo captured with bokeh effect often means objects in focus are sharp while the out-of-focus areas are all blurred. DSLR can easily render this kind of effect naturally. However, due to the limitation of sensors, smartphones cannot…

计算机视觉与模式识别 · 计算机科学 2021-04-13 Ming Qian , Congyu Qiao , Jiamin Lin , Zhenyu Guo , Chenghua Li , Cong Leng , Jian Cheng

Diffusion models (DMs) excel in photorealism, image editing, and solving inverse problems, aided by classifier-free guidance and image inversion techniques. However, rectified flow models (RFMs) remain underexplored for these tasks.…

计算机视觉与模式识别 · 计算机科学 2024-12-03 Maitreya Patel , Song Wen , Dimitris N. Metaxas , Yezhou Yang