中文
相关论文

相关论文: CA-Edit: Causality-Aware Condition Adapter for Hig…

200 篇论文

Recent advancements in video generation highlight that realistic audio-visual synchronization is crucial for engaging content creation. However, existing video editing methods largely overlook audio-visual synchronization and lack the…

计算机视觉与模式识别 · 计算机科学 2026-05-05 Haojie Zheng , Shuchen Weng , Jingqi Liu , Siqi Yang , Boxin Shi , Xinlong Wang

Diffusion models have recently enabled precise and photorealistic facial editing across a wide range of semantic attributes. Beyond single-step modifications, a growing class of applications now demands the ability to analyze and track…

计算机视觉与模式识别 · 计算机科学 2025-06-03 Yule Zhu , Ping Liu , Zhedong Zheng , Wei Liu

Facial makeup editing aims to realistically transfer makeup from a reference to a target face. Existing methods often produce low-quality results with coarse makeup details and struggle to preserve both identity and makeup fidelity, mainly…

计算机视觉与模式识别 · 计算机科学 2025-08-11 Huadong Wu , Yi Fu , Yunhao Li , Yuan Gao , Kang Du

Recently, diffusion models have exhibited superior performance in the area of image inpainting. Inpainting methods based on diffusion models can usually generate realistic, high-quality image content for masked areas. However, due to the…

计算机视觉与模式识别 · 计算机科学 2024-12-03 Ruichen Wang , Junliang Zhang , Qingsong Xie , Chen Chen , Haonan Lu

Conditional diffusion models can create unseen images in various settings, aiding image interpolation. Interpolation in latent spaces is well-studied, but interpolation with specific conditions like text or poses is less understood. Simple…

计算机视觉与模式识别 · 计算机科学 2024-10-07 Qiyuan He , Jinghao Wang , Ziwei Liu , Angela Yao

A context-aware recommender system (CARS) applies sensing and analysis of user context to provide personalized services. The contextual information can be driven from sensors in order to improve the accuracy of the recommendations. Yet,…

机器学习 · 计算机科学 2022-08-10 Amit Livne , Eliad Shem Tov , Adir Solomon , Achiya Elyasaf , Bracha Shapira , Lior Rokach

Large diffusion transformers (DiTs) follow global editing instructions well but consistently leak local edits into unrelated regions, because joint-attention architectures offer no explicit channel telling the network where to apply the…

计算机视觉与模式识别 · 计算机科学 2026-05-01 Honghao Cai , Xiangyuan Wang , Yunhao Bai , Haohua Chen , Tianze Zhou , Runqi Wang , Wei Zhu , Yibo Chen , Xu Tang , Yao Hu , Zhen Li

In computed tomography (CT), achieving high image quality while minimizing radiation exposure remains a key clinical challenge. This paper presents CAPRI-CT, a novel causal-aware deep learning framework for Causal Analysis and Predictive…

计算机视觉与模式识别 · 计算机科学 2025-07-24 Sneha George Gnanakalavathy , Hairil Abdul Razak , Robert Meertens , Jonathan E. Fieldsend , Xujiong Ye , Mohammed M. Abdelsamea

Most existing methods view makeup transfer as transferring color distributions of different facial regions and ignore details such as eye shadows and blushes. Besides, they only achieve controllable transfer within predefined fixed regions.…

计算机视觉与模式识别 · 计算机科学 2022-07-21 Chenyu Yang , Wanrong He , Yingqing Xu , Yang Gao

Existing domain adaptation methods assume that domain discrepancies are caused by a few discrete attributes and variations, e.g., art, real, painting, quickdraw, etc. We argue that this is not realistic as it is implausible to define the…

计算机视觉与模式识别 · 计算机科学 2022-08-30 Yinsong Xu , Zhuqing Jiang , Aidong Men , Yang Liu , Qingchao Chen

In this work, we propose a novel generative method to identify the causal impact and apply it to prediction tasks. We conduct causal impact analysis using interventional and counterfactual perspectives. First, applying interventions, we…

机器学习 · 计算机科学 2025-09-03 Soma Bandyopadhyay , Sudeshna Sarkar

Understanding human affect from facial behavior requires not only accurate recognition but also structured reasoning over the latent dependencies that drive muscle activations and their expressive outcomes. Although Action Units (AUs) have…

计算机视觉与模式识别 · 计算机科学 2025-12-02 Guanyu Hu , Tangzheng Lian , Dimitrios Kollias , Oya Celiktutan , Xinyu Yang

Diffusion-based point editing methods have gained significant traction in image editing tasks due to their ability to manipulate image semantics and fine details by applying localized perturbations on the manifold of noise latent. However,…

计算机视觉与模式识别 · 计算机科学 2026-05-14 Haoyang Hu , Masataka Seo , Yen-Wei Chen

In this paper we present a novel multi-attribute face manipulation method based on textual descriptions. Previous text-based image editing methods either require test-time optimization for each individual image or are restricted to single…

计算机视觉与模式识别 · 计算机科学 2023-03-28 Hao Wang , Guosheng Lin , Ana García del Molino , Anran Wang , Jiashi Feng , Zhiqi Shen

Semantic image editing requires inpainting pixels following a semantic map. It is a challenging task since this inpainting requires both harmony with the context and strict compliance with the semantic maps. The majority of the previous…

计算机视觉与模式识别 · 计算机科学 2023-09-26 Hakan Sivuk , Aysegul Dundar

Current text-driven image editing methods typically follow one of two directions: relying on large-scale, high-quality editing pair datasets to improve editing precision and diversity, or exploring alternative dataset-free techniques.…

计算机视觉与模式识别 · 计算机科学 2025-05-27 Chenrui Ma , Xi Xiao , Tianyang Wang , Yanning Shen

The latest deep learning-based approaches have shown promising results for the challenging task of inpainting missing regions of an image. However, the existing methods often generate contents with blurry textures and distorted structures…

计算机视觉与模式识别 · 计算机科学 2019-07-05 Hongyu Liu , Bin Jiang , Yi Xiao , Chao Yang

Published research highlights the presence of demographic bias in automated facial attribute classification algorithms, particularly impacting women and individuals with darker skin tones. Existing bias mitigation techniques typically…

计算机视觉与模式识别 · 计算机科学 2024-09-02 Ayesha Manzoor , Ajita Rattani

Human skin detection in images is a widely studied topic of Computer Vision for which it is commonly accepted that analysis of pixel color or local patches may suffice. This is because skin regions appear to be relatively uniform and many…

计算机视觉与模式识别 · 计算机科学 2020-03-31 Aloisio Dourado , Frederico Guth , Teofilo Emidio de Campos , Li Weigang

In recent years, learning-based color and tone enhancement methods for photos have become increasingly popular. However, most learning-based image enhancement methods just learn a mapping from one distribution to another based on one…

计算机视觉与模式识别 · 计算机科学 2024-06-25 Shiqi Gao , Huiyu Duan , Xinyue Li , Kang Fu , Yicong Peng , Qihang Xu , Yuanyuan Chang , Jia Wang , Xiongkuo Min , Guangtao Zhai