English
Related papers

Related papers: DepthArb: Training-Free Depth-Arbitrated Generatio…

200 papers

Diffusion-based text-to-image (T2I) models have made remarkable progress in generating photorealistic and semantically rich images. However, when the target concepts lie in low-density regions of the training distribution, these models…

Computer Vision and Pattern Recognition · Computer Science 2026-03-20 Kwanyoung Lee , SeungJu Cha , Yebin Ahn , Hyunwoo Oh , Sungho Koh , Dong-Jin Kim

Recent diffusion-based approaches have made substantial progress in image layer decomposition. However, accurately decomposing complex natural images remains challenging due to difficulties in occlusion completion, robust layer…

Computer Vision and Pattern Recognition · Computer Science 2026-05-13 Binhao Wang , Shihao Zhao , Bo Cheng , Qiuyu Ji , Yuhang Ma , Liebucha Wu , Shanyuan Liu , Dawei Leng , Yuhui Yin

Monocular depth estimation has seen significant advances through discriminative approaches, yet their performance remains constrained by the limitations of training datasets. While generative approaches have addressed this challenge by…

Computer Vision and Pattern Recognition · Computer Science 2025-07-01 Bulat Gabdullin , Nina Konovalova , Nikolay Patakin , Dmitry Senushkin , Anton Konushin

Occlusion is a long-standing problem in computer vision, particularly in instance segmentation. ACM MMSports 2023 DeepSportRadar has introduced a dataset that focuses on segmenting human subjects within a basketball context and a…

Computer Vision and Pattern Recognition · Computer Science 2023-11-22 Son Nguyen , Mikel Lainsa , Hung Dao , Daeyoung Kim , Giang Nguyen

We focus on explicitly learning disentangled representation for natural image generation, where the underlying spatial structure and the rendering on the structure can be independently controlled respectively, yet using no tuple…

Machine Learning · Computer Science 2019-10-01 Guang-Yuan Hao , Hong-Xing Yu , Wei-Shi Zheng

Conditional inference on arbitrary subsets of variables is a core problem in probabilistic inference with important applications such as masked language modeling and image inpainting. In recent years, the family of Any-Order Autoregressive…

Machine Learning · Computer Science 2022-10-25 Andy Shih , Dorsa Sadigh , Stefano Ermon

This research delves into the problem of interactive editing of human motion generation. Previous motion diffusion models lack explicit modeling of the word-level text-motion correspondence and good explainability, hence restricting their…

Computer Vision and Pattern Recognition · Computer Science 2025-01-23 Ling-Hao Chen , Shunlin Lu , Wenxun Dai , Zhiyang Dou , Xuan Ju , Jingbo Wang , Taku Komura , Lei Zhang

Diffusion Transformers (DiT)-based video generation models with 3D full attention exhibit strong generative capabilities. Trajectory control represents a user-friendly task in the field of controllable video generation. However, existing…

Computer Vision and Pattern Recognition · Computer Science 2025-09-30 Cheng Lei , Jiayu Zhang , Yue Ma , Xinyu Wang , Long Chen , Liang Tang , Yiqiang Yan , Fei Su , Zhicheng Zhao

Causal representation learning seeks to uncover causal relationships among high-level latent variables from low-level, entangled, and noisy observations. Existing approaches often either rely on deep neural networks, which lack…

Methodology · Statistics 2026-03-27 Wenjin Zhang , Yixin Wang , Yuqi Gu

This paper presents Deep ARTMAP, a novel extension of the ARTMAP architecture that generalizes the self-consistent modular ART (SMART) architecture to enable hierarchical learning (supervised and unsupervised) across arbitrary…

Machine Learning · Computer Science 2025-03-12 Niklas M. Melton , Leonardo Enzo Brito da Silva , Sasha Petrenko , Donald. C. Wunsch

Volumetric imaging by fluorescence microscopy is often limited by anisotropic spatial resolution from inferior axial resolution compared to the lateral resolution. To address this problem, here we present a deep-learning-enabled…

Computer Vision and Pattern Recognition · Computer Science 2022-07-06 Hyoungjun Park , Myeongsu Na , Bumju Kim , Soohyun Park , Ki Hean Kim , Sunghoe Chang , Jong Chul Ye

Deep Learning of neural networks has gained prominence in multiple life-critical applications like medical diagnoses and autonomous vehicle accident investigations. However, concerns about model transparency and biases persist. Explainable…

Computer Vision and Pattern Recognition · Computer Science 2025-11-25 Pedro Valois , Koichiro Niinuma , Kazuhiro Fukui

Template-free retrosynthesis methods treat the task as black-box sequence generation, limiting learning efficiency, while semi-template approaches rely on rigid reaction libraries that constrain generalization. We address this gap with a…

Machine Learning · Computer Science 2026-02-16 Chenguang Wang , Zihan Zhou , Lei Bai , Tianshu Yu

Text-to-image synthesis has achieved high-quality results with recent advances in diffusion models. However, text input alone has high spatial ambiguity and limited user controllability. Most existing methods allow spatial control through…

Computer Vision and Pattern Recognition · Computer Science 2023-10-31 Yuki Endo

Local feature matching is an essential component in many visual applications. In this work, we propose OAMatcher, a Tranformer-based detector-free method that imitates humans behavior to generate dense and accurate matches. Firstly,…

Computer Vision and Pattern Recognition · Computer Science 2024-06-18 Kun Dai , Tao Xie , Ke Wang , Zhiqiang Jiang , Ruifeng Li , Lijun Zhao

While many diffusion models perform well when controlling particular aspects such as style, character, and interaction, they struggle with fine-grained control due to dataset limitations and intricate model architecture design. This paper…

Computer Vision and Pattern Recognition · Computer Science 2026-01-29 Conghan Yue , Zhengwei Peng , Shiyan Du , Zhi Ji , Chuangjian Cai , Le Wan , Dongyu Zhang

Diffusion models have demonstrated remarkable capabilities in visual content generation but remain challenging to deploy due to their high computational cost during inference. This computational burden primarily arises from the quadratic…

Computer Vision and Pattern Recognition · Computer Science 2025-03-28 Ye Tian , Xin Xia , Yuxi Ren , Shanchuan Lin , Xing Wang , Xuefeng Xiao , Yunhai Tong , Ling Yang , Bin Cui

Deep generative neural networks have proven effective at both conditional and unconditional modeling of complex data distributions. Conditional generation enables interactive control, but creating new controls often requires expensive…

Machine Learning · Computer Science 2017-12-25 Jesse Engel , Matthew Hoffman , Adam Roberts

Deep convolutional neural networks (DCNNs) are powerful models that yield impressive results at object classification. However, recent work has shown that they do not generalize well to partially occluded objects and to mask attacks. In…

Computer Vision and Pattern Recognition · Computer Science 2020-01-30 Adam Kortylewski , Qing Liu , Huiyu Wang , Zhishuai Zhang , Alan Yuille

We formalize concepts around geometric occlusion in 2D images (i.e., ignoring semantics), and propose a novel unified formulation of both occlusion boundaries and occlusion orientations via a pixel-pair occlusion relation. The former…

Computer Vision and Pattern Recognition · Computer Science 2020-07-24 Xuchong Qiu , Yang Xiao , Chaohui Wang , Renaud Marlet
‹ Prev 1 4 5 6 7 8 10 Next ›