Related papers: Problems and related results in Algebraic Vision a…
In this short note we would like to show how it is possible to use techniques introduced in the theory of local dynamics of holomorphic germs tangent to the identity to study global meromorphic self-maps of the complex projective space. In…
Generative Adversarial Networks (GANs) have recently achieved impressive results for many real-world applications, and many GAN variants have emerged with improvements in sample quality and training stability. However, they have not been…
Detecting objects and estimating their pose remains as one of the major challenges of the computer vision research community. There exists a compromise between localizing the objects and estimating their viewpoints. The detector ideally…
Piece-wise 3D planar reconstruction provides holistic scene understanding of man-made environments, especially for indoor scenarios. Most recent approaches focused on improving the segmentation and reconstruction results by introducing…
Algebraic models for the reconstruction problem in X-ray computed tomography (CT) provide a flexible framework that applies to many measurement geometries. For large-scale problems we need to use iterative solvers, and we need stopping…
Whether is ground-based or space-based, any optical instrument suffers from some amount of optical geometric distortion. Recently, the diffraction-limited image quality afforded by space-based telescopes and by Adaptive Optics (AO)…
Given the complexities inherent in visual scenes, such as object occlusion, a comprehensive understanding often requires observation from multiple viewpoints. Existing multi-viewpoint object-centric learning methods typically employ random…
Multimodal large language models (MLLMs), such as GPT-4o, Gemini, LLaVA, and Flamingo, have made significant progress in integrating visual and textual modalities, excelling in tasks like visual question answering (VQA), image captioning,…
In this work, we aim to improve the 3D reasoning ability of Transformers in multi-view 3D human pose estimation. Recent works have focused on end-to-end learning-based transformer designs, which struggle to resolve geometric information…
In this paper, we address the problem of reconstructing an object's surface from a single image using generative networks. First, we represent a 3D surface with an aggregation of dense point clouds from multiple views. Each point cloud is…
This paper is a documentation of author's reseach, focusing on the topic Grassmann Algebra spanning over July, August 2025 under mentorship provided by DRP Turkiye 2025. Grassmann algebra is a fundamental structure in mathematics with…
The Riemannian geometry is one of the main theoretical pieces in Modern Mathematics and Physics. The study of Riemann Geometry in the relevant literature is performed by using a well defined analytical path. Usually it starts from the…
Let $Gr$ be a component of the Grassmann manifold of a $C^*$-algebra, presented as the unitary orbit of a given orthogonal projection $Gr=Gr(P)$. There are several natural connections in this manifold, and we first show that they all agree…
We prove a root stack valuative criterion for good moduli space maps and for gerbes for reductive groups under some mild assumptions on the residue characteristic. We give several applications to parahoric extension for torsors, rational…
Computed medical imaging systems require a computational reconstruction procedure for image formation. In order to recover a useful estimate of the object to-be-imaged when the recorded measurements are incomplete, prior knowledge about the…
Object pose refinement is essential for robust object pose estimation. Previous work has made significant progress towards instance-level object pose refinement. Yet, category-level pose refinement is a more challenging problem due to large…
Gabor wavelet is an essential tool for image analysis and computer vision tasks. Local structure tensors with multiple scales are widely used in local feature extraction. Our research indicates that the current corner detection method based…
Reconstructing 3D representations from 2D inputs is a fundamental task in computer vision and graphics, serving as a cornerstone for understanding and interacting with the physical world. While traditional methods achieve high fidelity,…
Robust visual localization under a wide range of viewing conditions is a fundamental problem in computer vision. Handling the difficult cases of this problem is not only very challenging but also of high practical relevance, e.g., in the…
Image-based 3D reconstruction is one of the most important tasks in Computer Vision with many solutions proposed over the last few decades. The objective is to extract metric information i.e. the geometry of scene objects directly from…