English
Related papers

Related papers: A software-based focus system for wide-field optic…

200 papers

Diffusion models generate samples by estimating the score function of the target distribution at various noise levels. The model is trained using samples drawn from the target distribution by progressively adding noise. Previous sample…

Machine Learning · Computer Science 2025-10-28 Syamantak Kumar , Dheeraj Nagaraj , Purnamrita Sarkar

We propose a framework that extends Blender to exploit Structure from Motion (SfM) and Multi-View Stereo (MVS) techniques for image-based modeling tasks such as sculpting or camera and motion tracking. Applying SfM allows us to determine…

Computer Vision and Pattern Recognition · Computer Science 2021-04-06 Sebastian Bullinger , Christoph Bodensteiner , Michael Arens

Focusing light through dynamically varying heterogeneous media is a sought-after goal with important applications ranging from free-space communication to nano-surgery. The underlying challenge is to control the optical wavefront with a…

Traditional glass-based optics are typically optimized for narrow spectral bands, such as the visible (400-700nm) or shortwave infrared (1000-1800nm). While the emergence of VIS-SWIR sensors (400-1700nm) offers transformative potential,…

Image and Video Processing · Electrical Eng. & Systems 2026-05-04 Vishwanath Saragadam , Niki Nezakati , Amit Roy-Chowdhury , Vivek Boominathan

Benchmark datasets in computer vision often contain off-topic images, near duplicates, and label errors, leading to inaccurate estimates of model performance. In this paper, we revisit the task of data cleaning and formalize it as either a…

Zero-shot learning (ZSL) which aims to recognize unseen classes with no labeled training sample, efficiently tackles the problem of missing labeled data in image retrieval. Nowadays there are mainly two types of popular methods for ZSL to…

Computer Vision and Pattern Recognition · Computer Science 2018-10-30 Gang Yang , Jinlu Liu , Xirong Li

Background subtraction (BGS) aims to extract all moving objects in the video frames to obtain binary foreground segmentation masks. Deep learning has been widely used in this field. Compared with supervised-based BGS methods, unsupervised…

Computer Vision and Pattern Recognition · Computer Science 2023-03-28 Yongqi An , Xu Zhao , Tao Yu , Haiyun Guo , Chaoyang Zhao , Ming Tang , Jinqiao Wang

Effective image deblurring typically relies on large and fully paired datasets of blurred and corresponding sharp images. However, obtaining such accurately aligned data in the real world poses a number of difficulties, limiting the…

Image and Video Processing · Electrical Eng. & Systems 2025-10-21 Alok Panigrahi , Jayaprakash Katual , Satish Mulleti

We propose a simple yet effective zero-shot framework for subject-driven image generation using a vanilla Flux model. By framing the task as grid-based image completion and simply replicating the subject image(s) in a mosaic layout, we…

Computer Vision and Pattern Recognition · Computer Science 2025-04-22 Hao Kang , Stathi Fotiadis , Liming Jiang , Qing Yan , Yumin Jia , Zichuan Liu , Min Jin Chong , Xin Lu

Quantitative analysis of the dynamics of tiny cellular and sub-cellular structures, known as particles, in time-lapse cell microscopy sequences requires the development of a reliable multi-target tracking method capable of tracking numerous…

Computer Vision and Pattern Recognition · Computer Science 2015-07-24 Seyed Hamid Rezatofighi , Stephen Gould , Ba Tuong Vo , Ba-Ngu Vo , Katarina Mele , Richard Hartley

The distribution of dark and luminous matter can be mapped around galaxies that gravitationally lens background objects into arcs or Einstein rings. New surveys will soon observe hundreds of thousands of galaxy lenses, and current,…

Luminescence imaging is invaluable for studying biological and material systems, particularly when advanced protocols that exploit temporal dynamics are employed. However, implementing such protocols often requires custom instrumentation,…

Instrumentation and Detectors · Physics 2025-09-16 Ian Coghill , Alienor Lahlou , Andrea Lodetti , Shizue Matsubara , Johann Boucle , Thomas Le Saux , Ludovic Jullien

Counting plant organs such as heads or tassels from outdoor imagery is a popular benchmark computer vision task in plant phenotyping, which has been previously investigated in the literature using state-of-the-art supervised deep learning…

Computer Vision and Pattern Recognition · Computer Science 2020-07-21 Jordan Ubbens , Tewodros Ayalew , Steve Shirtliffe , Anique Josuttes , Curtis Pozniak , Ian Stavness

Artefacts compromise clinical decision-making in the use of medical time series. Pulsatile waveforms offer probabilities for accurate artefact detection, yet most approaches rely on supervised manners and overlook patient-level distribution…

Signal Processing · Electrical Eng. & Systems 2025-05-01 Xuhang Chen , Ihsane Olakorede , Stefan Yu Bögli , Wenhao Xu , Erta Beqiri , Xuemeng Li , Chenyu Tang , Zeyu Gao , Shuo Gao , Ari Ercole , Peter Smielewski

Fast, volumetric imaging over large scales has been a long-standing goal in biological microscopy. Scanning techniques such as fluorescence confocal microscopy can acquire 2D images at high resolution and high speed, but extending the…

Pre-training image representations from the raw text about images enables zero-shot vision transfer to downstream tasks. Through pre-training on millions of samples collected from the internet, multimodal foundation models, such as CLIP,…

Machine Learning · Computer Science 2024-03-18 Chenguang Wang , Ruoxi Jia , Xin Liu , Dawn Song

The prominence of generalized foundation models in vision-language integration has witnessed a surge, given their multifarious applications. Within the natural domain, the procurement of vision-language datasets to construct these…

Computer Vision and Pattern Recognition · Computer Science 2024-09-12 Keumgang Cha , Donggeun Yu , Junghoon Seo

Brightfield microscopy of unstained live cells is challenging due to low contrast, dynamic morphology, uneven illumination, and lack of labels. Deep learning achieved SOTA performance on stained, high-contrast images but needs large labeled…

Image and Video Processing · Electrical Eng. & Systems 2025-10-15 Surajit Das , Pavel Zun

Vision language models such as CLIP have shown remarkable performance in zero shot classification, but remain susceptible to spurious correlations, where irrelevant visual features influence predictions. Existing debiasing methods often…

Computer Vision and Pattern Recognition · Computer Science 2025-11-04 Fangyu Wu , Yujun Cai

Editing images with diffusion models under strict training-free constraints remains a significant challenge. While recent optimisation-based methods achieve strong zero-shot edits from text, they struggle to preserve identity and capture…

Computer Vision and Pattern Recognition · Computer Science 2026-03-04 Niki Foteinopoulou , Ignas Budvytis , Stephan Liwicki