中文
相关论文

相关论文: Photonic Modes Prediction via Multi-Modal Diffusio…

200 篇论文

Structural coloration is commonly modeled using wave optics for reliable and photorealistic rendering of natural, quasi-periodic and complex nanostructures. Such models often rely on dense, preliminary or preprocessed data to accurately…

图形学 · 计算机科学 2025-07-03 Narayan Kandel , Daljit Singh J. S. Dhillon

We demonstrate that embedding physics-driven constraints into machine learning process can dramatically improve accuracy and generalizability of the resulting model. Physics-informed learning is illustrated on the example of analysis of…

计算物理 · 物理学 2021-12-16 Abantika Ghosh , Mohannad Elhamod , Jie Bu , Wei-Cheng Lee , Anuj Karpatne , Viktor A Podolskiy

Transformer models trained on massive text corpora have become the de facto models for a wide range of natural language processing tasks. However, learning effective word representations for function words remains challenging. Multimodal…

计算与语言 · 计算机科学 2022-10-25 Shashank Sonkar , Naiming Liu , Richard G. Baraniuk

Models for inferring monocular shape of surfaces with diffuse reflection -- shape from shading -- ought to produce distributions of outputs, because there are fundamental mathematical ambiguities of both continuous (e.g., bas-relief) and…

计算机视觉与模式识别 · 计算机科学 2024-11-05 Xinran Nicole Han , Todd Zickler , Ko Nishino

Vision-language pretrained models have seen remarkable success, but their application to safety-critical settings is limited by their lack of interpretability. To improve the interpretability of vision-language models such as CLIP, we…

计算机视觉与模式识别 · 计算机科学 2024-06-25 Ying Wang , Tim G. J. Rudner , Andrew Gordon Wilson

Diffusion models for image generation function by progressively adding noise to an image set and training a model to separate out the signal from the noise. The noise profile used by these models is white noise -- that is, noise based on…

计算机视觉与模式识别 · 计算机科学 2025-07-09 Andrew Randono

We present a novel learning-based method to build a differentiable computational model of a real fluorescence microscope. Our model can be used to calibrate a real optical setup directly from data samples and to engineer point spread…

图像与视频处理 · 电气工程与系统科学 2023-06-12 Josue Page , Paolo Favaro

Text-to-image diffusion models are now capable of generating images that are often indistinguishable from real images. To generate such images, these models must understand the semantics of the objects they are asked to generate. In this…

计算机视觉与模式识别 · 计算机科学 2023-12-29 Eric Hedlin , Gopal Sharma , Shweta Mahajan , Hossam Isack , Abhishek Kar , Andrea Tagliasacchi , Kwang Moo Yi

Understanding the structure of real data is paramount in advancing modern deep-learning methodologies. Natural data such as images are believed to be composed of features organized in a hierarchical and combinatorial manner, which neural…

机器学习 · 统计学 2024-12-25 Antonio Sclocchi , Alessandro Favero , Matthieu Wyart

Load forecasting plays a pivotal role in the safe and stable operation of power systems. Conventional deep learning methods often struggle to adapt to few-shot scenarios frequently encountered in industrial applications. Existing…

信号处理 · 电气工程与系统科学 2026-05-12 Yuxuan Chen , Shuo Dai , Ruoyi Xu , Haipeng Xie

In this article we address the theoretical study of a multiscale drift-diffusion (DD) model for the description of photoconversion mechanisms in organic solar cells. The multiscale nature of the formulation is based on the co-presence of…

应用物理 · 物理学 2018-02-14 Maurizio Verri , Matteo Porro , Riccardo Sacco , Sandro Salsa

We address the reconstruction of the full photon distribution of multimode fields generated by seeded parametric down-conversion (PDC). Our scheme is based on on/off avalanche photodetection assisted by maximum-likelihood (MaxLik)…

量子物理 · 物理学 2009-11-13 G. Brida , M. Genovese , A. Meda , S. Olivares , M. G. A. Paris , F. Piacentini

State-of-the-art learned reconstruction methods often rely on black-box modules that, despite their strong performance, raise questions about their interpretability and robustness. Here, we build on a recently proposed image reconstruction…

图像与视频处理 · 电气工程与系统科学 2026-05-19 Joshua Schulz , David Schote , Christoph Kolbitsch , Kostas Papafitsoros , Andreas Kofler

Multimodal semantic understanding often has to deal with uncertainty, which means the obtained messages tend to refer to multiple targets. Such uncertainty is problematic for our interpretation, including inter- and intra-modal uncertainty.…

计算机视觉与模式识别 · 计算机科学 2023-07-21 Yatai Ji , Junjie Wang , Yuan Gong , Lin Zhang , Yanru Zhu , Hongfa Wang , Jiaxing Zhang , Tetsuya Sakai , Yujiu Yang

Edge intelligence is constrained by the energy and latency costs of shuttling data through electronic memory hierarchies. Optical systems offer a fundamentally different computational regime: once an input wavefront is launched into a…

硬件体系结构 · 计算机科学 2026-04-20 Prakul Sunil Hiremath

Pre-trained diffusion models have demonstrated remarkable proficiency in synthesizing images across a wide range of scenarios with customizable prompts, indicating their effective capacity to capture universal features. Motivated by this,…

计算机视觉与模式识别 · 计算机科学 2024-11-22 Yuxiang Ji , Boyong He , Chenyuan Qu , Zhuoyue Tan , Chuan Qin , Liaoni Wu

When multimode optical fibers are perturbed, the data that is transmitted through them is scrambled. This presents a major difficulty for many possible applications, such as multimode fiber-based telecommunication and endoscopy. To overcome…

图像与视频处理 · 电气工程与系统科学 2022-02-17 Shachar Resisi , Sebastien M. Popoff , Yaron Bromberg

Advancements in generative models have sparked significant interest in generating images while adhering to specific structural guidelines. Scene graph to image generation is one such task of generating images which are consistent with the…

计算机视觉与模式识别 · 计算机科学 2024-07-23 Rameshwar Mishra , A V Subramanyam

The dislocation created in the topological material lays the foundation of many significant findings to control light but requires delicate fabrication of the material. To extend its flexibility and reconfigurability, we propose the…

光学 · 物理学 2024-11-18 Danying Yu , Kun Ding , Xianfeng Chen , Luqi Yuan

Professional-grade software applications are powerful but complicated$-$expert users can achieve impressive results, but novices often struggle to complete even basic tasks. Photo editing is a prime example: after loading a photo, the user…

‹ 上一页 1 8 9 10 下一页 ›