English
Related papers

Related papers: A multilayer level-set method for eikonal-based tr…

200 papers

Fine-tuning multilingual sequence-to-sequence large language models (msLLMs) has shown promise in developing neural machine translation (NMT) systems for low-resource languages (LRLs). However, conventional single-stage fine-tuning methods…

Computation and Language · Computer Science 2025-03-31 Sarubi Thillainathan , Songchen Yuan , En-Shiun Annie Lee , Sanath Jayasena , Surangika Ranathunga

Multimodal large language model (MLLM) inference splits into two phases with opposing hardware demands: vision encoding is compute-bound, while language generation is memory-bandwidth-bound. We show that under standard transformer KV…

Machine Learning · Computer Science 2026-03-16 Donglin Yu

Multimodal Large Language Models (MLLMs) have endowed LLMs with the ability to perceive and understand multi-modal signals. However, most of the existing MLLMs mainly adopt vision encoders pretrained on coarsely aligned image-text pairs,…

Computer Vision and Pattern Recognition · Computer Science 2023-11-28 Gongwei Chen , Leyang Shen , Rui Shao , Xiang Deng , Liqiang Nie

Reconstructing and understanding 3D structures from a limited number of images is a well-established problem in computer vision. Traditional methods usually break this task into multiple subtasks, each requiring complex transformations…

Computer Vision and Pattern Recognition · Computer Science 2024-11-01 Zhiwen Fan , Jian Zhang , Wenyan Cong , Peihao Wang , Renjie Li , Kairun Wen , Shijie Zhou , Achuta Kadambi , Zhangyang Wang , Danfei Xu , Boris Ivanovic , Marco Pavone , Yue Wang

We propose a new method of instance-level microtubule (MT) tracking in time-lapse image series using recurrent attention. Our novel deep learning algorithm segments individual MTs at each frame. Segmentation results from successive frames…

Computer Vision and Pattern Recognition · Computer Science 2020-01-22 Samira Masoudi , Afsaneh Razi , Cameron H. G. Wright , Jay C. Gatlin , Ulas Bagci

Although multimodal large language models (MLLMs) have achieved promising results on a wide range of vision-language tasks, their ability to perceive and understand human faces is rarely explored. In this work, we comprehensively evaluate…

Computer Vision and Pattern Recognition · Computer Science 2024-10-29 Haomiao Sun , Mingjie He , Tianheng Lian , Hu Han , Shiguang Shan

We study an inverse problem for Light Sheet Fluorescence Microscopy (LSFM), where the density of fluorescent molecules needs to be reconstructed. Our first step is to present a mathematical model to describe the measurements obtained by an…

Analysis of PDEs · Mathematics 2020-08-26 Evelyn Cueva , Matias Courdurier , Axel Osses , Victor Castañeda , Benjamin Palacios , Steffen Härtel

Multimodal LLMs (MLLMs) have reached remarkable levels of proficiency in understanding multimodal inputs. However, understanding and interpreting the behavior of such complex models is a challenging task, not to mention the dynamic shifts…

Artificial Intelligence · Computer Science 2025-08-14 Pegah Khayatan , Mustafa Shukor , Jayneel Parekh , Arnaud Dapogny , Matthieu Cord

Diffusion priors have been used for blind face restoration (BFR) by fine-tuning diffusion models (DMs) on restoration datasets to recover low-quality images. However, the naive application of DMs presents several key limitations. (i) The…

Computer Vision and Pattern Recognition · Computer Science 2025-03-25 Senmao Li , Kai Wang , Joost van de Weijer , Fahad Shahbaz Khan , Chun-Le Guo , Shiqi Yang , Yaxing Wang , Jian Yang , Ming-Ming Cheng

In this paper we present a multilevel projection-based iterative scheme for solving thermal radiative transfer problems that performs iteration cycles on the high-order Boltzmann transport equation (BTE) and low-order moment equations.…

Numerical Analysis · Mathematics 2026-03-18 Joseph M. Coale , Dmitriy Y. Anistratov

A phase retrieval technique using a spatial light modulator (SLM) and a phase diffuser for a fast reconstruction of smooth wave fronts is demonstrated experimentally. Diffuse illumination of a smooth test object with the aid of a phase…

Optics · Physics 2015-06-11 Mostafa Agour , Percival F. Almoro , Claas Falldorf

We present a robust numerical method for solving incompressible, immiscible two-phase flows. The method extends the monolithic phase conservative level set method with embedded redistancing by Quezada de Luna et al. [38] and a semi-implicit…

Numerical Analysis · Mathematics 2019-03-19 Manuel Quezada de Luna , J. Haydel Collins , Christopher E. Kees

Least-Squares Reverse-Time Migration (LSRTM) is a method that seismologists utilize to compute a high-resolution subsurface image. Nevertheless, LSRTM is a computationally demanding problem. One way to reduce the computational costs of the…

Geophysics · Physics 2025-09-12 Aydin Shoja , Joost van der Neut , Kees Wapenaar

Previous methods based on 3DCNN, convLSTM, or optical flow have achieved great success in video salient object detection (VSOD). However, they still suffer from high computational costs or poor quality of the generated saliency maps. To…

Computer Vision and Pattern Recognition · Computer Science 2024-01-02 Xing Zhao , Haoran Liang , Peipei Li , Guodao Sun , Dongdong Zhao , Ronghua Liang , Xiaofei He

Contemporary Vision-Language Models (VLMs) achieve strong performance on a wide range of tasks by pairing a vision encoder with a pre-trained language model, fine-tuned for visual-text inputs. Yet despite these gains, it remains unclear how…

Computer Vision and Pattern Recognition · Computer Science 2026-02-10 Lachin Naghashyar , Hunar Batra , Ashkan Khakzar , Philip Torr , Ronald Clark , Christian Schroeder de Witt , Constantin Venhoff

The development of Multi-modal Large Language Models (MLLMs) enhances Large Language Models (LLMs) with the ability to perceive data formats beyond text, significantly advancing a range of downstream applications, such as visual question…

Computer Vision and Pattern Recognition · Computer Science 2024-12-03 Minbin Huang , Runhui Huang , Han Shi , Yimeng Chen , Chuanyang Zheng , Xiangguo Sun , Xin Jiang , Zhenguo Li , Hong Cheng

Quantum computing shows substantial potential in accelerating simulations and alleviating memory bottlenecks in computational fluid dynamics (CFD), owing to its inherent properties of superposition and entanglement. The lattice Boltzmann…

Quantum Physics · Physics 2026-03-03 Yang Xiao , Liming Yang , Chang Shu , Yinjie Du

Multimodal Large Language Models (MLLMs) rely on strong linguistic reasoning inherited from their base language models. However, multimodal instruction fine-tuning paradoxically degrades this text's reasoning capability, undermining…

Computation and Language · Computer Science 2026-01-13 Zijing Wang , Yongkang Liu , Mingyang Wang , Ercong Nie , Deyuan Chen , Zhengjie Zhao , Shi Feng , Daling Wang , Xiaocui Yang , Yifei Zhang , Hinrich Schütze

Multimodal irregular time series (MITS) consist of asynchronous and irregularly sampled observations from heterogeneous numerical and textual channels. In healthcare, for example, patients' electronic health records (EHR) include irregular…

Machine Learning · Computer Science 2026-05-14 Hsing-Huan Chung , Shijun Li , Yoav Wald , Xing Han , Suchi Saria , Joydeep Ghosh

We develop a framework that allows the use of the multi-level Monte Carlo (MLMC) methodology (Giles2015) to calculate expectations with respect to the invariant measure of an ergodic SDE. In that context, we study the (over-damped) Langevin…

Numerical Analysis · Mathematics 2019-08-13 Michael B. Giles , Mateusz B. Majka , Lukasz Szpruch , Sebastian Vollmer , Konstantinos Zygalakis
‹ Prev 1 8 9 10 Next ›