English
Related papers

Related papers: A universal vision transformer for fast calorimete…

200 papers

Event-based vision sensors offer high time resolution, high dynamic range, and low power consumption, yet event-based vision models lag behind conventional frame-based vision methods. We argue that this gap is partly due to the lack of…

Computer Vision and Pattern Recognition · Computer Science 2026-04-01 Jens Egholm Pedersen , Dimitris Korakovounis , Jörg Conradt

The advent of Vision Transformers (ViTs) marks a substantial paradigm shift in the realm of computer vision. ViTs capture the global information of images through self-attention modules, which perform dot product computations among…

Computer Vision and Pattern Recognition · Computer Science 2024-06-04 Shuoxi Zhang , Hanpeng Liu , Stephen Lin , Kun He

Deep learning models have shown remarkable success in dermatological image analysis, offering potential for automated skin disease diagnosis. Previously, convolutional neural network(CNN) based architectures have achieved immense popularity…

Computer Vision and Pattern Recognition · Computer Science 2026-02-10 Rifat Sadik , Tanvir Rahman , Arpan Bhattacharjee , Bikash Chandra Halder , Ismail Hossain , Mridul Banik , Jia Uddin

The Vision Transformer (ViT) architecture has become widely recognized in computer vision, leveraging its self-attention mechanism to achieve remarkable success across various tasks. Despite its strengths, ViT's optimization remains…

Computer Vision and Pattern Recognition · Computer Science 2025-08-26 Haoyu Yun , Hamid Krim

Learning Bird's Eye View (BEV) representation from surrounding-view cameras is of great importance for autonomous driving. In this work, we propose a Geometry-guided Kernel Transformer (GKT), a novel 2D-to-BEV representation learning…

Computer Vision and Pattern Recognition · Computer Science 2022-06-10 Shaoyu Chen , Tianheng Cheng , Xinggang Wang , Wenming Meng , Qian Zhang , Wenyu Liu

Vision Transformers (ViTs) have emerged as the state-of-the-art architecture in representation learning, leveraging self-attention mechanisms to excel in various tasks. ViTs split images into fixed-size patches, constraining them to a…

Computer Vision and Pattern Recognition · Computer Science 2026-02-17 Aswathi Varma , Suprosanna Shit , Chinmay Prabhakar , Daniel Scholz , Hongwei Bran Li , Bjoern Menze , Daniel Rueckert , Benedikt Wiestler

Although Vision Transformers (ViTs) have achieved significant success, their intensive computations and substantial memory overheads challenge their deployment on edge devices. To address this, efficient ViTs have emerged, typically…

Hardware Architecture · Computer Science 2024-10-15 Yanbiao Liang , Huihong Shi , Zhongfeng Wang

Vision Transformer (ViT) is emerging as the state-of-the-art architecture for image recognition. While recent studies suggest that ViTs are more robust than their convolutional counterparts, our experiments find that ViTs trained on…

Computer Vision and Pattern Recognition · Computer Science 2022-04-05 Chengzhi Mao , Lu Jiang , Mostafa Dehghani , Carl Vondrick , Rahul Sukthankar , Irfan Essa

We study whether machine-learning models for fast calorimeter simulations can learn meaningful representations of calorimeter signatures that account for variations in the full particle detector's configuration. This may open new…

Instrumentation and Detectors · Physics 2025-08-29 Johannes Erdmann , Jonas Kann , Florian Mausolf , Peter Wissmann

Recently, Vision Transformers (ViTs) have attracted a lot of attention in the field of computer vision. Generally, the powerful representative capacity of ViTs mainly benefits from the self-attention mechanism, which has a high computation…

Computer Vision and Pattern Recognition · Computer Science 2023-10-12 Deli Yu , Teng Xi , Jianwei Li , Baopu Li , Gang Zhang , Haocheng Feng , Junyu Han , Jingtuo Liu , Errui Ding , Jingdong Wang

Vision transformers (ViTs) have found only limited practical use in processing images, in spite of their state-of-the-art accuracy on certain benchmarks. The reason for their limited use include their need for larger training datasets and…

Computer Vision and Pattern Recognition · Computer Science 2022-01-26 Pranav Jeevan , Amit sethi

In this paper, we design and train a Generative Image-to-text Transformer, GIT, to unify vision-language tasks such as image/video captioning and question answering. While generative models provide a consistent network architecture between…

Computer Vision and Pattern Recognition · Computer Science 2022-12-19 Jianfeng Wang , Zhengyuan Yang , Xiaowei Hu , Linjie Li , Kevin Lin , Zhe Gan , Zicheng Liu , Ce Liu , Lijuan Wang

Calorimeters with a high granularity are a fundamental requirement of the Particle Flow paradigm. This paper focuses on the prototype of a hadron calorimeter with analog readout, consisting of thirty-eight scintillator layers alternating…

Instrumentation and Detectors · Physics 2014-06-17 C. Adloff , J. Blaha , J. -J. Blaising , C. Drancourt , A. Espargilière , R. Gaglione , N. Geffroy , Y. Karyotakis , J. Prast , G. Vouters , K. Francis , J. Repond , J. Schlereth , J. Smith , L. Xia , E. Baldolemar , J. Li , S. T. Park , M. Sosebee , A. P. White , J. Yu , T. Buanes , G. Eigen , Y. Mikami , N. K. Watson , G. Mavromanolakis , M. A. Thomson , D. R. Ward , W. Yan , D. Benchekroun , A. Hoummada , Y. Khoulaki , J. Apostolakis , A. Dotti , G. Folger , V. Ivantchenko , V. Uzhinskiy , M. Benyamna , C. Cârloganu , F. Fehr , P. Gay , S. Manen , L. Royer , G. C. Blazey , A. Dyshkant , J. G. R. Lima , V. Zutshi , J. -Y. Hostachy , L. Morin , U. Cornett , D. David , G. Falley , K. Gadow , P. Göttlicher , C. Günter , B. Hermberg , S. Karstensen , F. Krivan , A. -I. Lucaci-Timoce , S. Lu , B. Lutz , S. Morozov , V. Morgunov , M. Reinecke , F. Sefkow , P. Smirnov , M. Terwort , A. Vargas-Trevino , N. Feege , E. Garutti , I. Marchesinik , M. Ramilli , P. Eckert , T. Harion , A. Kaplan , H. -Ch. Schultz-Coulon , W. Shen , R. Stamen , B. Bilki , E. Norbeck , Y. Onel , G. W. Wilson , K. Kawagoe , P. D. Dauncey , A. -M. Magnan , V. Bartsch , M. Wing , F. Salvatore , E. Calvo Alamillo , M. -C. Fouz , J. Puerta-Pelayo , B. Bobchenko , M. Chadeeva , M. Danilov , A. Epifantsev , O. Markin , R. Mizuk , E. Novikov , V. Popov , V. Rusinov , E. Tarkovsky , N. Kirikova , V. Kozlov , P. Smirnov , Y. Soloviev , P. Buzhan , A. Ilyin , V. Kantserov , V. Kaplin , A. Karakash , E. Popova , V. Tikhomirov , C. Kiesling , K. Seidel , F. Simon , C. Soldner , M. Szalay , M. Tesar , L. Weuste , M. S. Amjad , J. Bonis , S. Callier , S. Conforti di Lorenzo , P. Cornebise , Ph. Doublet , F. Dulucq , J. Fleury , T. Frisson , N. van der Kolk , H. Li , G. Martin-Chassard , F. Richard , Ch. de la Taille , R. Pöschl , L. Raux , J. Rouëné , N. Seguin-Moreau , M. Anduze , V. Boudry , J-C. Brient , D. Jeans , P. Mora de Freitas , G. Musat , M. Reinhard , M. Ruan , H. Videau , B. Bulanek , J. Zacek , J. Cvach , P. Gallus , M. Havranek , M. Janata , J. Kvasnicka , D. Lednicky , M. Marcisovsky , I. Polak , J. Popule , L. Tomasek , M. Tomasek , P. Ruzicka , P. Sicho , J. Smolik , V. Vrba , J. Zalesak , B. Belhorma , H. Ghazlane , T. Takeshita , S. Uozumi , M. Götze , O. Hartbrich , J. Sauer , S. Weber , C. Zeitnitz

Vision Transformers (ViT) have made many breakthroughs in computer vision tasks. However, considerable redundancy arises in the spatial dimension of an input image, leading to massive computational costs. Therefore, We propose a…

Computer Vision and Pattern Recognition · Computer Science 2022-11-22 Mengzhao Chen , Mingbao Lin , Ke Li , Yunhang Shen , Yongjian Wu , Fei Chao , Rongrong Ji

Vision Transformers (ViTs), with their ability to model long-range dependencies through self-attention mechanisms, have become a standard architecture in computer vision. However, the interpretability of these models remains a challenge. To…

Computer Vision and Pattern Recognition · Computer Science 2025-01-09 Walid Bousselham , Angie Boggust , Sofian Chaybouti , Hendrik Strobelt , Hilde Kuehne

Vision Transformers (ViT) have achieved remarkable success in large-scale image recognition. They split every 2D image into a fixed number of patches, each of which is treated as a token. Generally, representing an image with more tokens…

Computer Vision and Pattern Recognition · Computer Science 2021-10-27 Yulin Wang , Rui Huang , Shiji Song , Zeyi Huang , Gao Huang

Plant phenology-the study of recurrent life cycle events-is essential for understanding ecosystem dynamics and their responses to climate change impacts. While Unmanned Aerial Vehicles (UAVs) and near-surface cameras enable high-resolution…

Vision Transformers (ViTs) have demonstrated remarkable potential in image processing tasks by utilizing self-attention mechanisms to capture global relationships within data. However, their scalability is hindered by significant…

Machine Learning · Computer Science 2026-02-25 Huy Trinh , Rebecca Ma , Zeqi Yu , Tahsin Reza

Recently, vision transformers (ViTs) have superseded convolutional neural networks in numerous applications, including classification, detection, and segmentation. However, the high computational requirements of ViTs hinder their widespread…

Computer Vision and Pattern Recognition · Computer Science 2024-05-20 Jemin Lee , Yongin Kwon , Sihyeong Park , Misun Yu , Jeman Park , Hwanjun Song

Operator learning, which aims to approximate maps between infinite-dimensional function spaces, is an important area in scientific machine learning with applications across various physical domains. Here we introduce the Continuous Vision…

Machine Learning · Computer Science 2025-02-18 Sifan Wang , Jacob H Seidman , Shyam Sankaran , Hanwen Wang , George J. Pappas , Paris Perdikaris