English
Related papers

Related papers: CSP-Atlas: Concept-Specific Neural Circuits in a S…

200 papers

Explicit Chain-of-Thought improves the reasoning performance of large language models but often incurs high inference cost due to verbose token-level traces. While recent approaches reduce this overhead via concise prompting or step…

Computation and Language · Computer Science 2026-03-09 Yunlong Chu , Minglai Shao , Yuhang Liu , Bing Hao , Yumeng Lin , Jialu Wang , Ruijie Wang

We introduce the Convolutional Set Transformer (CST), a novel neural architecture designed to process image sets of arbitrary cardinality that are visually heterogeneous yet share high-level semantics - such as a common category, scene, or…

Computer Vision and Pattern Recognition · Computer Science 2025-09-30 Federico Chinello , Giacomo Boracchi

Automated construction is one of the most promising areas that can improve efficiency, reduce costs and minimize errors in the process of building construction. In this paper, a comparative analysis of three neural network models for…

Computer Vision and Pattern Recognition · Computer Science 2025-03-31 Ivan Beleacov

Weakly-supervised semantic segmentation aims to assign each pixel a semantic category under weak supervisions, such as image-level tags. Most of existing weakly-supervised semantic segmentation methods do not use any feedback from…

Computer Vision and Pattern Recognition · Computer Science 2019-05-30 Zhengqiang Zhang , Shujian Yu , Shi Yin , Qinmu Peng , Xinge You

Sparse coding, which is the decomposition of a vector using only a few basis elements, is widely used in machine learning and image processing. The basis set, also called dictionary, is learned to adapt to specific data. This approach has…

Machine Learning · Computer Science 2011-12-13 Louise Benoît , Julien Mairal , Francis Bach , Jean Ponce

We propose a new unbiased threshold for network analysis named the Cluster-Span Threshold (CST). This is based on the clustering coefficient, C, following logic that a balance of `clustering' to `spanning' triples results in a useful…

Neurons and Cognition · Quantitative Biology 2016-04-11 Keith Smith , Hamed Azami , Mario A. Parra , John M. Starr , Javier Escudero

A widely used strategy to discover and understand language model mechanisms is circuit analysis. A circuit is a minimal subgraph of a model's computation graph that executes a specific task. We identify a gap in existing circuit discovery…

Machine Learning · Computer Science 2025-02-10 Tal Haklay , Hadas Orgad , David Bau , Aaron Mueller , Yonatan Belinkov

A fundamental question in interpretability research is to what extent neural networks, particularly language models, implement reusable functions through subnetworks that can be composed to perform more complex tasks. Recent advances in…

Machine Learning · Computer Science 2025-06-24 Philipp Mondorf , Sondre Wold , Barbara Plank

We study how reliably sparse autoencoders (SAEs) support claims about reasoning-related internal features in large language models. We first give a stylized analysis showing that sparsity-regularized decoding can preferentially retain…

Machine Learning · Computer Science 2026-05-19 George Ma , Zhongyuan Liang , Irene Y. Chen , Somayeh Sojoudi

Recently, large models, such as Vision Transformer and BERT, have garnered significant attention due to their exceptional performance. However, their extensive computational requirements lead to considerable power and hardware resource…

Hardware Architecture · Computer Science 2025-01-15 Zhengke Li , Wendong Mao , Siyu Zhang , Qiwei Dong , Zhongfeng Wang

We present the performance of a semantic segmentation network, SparseSSNet, that provides pixel-level classification of MicroBooNE data. The MicroBooNE experiment employs a liquid argon time projection chamber for the study of neutrino…

Instrumentation and Detectors · Physics 2021-04-07 MicroBooNE collaboration , P. Abratenko , M. Alrashed , R. An , J. Anthony , J. Asaadi , A. Ashkenazi , S. Balasubramanian , B. Baller , C. Barnes , G. Barr , V. Basque , L. Bathe-Peters , O. Benevides Rodrigues , S. Berkman , A. Bhanderi , A. Bhat , M. Bishai , A. Blake , T. Bolton , L. Camilleri , D. Caratelli , I. Caro Terrazas , R. Castillo Fernandez , F. Cavanna , G. Cerati , Y. Chen , E. Church , D. Cianci , J. M. Conrad , M. Convery , L. Cooper-Troendle , J. I. Crespo-Anadon , M. Del Tutto , S. R. Dennis , D. Devitt , R. Diurba , R. Dorrill , K. Duffy , S. Dytman , B. Eberly , A. Ereditato , J. J. Evans , G. A. Fiorentini Aguirre , R. S. Fitzpatrick , B. T. Fleming , N. Foppiani , D. Franco , A. P. Furmanski , D. Garcia-Gamez , S. Gardiner , G. Ge , S. Gollapinni , O. Goodwin , E. Gramellini , P. Green , H. Greenlee , W. Gu , R. Guenette , P. Guzowski , L. Hagaman , E. Hall , P. Hamilton , O. Hen , G. A. Horton-Smith , A. Hourlier , R. Itay , C. James , J. Jan de Vries , X. Ji , L. Jiang , J. H. Jo , R. A. Johnson , Y. J. Jwa , N. Kamp , N. Kaneshige , G. Karagiorgi , W. Ketchum , B. Kirby , M. Kirby , T. Kobilarcik , I. Kreslo , R. LaZur , I. Lepetic , K. Li , Y. Li , B. R. Littlejohn , W. C. Louis , X. Luo , A. Marchionni , C. Mariani , D. Marsden , J. Marshall , J. Martin-Albo , D. A. Martinez Caicedo , K. Mason , A. Mastbaum , N. McConkey , V. Meddage , T. Mettler , K. Miller , J. Mills , K. Mistry , T. Mohayai , A. Mogan , J. Moon , M. Mooney , A. F. Moor , C. D. Moore , L. Mora Lepin , J. Mousseau , M. Murphy , D. Naples , A. Navrer-Agasson , R. K. Neely , P. Nienaber , J. Nowak , O. Palamara , V. Paolone , A. Papadopoulou , V. Papavassiliou , S. F. Pate , A. Paudel , Z. Pavlovic , E. Piasetzky , I. Ponce-Pinto , S. Prince , X. Qian , J. L. Raaf , V. Radeka , A. Rafique , M. Reggiani-Guzzo , L. Ren , L. Rochester , J. Rodriguez Rondon , H. E. Rogers , M. Rosenberg , M. Ross-Lonergan , B. Russell , G. Scanavini , D. W. Schmitz , A. Schukraft , W. Seligman , M. H. Shaevitz , R. Sharankova , J. Sinclair , A. Smith , E. L. Snider , M. Soderberg , S. Soldner-Rembold , S. R. Soleti , P. Spentzouris , J. Spitz , M. Stancari , J. St. John , T. Strauss , K. Sutton , S. Sword-Fehlberg , A. M. Szelc , N. Tagg , W. Tang , K. Terao , C. Thorpe , M. Toups , Y. -T. Tsai , M. A. Uchida , T. Usher , W. Van De Pontseele , B. Viren , M. Weber , H. Wei , Z. Williams , S. Wolbers , T. Wongjirad , M. Wospakrik , W. Wu , E. Yandel , T. Yang , G. Yarbrough , L. E. Yates , G. P. Zeller , J. Zennamo , C. Zhang

Boolean circuits form the foundational computational substrate of symmetric cryptography, yet the exploration of their architectural design space has remained largely confined to a handful of canonical paradigms - SPN, Feistel networks, and…

Cryptography and Security · Computer Science 2026-05-01 Arnaud Valence

Attention improves representation learning over RNNs, but its discrete nature limits continuous-time (CT) modeling. We introduce Neuronal Attention Circuit (NAC), a novel, biologically inspired CT-Attention mechanism that reformulates…

Artificial Intelligence · Computer Science 2026-01-07 Waleed Razzaq , Izis Kanjaraway , Yun-Bo Zhao

Sparse convolutional neural networks (CNNs) have gained significant traction over the past few years as sparse CNNs can drastically decrease the model size and computations, if exploited befittingly, as compared to their dense counterparts.…

Hardware Architecture · Computer Science 2021-11-10 Mahmood Azhar Qureshi , Arslan Munir

We propose a novel weakly-supervised semantic segmentation algorithm based on Deep Convolutional Neural Network (DCNN). Contrary to existing weakly-supervised approaches, our algorithm exploits auxiliary segmentation annotations available…

Computer Vision and Pattern Recognition · Computer Science 2015-12-29 Seunghoon Hong , Junhyuk Oh , Bohyung Han , Honglak Lee

Deep vision models have achieved remarkable classification performance by leveraging a hierarchical architecture in which human-interpretable concepts emerge through the composition of individual neurons across layers. Given the distributed…

Computer Vision and Pattern Recognition · Computer Science 2025-08-05 Dahee Kwon , Sehyun Lee , Jaesik Choi

The functional and structural representation of the brain as a complex network is marked by the fact that the comparison of noisy and intrinsically correlated high-dimensional structures between experimental conditions or groups shuns…

Neurons and Cognition · Quantitative Biology 2013-10-25 Tommaso Furlanello , Marco Cristoforetti , Cesare Furlanello , Giuseppe Jurman

Large language models handle single-turn generation well, but multi-turn interactions still require the model to reconstruct user intent and task state from an expanding token history because internal representations do not persist across…

Computation and Language · Computer Science 2025-12-11 Vishwas Hegde , Vindhya Shigehalli

To enhance developer productivity, all modern integrated development environments (IDEs) include code suggestion functionality that proposes likely next tokens at the cursor. While current IDEs work well for statically-typed languages,…

Neural and Evolutionary Computing · Computer Science 2016-11-28 Avishkar Bhoopchand , Tim Rocktäschel , Earl Barr , Sebastian Riedel

Semantic segmentation, like other fields of computer vision, has seen a remarkable performance advance by the use of deep convolution neural networks. However, considering that neighboring pixels are heavily dependent on each other, both…

Computer Vision and Pattern Recognition · Computer Science 2017-08-08 Hyojin Park , Jisoo Jeong , Youngjoon Yoo , Nojun Kwak
‹ Prev 1 3 4 5 6 7 10 Next ›