中文
相关论文

相关论文: Layers at Similar Depths Generate Similar Activati…

200 篇论文

Neural models learn representations of high-dimensional data on low-dimensional manifolds. Multiple factors, including stochasticities in the training process, model architectures, and additional inductive biases, may induce different…

机器学习 · 计算机科学 2025-12-02 Hanlin Yu , Berfin Inal , Georgios Arvanitidis , Soren Hauberg , Francesco Locatello , Marco Fumero

The increased use of Large Language Models (LLMs) in geography raises substantial questions about the safety of integrating these tools across a wide range of processes and analyses, given our very limited understanding of their inner…

机器学习 · 计算机科学 2026-05-15 Stef De Sabbata , Rahul Baiju , Stefano Mizzaro , Kevin Roitero

Dynamic multilayer networks frequently represent the structure of multiple co-evolving relations; however, statistical models are not well-developed for this prevalent network type. Here, we propose a new latent space model for dynamic…

统计方法学 · 统计学 2021-03-25 Joshua Daniel Loyal , Yuguo Chen

Complex networks, such as transportation networks, social networks, or biological networks, capture the complex system they model often by representing only one type of interactions. In real world systems, there may be many different…

物理与社会 · 物理学 2020-10-16 Blaž Škrlj , Benjamin Renoust

Many real-world network are multilayer, with nontrivial correlations across layers. Here we show that these correlations amplify geometry in networks. We focus on mutual clustering--a measure of the amount of triangles that are present in…

物理与社会 · 物理学 2026-02-24 Jasper van der Kolk , Dmitri Krioukov , Marián Boguñá , M. Ángeles Serrano

Large language models (LLMs) are increasingly used to model human social behavior, with recent research exploring their ability to simulate social dynamics. Here, we test whether LLMs mirror human behavior in social dilemmas, where…

社会与信息网络 · 计算机科学 2024-11-18 Jin Han , Balaraju Battu , Ivan Romić , Talal Rahwan , Petter Holme

Large Language Models (LLMs) have demonstrated impressive generalization capabilities across various tasks, but their claim to practical relevance is still mired by concerns on their reliability. Recent works have proposed examining the…

机器学习 · 计算机科学 2025-07-08 Waiss Azizian , Michael Kirchhof , Eugene Ndiaye , Louis Bethune , Michal Klein , Pierre Ablin , Marco Cuturi

The coexistence of multiple types of interactions within social, technological and biological networks has moved the focus of the physics of complex systems towards a multiplex description of the interactions between their constituents.…

Real networks often form interacting parts of larger and more complex systems. Examples can be found in different domains, ranging from the Internet to structural and functional brain networks. Here, we show that these multiplex systems are…

物理与社会 · 物理学 2017-09-11 Kaj-Kolja Kleineberg , Marian Boguna , M. Angeles Serrano , Fragkiskos Papadopoulos

While large language models (LLMs) are trained purely on textual data, prior work has shown that their internal representations can exhibit rich geometric structure in embedding space. Building on this line of work, we investigate whether…

人工智能 · 计算机科学 2026-05-28 Simardeep Singh , Paras Chopra

Weight matrices in deep networks exhibit geometric continuity -- principal singular vectors of adjacent layers point in similar directions. While this property has been widely observed, its origin remains unexplained. Through experiments on…

机器学习 · 计算机科学 2026-05-07 Kyungwon Jeong , Won-Gi Paeng , Honggyo Suh

The geometric structure of latent representations in large language models (LLMs) is an active area of research, driven in part by its implications for model transparency and AI safety. Existing literature has focused mainly on general…

机器学习 · 计算机科学 2026-04-14 Benjamin J. Choi , Melanie Weber

Large language models (LLMs) often exhibit sycophantic behaviors -- such as excessive agreement with or flattery of the user -- but it is unclear whether these behaviors arise from a single mechanism or multiple distinct processes. We…

计算与语言 · 计算机科学 2026-03-24 Daniel Vennemeyer , Phan Anh Duong , Tiffany Zhan , Tianyu Jiang

We investigate the extent to which an LLM's hidden-state geometry can be recovered from its behavior in psycholinguistic experiments. Across eight instruction-tuned transformer models, we run two experimental paradigms -- similarity-based…

机器学习 · 计算机科学 2026-02-17 Louis Schiekiera , Max Zimmer , Christophe Roux , Sebastian Pokutta , Fritz Günther

Latent space models are frequently used for modeling single-layer networks and include many popular special cases, such as the stochastic block model and the random dot product graph. However, they are not well-developed for more complex…

统计方法学 · 统计学 2021-07-09 Peter W. MacDonald , Elizaveta Levina , Ji Zhu

Despite the architectural similarities between tabular in-context learning (ICL) models and large language models (LLMs), little is known about how individual layers contribute to tabular prediction. In this paper, we investigate how the…

机器学习 · 计算机科学 2025-11-20 Amir Rezaei Balef , Mykhailo Koshil , Katharina Eggensperger

Large language model (LLM) architectures are often described as functionally hierarchical: Early layers process syntax, middle layers begin to parse semantics, and late layers integrate information. The present work revisits these ideas.…

计算与语言 · 计算机科学 2025-01-14 Paul C. Bogdan

We investigate the origins of massive activations in large language models (LLMs) and identify a specific layer named the \textbf{Massive Emergence Layer (ME Layer)}, that is consistently observed across model families, where massive…

计算与语言 · 计算机科学 2026-05-14 Zeru Shi , Zhenting Wang , Fan Yang , Qifan Wang , Ruixiang Tang

Social networks profoundly influence how humans form opinions, exchange information, and organize collectively. As large language models (LLMs) are increasingly embedded into social and professional environments, it is critical to…

社会与信息网络 · 计算机科学 2025-10-07 Marios Papachristou , Yuan Yuan

Large language models (LLMs) have revolutionized the field of natural language processing (NLP), and recent studies have aimed to understand their underlying mechanisms. However, most of this research is conducted within a monolingual…

计算与语言 · 计算机科学 2025-09-29 Weixuan Wang , Barry Haddow , Minghao Wu , Wei Peng , Alexandra Birch
‹ 上一页 1 2 3 10 下一页 ›