English
Related papers

Related papers: Richer Countries and Richer Representations

200 papers

Rule-based machine translation is more data efficient than the big data-based machine translation approaches, making it appropriate for languages with low bilingual corpus resources -- i.e., minority languages. However, the rule-based…

Computation and Language · Computer Science 2019-04-29 Patrick Connor

Your name tells a lot about you: your gender, ethnicity and so on. It has been shown that name embeddings are more effective in representing names than traditional substring features. However, our previous name embedding model is trained on…

Social and Information Networks · Computer Science 2019-05-14 Junting Ye , Steven Skiena

In meetings where important decisions get made, what items receive more attention may influence the outcome. We examine how different types of rhetorical (de-)emphasis -- including hedges, superlatives, and contrastive conjunctions --…

Social and Information Networks · Computer Science 2016-12-21 Chenhao Tan , Lillian Lee

We ask whether demographic identity, signaled by a name alone, systematically reshapes the generative distribution of a language model. Measuring full-vocabulary Shannon entropy at temperature zero across six open-weight base models and…

Computation and Language · Computer Science 2026-05-08 Messi H. J. Lee

Word embeddings and pre-trained language models allow to build rich representations of text and have enabled improvements across most NLP tasks. Unfortunately they are very expensive to train, and many small companies and research groups…

Computation and Language · Computer Science 2020-04-03 Rodrigo Agerri , Iñaki San Vicente , Jon Ander Campos , Ander Barrena , Xabier Saralegi , Aitor Soroa , Eneko Agirre

Word embeddings are a fixed, distributional representation of the context of words in a corpus learned from word co-occurrences. Despite their proven utility in machine learning tasks, word embedding models may capture uneven semantic and…

Computation and Language · Computer Science 2021-10-07 James Powell , Kari Sentz , Martin Klein

Dense vector representations for sentences made significant progress in recent years as can be seen on sentence similarity tasks. Real-world phrase retrieval applications, on the other hand, still encounter challenges for effective use of…

Computation and Language · Computer Science 2024-05-14 Eyal Orbach , Lev Haikin , Nelly David , Avi Faizakof

Mounting evidences are being gathered suggesting that income and wealth distribution in various countries or societies follow a robust pattern, close to the Gibbs distribution of energy in an ideal gas in equilibrium, but also deviating…

Physics and Society · Physics 2008-12-02 Arnab Chatterjee , Sitabhra Sinha , Bikas K. Chakrabarti

Word embeddings are commonly obtained as optimizers of a criterion function $f$ of a text corpus, but assessed on word-task performance using a different evaluation function $g$ of the test data. We contend that a possible source of…

Machine Learning · Statistics 2019-11-11 Rachel Carrington , Karthik Bharath , Simon Preston

Recent advancements in Large Language Models (LLMs) have made them a popular information-seeking tool among end users. However, the statistical training methods for LLMs have raised concerns about their representation of under-represented…

Computation and Language · Computer Science 2025-04-09 Shiran Dudy , Thulasi Tholeti , Resmi Ramachandranpillai , Muhammad Ali , Toby Jia-Jun Li , Ricardo Baeza-Yates

Approaches to signal representation and coding theory have traditionally focused on how to best represent signals using parsimonious representations that incur the lowest possible distortion. Classical examples include linear and non-linear…

Information Theory · Computer Science 2015-12-25 Petros T Boufounos , Shantanu Rane , Hassan Mansour

Human-annotated datasets with explicit difficulty ratings are essential in intelligent educational systems. Although embedding vector spaces are widely used to represent semantic closeness and are promising for analyzing text difficulty,…

Artificial Intelligence · Computer Science 2025-12-05 Yo Ehara

Reinforcement Learning (RL) has achieved tremendous development in recent years, but still faces significant obstacles in addressing complex real-life problems due to the issues of poor system generalization, low sample efficiency as well…

Artificial Intelligence · Computer Science 2025-02-25 Chao Yu , Shicheng Ye , Hankz Hankui Zhuo

Economies grow by upgrading the type of products they produce and export. The technology, capital, institutions and skills needed to make such new products are more easily adapted from some products than others. We study the network of…

General Finance · Quantitative Finance 2009-11-13 C. A. Hidalgo , B. Klinger , A. -L. Barabasi , R. Hausmann

Various papers demonstrate the importance of inequality, poverty and the size of the middle class for economic growth. When explaining why these measures of the income distribution are added to the growth regression, it is often mentioned…

Econometrics · Economics 2019-03-07 Max Köhler , Stefan Sperlich , Jisu Yoon

The low-level sensory and motor signals in deep reinforcement learning, which exist in high-dimensional spaces such as image observations or motor torques, are inherently challenging to understand or utilize directly for downstream tasks.…

Artificial Intelligence · Computer Science 2023-03-07 Pu Hua , Yubei Chen , Huazhe Xu

CNNs exhibit many behaviors different from humans, one of which is the capability of employing high-frequency components. This paper discusses the frequency bias phenomenon in image classification tasks: the high-frequency components are…

Computer Vision and Pattern Recognition · Computer Science 2022-08-17 Zhiyu Lin , Yifei Gao , Jitao Sang

In reinforcement learning (RL), state representations are key to dealing with large or continuous state spaces. While one of the promises of deep learning algorithms is to automatically construct features well-tuned for the task they try to…

Machine Learning · Computer Science 2023-06-21 Charline Le Lan , Stephen Tu , Mark Rowland , Anna Harutyunyan , Rishabh Agarwal , Marc G. Bellemare , Will Dabney

Direct democracy is a special case of an ensemble of classifiers, where every person (classifier) votes on every issue. This fails when the average voter competence (classifier accuracy) falls below 50%, which can happen in noisy settings…

Computer Science and Game Theory · Computer Science 2018-07-23 Malik Magdon-Ismail , Lirong Xia

Replacing static word embeddings with contextualized word representations has yielded significant improvements on many NLP tasks. However, just how contextual are the contextualized representations produced by models such as ELMo and BERT?…

Computation and Language · Computer Science 2019-09-04 Kawin Ethayarajh
‹ Prev 1 8 9 10 Next ›