English
Related papers

Related papers: Idiom Understanding as a Tool to Measure the Diale…

200 papers

Typologically diverse benchmarks are increasingly created to track the progress achieved in multilingual NLP. Linguistic diversity of these data sets is typically measured as the number of languages or language families included in the…

Computation and Language · Computer Science 2024-04-17 Tanja Samardzic , Ximena Gutierrez , Christian Bentz , Steven Moran , Olga Pelloni

A few benchmarking datasets have been released to evaluate the factual knowledge of pretrained language models. These benchmarks (e.g., LAMA, and ParaRel) are mainly developed in English and later are translated to form new multilingual…

Computation and Language · Computer Science 2023-06-09 Amr Keleg , Walid Magdy

The Bangla linguistic variety is a fascinating mix of regional dialects that contributes to the cultural diversity of the Bangla-speaking community. Despite extensive study into translating Bangla to English, English to Bangla, and Banglish…

Computation and Language · Computer Science 2025-11-18 Fatema Tuj Johora Faria , Mukaffi Bin Moin , Ahmed Al Wase , Mehidi Ahmmed , Md. Rabius Sani , Tashreef Muhammad

Large Language Models (LLMs) exhibit significant performance variations depending on the linguistic and cultural context in which they are applied. This disparity signals the necessity of mature evaluation frameworks that can assess their…

Computation and Language · Computer Science 2025-09-12 Thales Sales Almeida , Giovana Kerche Bonás , João Guilherme Alves Santos

Metaphor and sarcasm are common figurative expressions in people's communication, especially on the Internet or the memes popular among teenagers. We create a new benchmark named NYK-MS (NewYorKer for Metaphor and Sarcasm), which contains…

Computation and Language · Computer Science 2024-09-04 Ke Chang , Hao Li , Junzhao Zhang , Yunfang Wu

Yor\`ub\'a an African language with roughly 47 million speakers encompasses a continuum with several dialects. Recent efforts to develop NLP technologies for African languages have focused on their standard dialects, resulting in…

Computation and Language · Computer Science 2024-07-01 Orevaoghene Ahia , Anuoluwapo Aremu , Diana Abagyan , Hila Gonen , David Ifeoluwa Adelani , Daud Abolade , Noah A. Smith , Yulia Tsvetkov

Multi-lingual competence in large language models is often evaluated via static data benchmarks such as Belebele, M-MMLU and M-GSM. However, these evaluations often fail to provide an adequate understanding of the practical performance and…

Computation and Language · Computer Science 2026-03-13 Victor Ojewale , Inioluwa Deborah Raji , Suresh Venkatasubramanian

Spoken Language Models (SLMs) aim to learn linguistic competence directly from speech using discrete units, widening access to Natural Language Processing (NLP) technologies for languages with limited written resources. However, progress…

Computation and Language · Computer Science 2026-02-23 Adel Moumen , Guangzhi Sun , Philip C. Woodland

The pre-trained multi-lingual XLSR model generalizes well for language identification after fine-tuning on unseen languages. However, the performance significantly degrades when the languages are not very distinct from each other, for…

Machine Learning · Computer Science 2023-02-17 Shangeth Rajaa , Kriti Anandan , Swaraj Dalmia , Tarun Gupta , Eng Siong Chng

Multilingual pre-trained Large Language Models (LLMs) are incredibly effective at Question Answering (QA), a core task in Natural Language Understanding, achieving high accuracies on several multilingual benchmarks. However, little is known…

Computation and Language · Computer Science 2024-04-16 Yahan Yang , Soham Dan , Dan Roth , Insup Lee

We propose a new benchmark evaluating the performance of multimodal large language models on rebus puzzles. The dataset covers 333 original examples of image-based wordplay, cluing 13 categories such as movies, composers, major cities, and…

The goal of this paper is to learn more about how idiomatic information is structurally encoded in embeddings, using a structural probing method. We repurpose an existing English verbal multi-word expression (MWE) dataset to suit the…

Computation and Language · Computer Science 2023-04-28 Filip Klubička , Vasudevan Nedumpozhimana , John D. Kelleher

To date, there exist almost no culturally-specific evaluation benchmarks for large language models (LLMs) that cover a large number of languages and cultures. In this paper, we present Global PIQA, a participatory commonsense reasoning…

Computation and Language · Computer Science 2025-10-29 Tyler A. Chang , Catherine Arnett , Abdelrahman Eldesokey , Abdelrahman Sadallah , Abeer Kashar , Abolade Daud , Abosede Grace Olanihun , Adamu Labaran Mohammed , Adeyemi Praise , Adhikarinayum Meerajita Sharma , Aditi Gupta , Afitab Iyigun , Afonso Simplício , Ahmed Essouaied , Aicha Chorana , Akhil Eppa , Akintunde Oladipo , Akshay Ramesh , Aleksei Dorkin , Alfred Malengo Kondoro , Alham Fikri Aji , Ali Eren Çetintaş , Allan Hanbury , Alou Dembele , Alp Niksarli , Álvaro Arroyo , Amin Bajand , Amol Khanna , Ana Chkhaidze , Ana Condez , Andiswa Mkhonto , Andrew Hoblitzell , Andrew Tran , Angelos Poulis , Anirban Majumder , Anna Vacalopoulou , Annette Kuuipolani Kanahele Wong , Annika Simonsen , Anton Kovalev , Ashvanth. S , Ayodeji Joseph Lana , Barkin Kinay , Bashar Alhafni , Benedict Cibalinda Busole , Bernard Ghanem , Bharti Nathani , Biljana Stojanovska Đurić , Bola Agbonile , Bragi Bergsson , Bruce Torres Fischer , Burak Tutar , Burcu Alakuş Çınar , Cade J. Kanoniakapueo Kane , Can Udomcharoenchaikit , Catherine Arnett , Chadi Helwe , Chaithra Reddy Nerella , Chen Cecilia Liu , Chiamaka Glory Nwokolo , Cristina España-Bonet , Cynthia Amol , DaeYeop Lee , Dana Arad , Daniil Dzenhaliou , Daria Pugacheva , Dasol Choi , Daud Abolade , David Liu , David Semedo , Deborah Popoola , Deividas Mataciunas , Delphine Nyaboke , Dhyuthy Krishna Kumar , Diogo Glória-Silva , Diogo Tavares , Divyanshu Goyal , DongGeon Lee , Ebele Nwamaka Anajemba , Egonu Ngozi Grace , Elena Mickel , Elena Tutubalina , Elias Herranen , Emile Anand , Emmanuel Habumuremyi , Emuobonuvie Maria Ajiboye , Eryawan Presma Yulianrifat , Esther Adenuga , Ewa Rudnicka , Faith Olabisi Itiola , Faran Taimoor Butt , Fathima Thekkekara , Fatima Haouari , Filbert Aurelian Tjiaranata , Firas Laakom , Francesca Grasso , Francesco Orabona , Francesco Periti , Gbenga Kayode Solomon , Gia Nghia Ngo , Gloria Udhehdhe-oze , Gonçalo Martins , Gopi Naga Sai Ram Challagolla , Guijin Son , Gulnaz Abdykadyrova , Hafsteinn Einarsson , Hai Hu , Hamidreza Saffari , Hamza Zaidi , Haopeng Zhang , Harethah Abu Shairah , Harry Vuong , Hele-Andra Kuulmets , Houda Bouamor , Hwanjo Yu , Iben Nyholm Debess , İbrahim Ethem Deveci , Ikhlasul Akmal Hanif , Ikhyun Cho , Inês Calvo , Inês Vieira , Isaac Manzi , Ismail Daud , Itay Itzhak , Iuliia , Alekseenko , Ivan Belashkin , Ivan Spada , Ivan Zhelyazkov , Jacob Brinton , Jafar Isbarov , Jaka Čibej , Jan Čuhel , Jan Kocoń , Jauza Akbar Krito , Jebish Purbey , Jennifer Mickel , Jennifer Za , Jenny Kunz , Jihae Jeong , Jimena Tena Dávalos , Jinu Lee , João Magalhães , John Yi , Jongin Kim , Joseph Chataignon , Joseph Marvin Imperial , Jubeerathan Thevakumar , Judith Land , Junchen Jiang , Jungwhan Kim , Kairit Sirts , Kamesh R , Kamesh V , Kanda Patrick Tshinu , Kätriin Kukk , Kaustubh Ponkshe , Kavsar Huseynova , Ke He , Kelly Buchanan , Kengatharaiyer Sarveswaran , Kerem Zaman , Khalil Mrini , Kian Kyars , Krister Kruusmaa , Kusum Chouhan , Lainitha Krishnakumar , Laura Castro Sánchez , Laura Porrino Moscoso , Leshem Choshen , Levent Sencan , Lilja Øvrelid , Lisa Alazraki , Lovina Ehimen-Ugbede , Luheerathan Thevakumar , Luxshan Thavarasa , Mahnoor Malik , Mamadou K. Keita , Mansi Jangid , Marco De Santis , Marcos García , Marek Suppa , Mariam D'Ciofalo , Marii Ojastu , Maryam Sikander , Mausami Narayan , Maximos Skandalis , Mehak Mehak , Mehmet İlteriş Bozkurt , Melaku Bayu Workie , Menan Velayuthan , Michael Leventhal , Michał Marcińczuk , Mirna Potočnjak , Mohammadamin Shafiei , Mridul Sharma , Mrityunjaya Indoria , Muhammad Ravi Shulthan Habibi , Murat Kolić , Nada Galant , Naphat Permpredanun , Narada Maugin , Nicholas Kluge Corrêa , Nikola Ljubešić , Nirmal Thomas , Nisansa de Silva , Nisheeth Joshi , Nitish Ponkshe , Nizar Habash , Nneoma C. Udeze , Noel Thomas , Noémi Ligeti-Nagy , Nouhoum Coulibaly , Nsengiyumva Faustin , Odunayo Kareemat Buliaminu , Odunayo Ogundepo , Oghojafor Godswill Fejiro , Ogundipe Blessing Funmilola , Okechukwu God'spraise , Olanrewaju Samuel , Olaoye Deborah Oluwaseun , Olasoji Akindejoye , Olga Popova , Olga Snissarenko , Onyinye Anulika Chiemezie , Orkun Kinay , Osman Tursun , Owoeye Tobiloba Moses , Oyelade Oluwafemi Joshua , Oyesanmi Fiyinfoluwa , Pablo Gamallo , Pablo Rodríguez Fernández , Palak Arora , Pedro Valente , Peter Rupnik , Philip Oghenesuowho Ekiugbo , Pramit Sahoo , Prokopis Prokopidis , Pua Niau-Puhipau , Quadri Yahya , Rachele Mignone , Raghav Singhal , Ram Mohan Rao Kadiyala , Raphael Merx , Rapheal Afolayan , Ratnavel Rajalakshmi , Rishav Ghosh , Romina Oji , Ron Kekeha Solis , Rui Guerra , Rushikesh Zawar , Sa'ad Nasir Bashir , Saeed Alzaabi , Sahil Sandeep , Sai Pavan Batchu , SaiSandeep Kantareddy , Salsabila Zahirah Pranida , Sam Buchanan , Samuel Rutunda , Sander Land , Sarah Sulollari , Sardar Ali , Saroj Sapkota , Saulius Tautvaisas , Sayambhu Sen , Sayantani Banerjee , Sebastien Diarra , SenthilNathan. M , Sewoong Lee , Shaan Shah , Shankar Venkitachalam , Sharifa Djurabaeva , Sharon Ibejih , Shivanya Shomir Dutta , Siddhant Gupta , Silvia Paniagua Suárez , Sina Ahmadi , Sivasuthan Sukumar , Siyuan Song , Snegha A. , Sokratis Sofianopoulos , Sona Elza Simon , Sonja Benčina , Sophie Gvasalia , Sphurti Kirit More , Spyros Dragazis , Stephan P. Kaufhold , Suba. S , Sultan AlRashed , Surangika Ranathunga , Taiga Someya , Taja Kuzman Pungeršek , Tal Haklay , Tasi'u Jibril , Tatsuya Aoyama , Tea Abashidze , Terenz Jomar Dela Cruz , Terra Blevins , Themistoklis Nikas , Theresa Dora Idoko , Thu Mai Do , Tilek Chubakov , Tommaso Gargiani , Uma Rathore , Uni Johannesen , Uwuma Doris Ugwu , Vallerie Alexandra Putra , Vanya Bannihatti Kumar , Varsha Jeyarajalingam , Varvara Arzt , Vasudevan Nedumpozhimana , Viktoria Ondrejova , Viktoryia Horbik , Vishnu Vardhan Reddy Kummitha , Vuk Dinić , Walelign Tewabe Sewunetie , Winston Wu , Xiaojing Zhao , Yacouba Diarra , Yaniv Nikankin , Yash Mathur , Yixi Chen , Yiyuan Li , Yolanda Xavier , Yonatan Belinkov , Yusuf Ismail Abayomi , Zaid Alyafeai , Zhengyang Shan , Zhi Rui Tam , Zilu Tang , Zuzana Nadova , Baber Abbasi , Stella Biderman , David Stap , Duygu Ataman , Fabian Schmidt , Hila Gonen , Jiayi Wang , David Ifeoluwa Adelani

Large language models (LLMs) frequently exhibit performance biases against regional dialects of low-resource languages. However, frameworks to quantify these disparities remain scarce. We propose a two-phase framework to evaluate dialectal…

Computation and Language · Computer Science 2026-03-24 K. M. Jubair Sami , Dipto Sumit , Ariyan Hossain , Farig Sadeque

This paper presents a novel Dialectal Sound and Vowelization Recovery framework, designed to recognize borrowed and dialectal sounds within phonologically diverse and dialect-rich languages, that extends beyond its standard orthographic…

Audio and Speech Processing · Electrical Eng. & Systems 2024-08-06 Yassine El Kheir , Hamdy Mubarak , Ahmed Ali , Shammur Absar Chowdhury

While cross-lingual word embeddings have been studied extensively in recent years, the qualitative differences between the different algorithms remain vague. We observe that whether or not an algorithm uses a particular feature set…

Computation and Language · Computer Science 2017-01-11 Omer Levy , Anders Søgaard , Yoav Goldberg

The traditional approach to morphological inflection (the task of modifying a base word (lemma) to express grammatical categories) has been, for decades, to consider lexical entries of lemma-tag-form triples uniformly, lacking any…

Computation and Language · Computer Science 2025-10-28 Tomáš Sourada , Jana Straková

This paper presents insights from evaluating 16 frontier large language models (LLMs) on the WebApp1K benchmark, a test suite designed to assess the ability of LLMs to generate web application code. The results reveal that while all models…

Software Engineering · Computer Science 2024-09-10 Yi Cui

Large language models (LLMs) are increasingly deployed in real-world communication settings, yet their ability to resolve context-dependent ambiguity remains underexplored. In this work, we present EMODIS, a new benchmark for evaluating…

Computation and Language · Computer Science 2025-11-11 Jiacheng Huang , Ning Yu , Xiaoyin Yi

Idiom translation is a challenging problem in machine translation because the meaning of idioms is non-compositional, and a literal (word-by-word) translation is likely to be wrong. In this paper, we focus on evaluating the quality of idiom…

Computation and Language · Computer Science 2018-02-21 Yutong Shao , Rico Sennrich , Bonnie Webber , Federico Fancellu