English
Related papers

Related papers: Voices of Civilizations: A Multilingual QA Benchma…

200 papers

The evaluation of music understanding in Large Audio-Language Models (LALMs) requires a rigorously defined benchmark that truly tests whether models can perceive and interpret music, a standard that current data methodologies frequently…

Computation and Language · Computer Science 2026-03-31 Benno Weck , Pablo Puentes , Andrea Poltronieri , Satyajeet Prabhu , Dmitry Bogdanov

Large audio language models (LALMs) leverage multimodal representations to generate open-ended answers to natural language queries about audio. In this paper, we (1) provide empirical evidence that assessment of LALMs using the popular…

Sound · Computer Science 2026-05-28 Daniel Chenyu Lin , Michael Freeman , John Thickstun

Language Models (LMs) are primarily evaluated on globally popular sports, often overlooking regional and indigenous sporting traditions. To address this gap, we introduce \textbf{\textit{CultSportQA}}, a benchmark designed to assess LMs'…

Existing benchmarks that measure cultural adaptation in LLMs are misaligned with the actual challenges these models face when interacting with users from diverse cultural backgrounds. In this work, we introduce the first framework and…

Computation and Language · Computer Science 2025-10-14 Shreya Havaldar , Sunny Rai , Young-Min Cho , Lyle Ungar

Visual Question Answering (VQA) is an important task in multimodal AI, and it is often used to test the ability of vision-language models to understand and reason on knowledge present in both visual and textual data. However, most of the…

Computer Vision and Pattern Recognition · Computer Science 2024-11-05 David Romero , Chenyang Lyu , Haryo Akbarianto Wibowo , Teresa Lynn , Injy Hamed , Aditya Nanda Kishore , Aishik Mandal , Alina Dragonetti , Artem Abzaliev , Atnafu Lambebo Tonja , Bontu Fufa Balcha , Chenxi Whitehouse , Christian Salamea , Dan John Velasco , David Ifeoluwa Adelani , David Le Meur , Emilio Villa-Cueva , Fajri Koto , Fauzan Farooqui , Frederico Belcavello , Ganzorig Batnasan , Gisela Vallejo , Grainne Caulfield , Guido Ivetta , Haiyue Song , Henok Biadglign Ademtew , Hernán Maina , Holy Lovenia , Israel Abebe Azime , Jan Christian Blaise Cruz , Jay Gala , Jiahui Geng , Jesus-German Ortiz-Barajas , Jinheon Baek , Jocelyn Dunstan , Laura Alonso Alemany , Kumaranage Ravindu Yasas Nagasinghe , Luciana Benotti , Luis Fernando D'Haro , Marcelo Viridiano , Marcos Estecha-Garitagoitia , Maria Camila Buitrago Cabrera , Mario Rodríguez-Cantelar , Mélanie Jouitteau , Mihail Mihaylov , Mohamed Fazli Mohamed Imam , Muhammad Farid Adilazuarda , Munkhjargal Gochoo , Munkh-Erdene Otgonbold , Naome Etori , Olivier Niyomugisha , Paula Mónica Silva , Pranjal Chitale , Raj Dabre , Rendi Chevi , Ruochen Zhang , Ryandito Diandaru , Samuel Cahyawijaya , Santiago Góngora , Soyeong Jeong , Sukannya Purkayastha , Tatsuki Kuribayashi , Teresa Clifford , Thanmay Jayakumar , Tiago Timponi Torrent , Toqeer Ehsan , Vladimir Araujo , Yova Kementchedjhieva , Zara Burzo , Zheng Wei Lim , Zheng Xin Yong , Oana Ignat , Joan Nwatu , Rada Mihalcea , Thamar Solorio , Alham Fikri Aji

Large language models (LLMs) are now deployed worldwide, inspiring a surge of benchmarks that measure their multilingual and multicultural abilities. However, these benchmarks prioritize generic language understanding or superficial…

The rapid progress of large language models (LLMs) raises concerns about cultural bias, fairness, and performance in diverse languages and underrepresented regions. Addressing these gaps requires large-scale resources grounded in…

Computation and Language · Computer Science 2026-04-08 Firoj Alam , Md Arid Hasan , Sahinur Rahman Laskar , Mucahid Kutlu , Kareem Darwish , Shammur Absar Chowdhury

Most multilingual question-answering benchmarks, while covering a diverse pool of languages, do not factor in regional diversity in the information they capture and tend to be Western-centric. This introduces a significant gap in fairly…

Computation and Language · Computer Science 2025-11-04 Eshaan Tanwar , Anwoy Chatterjee , Michael Saxon , Alon Albalak , William Yang Wang , Tanmoy Chakraborty

Vision-language models (VLMs) have advanced human-AI interaction but struggle with cultural understanding, often misinterpreting symbols, gestures, and artifacts due to biases in predominantly Western-centric training data. In this paper,…

Artificial Intelligence · Computer Science 2025-01-03 Shudong Liu , Yiqiao Jin , Cheng Li , Derek F. Wong , Qingsong Wen , Lichao Sun , Haipeng Chen , Xing Xie , Jindong Wang

Internet audio-visual clips convey meaning through time-varying sound and motion, which extend beyond what text alone can represent. To examine whether AI models can understand such signals in human cultural contexts, we introduce AVMeme…

Music understanding is a complex task that often requires reasoning over both structural and semantic elements of audio. We introduce BASS, designed to evaluate music understanding and reasoning in audio language models across four broad…

Sound · Computer Science 2026-02-05 Min Jang , Orevaoghene Ahia , Nazif Tamer , Sachin Kumar , Yulia Tsvetkov , Noah A. Smith

Large language models (LLMs) are now used worldwide, yet their multimodal understanding and reasoning often degrade outside Western, high-resource settings. We propose MMA-ASIA, a comprehensive framework to evaluate LLMs' cultural awareness…

Foundation models and vision-language pre-training have notably advanced Vision Language Models (VLMs), enabling multimodal processing of visual and linguistic data. However, their performance has been typically assessed on general scene…

Computer Vision and Pattern Recognition · Computer Science 2024-10-15 Shravan Nayak , Kanishk Jain , Rabiul Awal , Siva Reddy , Sjoerd van Steenkiste , Lisa Anne Hendricks , Karolina Stańczak , Aishwarya Agrawal

Question-answering (QA) is a natural approach for humans to understand a piece of music audio. However, for machines, accessing a large-scale dataset covering diverse aspects of music is crucial, yet challenging, due to the scarcity of…

Sound · Computer Science 2025-08-28 Zhihao Ouyang , Ju-Chiang Wang , Daiyu Zhang , Bin Chen , Shangjie Li , Quan Lin

To date, there exist almost no culturally-specific evaluation benchmarks for large language models (LLMs) that cover a large number of languages and cultures. In this paper, we present Global PIQA, a participatory commonsense reasoning…

Computation and Language · Computer Science 2025-10-29 Tyler A. Chang , Catherine Arnett , Abdelrahman Eldesokey , Abdelrahman Sadallah , Abeer Kashar , Abolade Daud , Abosede Grace Olanihun , Adamu Labaran Mohammed , Adeyemi Praise , Adhikarinayum Meerajita Sharma , Aditi Gupta , Afitab Iyigun , Afonso Simplício , Ahmed Essouaied , Aicha Chorana , Akhil Eppa , Akintunde Oladipo , Akshay Ramesh , Aleksei Dorkin , Alfred Malengo Kondoro , Alham Fikri Aji , Ali Eren Çetintaş , Allan Hanbury , Alou Dembele , Alp Niksarli , Álvaro Arroyo , Amin Bajand , Amol Khanna , Ana Chkhaidze , Ana Condez , Andiswa Mkhonto , Andrew Hoblitzell , Andrew Tran , Angelos Poulis , Anirban Majumder , Anna Vacalopoulou , Annette Kuuipolani Kanahele Wong , Annika Simonsen , Anton Kovalev , Ashvanth. S , Ayodeji Joseph Lana , Barkin Kinay , Bashar Alhafni , Benedict Cibalinda Busole , Bernard Ghanem , Bharti Nathani , Biljana Stojanovska Đurić , Bola Agbonile , Bragi Bergsson , Bruce Torres Fischer , Burak Tutar , Burcu Alakuş Çınar , Cade J. Kanoniakapueo Kane , Can Udomcharoenchaikit , Catherine Arnett , Chadi Helwe , Chaithra Reddy Nerella , Chen Cecilia Liu , Chiamaka Glory Nwokolo , Cristina España-Bonet , Cynthia Amol , DaeYeop Lee , Dana Arad , Daniil Dzenhaliou , Daria Pugacheva , Dasol Choi , Daud Abolade , David Liu , David Semedo , Deborah Popoola , Deividas Mataciunas , Delphine Nyaboke , Dhyuthy Krishna Kumar , Diogo Glória-Silva , Diogo Tavares , Divyanshu Goyal , DongGeon Lee , Ebele Nwamaka Anajemba , Egonu Ngozi Grace , Elena Mickel , Elena Tutubalina , Elias Herranen , Emile Anand , Emmanuel Habumuremyi , Emuobonuvie Maria Ajiboye , Eryawan Presma Yulianrifat , Esther Adenuga , Ewa Rudnicka , Faith Olabisi Itiola , Faran Taimoor Butt , Fathima Thekkekara , Fatima Haouari , Filbert Aurelian Tjiaranata , Firas Laakom , Francesca Grasso , Francesco Orabona , Francesco Periti , Gbenga Kayode Solomon , Gia Nghia Ngo , Gloria Udhehdhe-oze , Gonçalo Martins , Gopi Naga Sai Ram Challagolla , Guijin Son , Gulnaz Abdykadyrova , Hafsteinn Einarsson , Hai Hu , Hamidreza Saffari , Hamza Zaidi , Haopeng Zhang , Harethah Abu Shairah , Harry Vuong , Hele-Andra Kuulmets , Houda Bouamor , Hwanjo Yu , Iben Nyholm Debess , İbrahim Ethem Deveci , Ikhlasul Akmal Hanif , Ikhyun Cho , Inês Calvo , Inês Vieira , Isaac Manzi , Ismail Daud , Itay Itzhak , Iuliia , Alekseenko , Ivan Belashkin , Ivan Spada , Ivan Zhelyazkov , Jacob Brinton , Jafar Isbarov , Jaka Čibej , Jan Čuhel , Jan Kocoń , Jauza Akbar Krito , Jebish Purbey , Jennifer Mickel , Jennifer Za , Jenny Kunz , Jihae Jeong , Jimena Tena Dávalos , Jinu Lee , João Magalhães , John Yi , Jongin Kim , Joseph Chataignon , Joseph Marvin Imperial , Jubeerathan Thevakumar , Judith Land , Junchen Jiang , Jungwhan Kim , Kairit Sirts , Kamesh R , Kamesh V , Kanda Patrick Tshinu , Kätriin Kukk , Kaustubh Ponkshe , Kavsar Huseynova , Ke He , Kelly Buchanan , Kengatharaiyer Sarveswaran , Kerem Zaman , Khalil Mrini , Kian Kyars , Krister Kruusmaa , Kusum Chouhan , Lainitha Krishnakumar , Laura Castro Sánchez , Laura Porrino Moscoso , Leshem Choshen , Levent Sencan , Lilja Øvrelid , Lisa Alazraki , Lovina Ehimen-Ugbede , Luheerathan Thevakumar , Luxshan Thavarasa , Mahnoor Malik , Mamadou K. Keita , Mansi Jangid , Marco De Santis , Marcos García , Marek Suppa , Mariam D'Ciofalo , Marii Ojastu , Maryam Sikander , Mausami Narayan , Maximos Skandalis , Mehak Mehak , Mehmet İlteriş Bozkurt , Melaku Bayu Workie , Menan Velayuthan , Michael Leventhal , Michał Marcińczuk , Mirna Potočnjak , Mohammadamin Shafiei , Mridul Sharma , Mrityunjaya Indoria , Muhammad Ravi Shulthan Habibi , Murat Kolić , Nada Galant , Naphat Permpredanun , Narada Maugin , Nicholas Kluge Corrêa , Nikola Ljubešić , Nirmal Thomas , Nisansa de Silva , Nisheeth Joshi , Nitish Ponkshe , Nizar Habash , Nneoma C. Udeze , Noel Thomas , Noémi Ligeti-Nagy , Nouhoum Coulibaly , Nsengiyumva Faustin , Odunayo Kareemat Buliaminu , Odunayo Ogundepo , Oghojafor Godswill Fejiro , Ogundipe Blessing Funmilola , Okechukwu God'spraise , Olanrewaju Samuel , Olaoye Deborah Oluwaseun , Olasoji Akindejoye , Olga Popova , Olga Snissarenko , Onyinye Anulika Chiemezie , Orkun Kinay , Osman Tursun , Owoeye Tobiloba Moses , Oyelade Oluwafemi Joshua , Oyesanmi Fiyinfoluwa , Pablo Gamallo , Pablo Rodríguez Fernández , Palak Arora , Pedro Valente , Peter Rupnik , Philip Oghenesuowho Ekiugbo , Pramit Sahoo , Prokopis Prokopidis , Pua Niau-Puhipau , Quadri Yahya , Rachele Mignone , Raghav Singhal , Ram Mohan Rao Kadiyala , Raphael Merx , Rapheal Afolayan , Ratnavel Rajalakshmi , Rishav Ghosh , Romina Oji , Ron Kekeha Solis , Rui Guerra , Rushikesh Zawar , Sa'ad Nasir Bashir , Saeed Alzaabi , Sahil Sandeep , Sai Pavan Batchu , SaiSandeep Kantareddy , Salsabila Zahirah Pranida , Sam Buchanan , Samuel Rutunda , Sander Land , Sarah Sulollari , Sardar Ali , Saroj Sapkota , Saulius Tautvaisas , Sayambhu Sen , Sayantani Banerjee , Sebastien Diarra , SenthilNathan. M , Sewoong Lee , Shaan Shah , Shankar Venkitachalam , Sharifa Djurabaeva , Sharon Ibejih , Shivanya Shomir Dutta , Siddhant Gupta , Silvia Paniagua Suárez , Sina Ahmadi , Sivasuthan Sukumar , Siyuan Song , Snegha A. , Sokratis Sofianopoulos , Sona Elza Simon , Sonja Benčina , Sophie Gvasalia , Sphurti Kirit More , Spyros Dragazis , Stephan P. Kaufhold , Suba. S , Sultan AlRashed , Surangika Ranathunga , Taiga Someya , Taja Kuzman Pungeršek , Tal Haklay , Tasi'u Jibril , Tatsuya Aoyama , Tea Abashidze , Terenz Jomar Dela Cruz , Terra Blevins , Themistoklis Nikas , Theresa Dora Idoko , Thu Mai Do , Tilek Chubakov , Tommaso Gargiani , Uma Rathore , Uni Johannesen , Uwuma Doris Ugwu , Vallerie Alexandra Putra , Vanya Bannihatti Kumar , Varsha Jeyarajalingam , Varvara Arzt , Vasudevan Nedumpozhimana , Viktoria Ondrejova , Viktoryia Horbik , Vishnu Vardhan Reddy Kummitha , Vuk Dinić , Walelign Tewabe Sewunetie , Winston Wu , Xiaojing Zhao , Yacouba Diarra , Yaniv Nikankin , Yash Mathur , Yixi Chen , Yiyuan Li , Yolanda Xavier , Yonatan Belinkov , Yusuf Ismail Abayomi , Zaid Alyafeai , Zhengyang Shan , Zhi Rui Tam , Zilu Tang , Zuzana Nadova , Baber Abbasi , Stella Biderman , David Stap , Duygu Ataman , Fabian Schmidt , Hila Gonen , Jiayi Wang , David Ifeoluwa Adelani

Africa is home to over one-third of the world's languages, yet remains underrepresented in AI research. We introduce Afri-MCQA, the first Multilingual Cultural Question-Answering benchmark covering 7.5k Q&A pairs across 15 African languages…

Natural Question Answering (QA) datasets play a crucial role in evaluating the capabilities of large language models (LLMs), ensuring their effectiveness in real-world applications. Despite the numerous QA datasets that have been developed…

Multimodal models that jointly process audio and language hold great promise in audio understanding and are increasingly being adopted in the music domain. By allowing users to query via text and obtain information about a given audio…

Sound · Computer Science 2024-08-05 Benno Weck , Ilaria Manco , Emmanouil Benetos , Elio Quinton , George Fazekas , Dmitry Bogdanov

To create culturally inclusive vision-language models (VLMs), developing a benchmark that tests their ability to address culturally relevant questions is essential. Existing approaches typically rely on human annotators, making the process…

Computation and Language · Computer Science 2025-06-02 ChaeHun Park , Yujin Baek , Jaeseok Kim , Yu-Jung Heo , Du-Seong Chang , Jaegul Choo
‹ Prev 1 2 3 10 Next ›