English
Related papers

Related papers: Mapping the Podcast Ecosystem with the Structured …

200 papers

Whether the media faithfully communicate scientific information has long been a core issue to the science community. Automatically identifying paraphrased scientific findings could enable large-scale tracking and analysis of information…

Computation and Language · Computer Science 2022-10-25 Dustin Wright , Jiaxin Pei , David Jurgens , Isabelle Augenstein

The rapid advances in text-to-speech (TTS) technologies have made audio deepfakes increasingly realistic and accessible, raising significant security and trust concerns. While existing research has largely focused on detecting…

Sound · Computer Science 2026-02-03 Alabi Ahmed , Vandana Janeja , Sanjay Purushotham

Systems that can automatically define unfamiliar terms hold the promise of improving the accessibility of scientific texts, especially for readers who may lack prerequisite background knowledge. However, current systems assume a single…

Sound designers search for sounds in large sound effects libraries using aspects such as sound class or visual context. However, the metadata needed for such search is often missing or incomplete, and requires significant manual effort to…

Sound · Computer Science 2026-02-17 Sripathi Sridhar , Prem Seetharaman , Oriol Nieto , Mark Cartwright , Justin Salamon

Recommender systems shape music listening worldwide due to their widespread adoption on online platforms. Growing concerns about representational harms that these systems may cause are increasingly part of the scientific and public debate,…

Human-Computer Interaction · Computer Science 2026-03-12 Lorenzo Porcaro , Chiara Monaldi

This paper presents an overview of a program designed to address the growing need for developing freely available speech resources for under-represented languages. At present we have released 38 datasets for building text-to-speech and…

Mining textual patterns in news, tweets, papers, and many other kinds of text corpora has been an active theme in text mining and NLP research. Previous studies adopt a dependency parsing-based pattern discovery approach. However, the…

Computation and Language · Computer Science 2017-03-16 Meng Jiang , Jingbo Shang , Taylor Cassidy , Xiang Ren , Lance M. Kaplan , Timothy P. Hanratty , Jiawei Han

Audio captioning is an important research area that aims to generate meaningful descriptions for audio clips. Most of the existing research extracts acoustic features of audio clips as input to encoder-decoder and transformer architectures…

Sound · Computer Science 2022-04-20 Ayşegül Özkaya Eren , Mustafa Sert

Idioms are figurative expressions whose meanings often cannot be inferred from their individual words, making them difficult to process computationally and posing challenges for human experimental studies. This survey reviews datasets…

Computation and Language · Computer Science 2025-08-19 Michael Flor , Xinyi Liu , Anna Feldman

We introduce Paralinguistic Speech Captions (ParaSpeechCaps), a large-scale dataset that annotates speech utterances with rich style captions. While rich abstract tags (e.g. guttural, nasal, pained) have been explored in small-scale…

Audio and Speech Processing · Electrical Eng. & Systems 2025-09-26 Anuj Diwan , Zhisheng Zheng , David Harwath , Eunsol Choi

Podcasts are conversational in nature and speaker changes are frequent -- requiring speaker diarization for content understanding. We propose an unsupervised technique for speaker diarization without relying on language-specific components.…

Computation and Language · Computer Science 2022-07-27 M. Iftekhar Tanveer , Diego Casabuena , Jussi Karlgren , Rosie Jones

Data scarcity has been a long standing issue in the field of open-domain social dialogue. To quench this thirst, we present SODA: the first publicly available, million-scale high-quality social dialogue dataset. By contextualizing social…

Computation and Language · Computer Science 2023-10-25 Hyunwoo Kim , Jack Hessel , Liwei Jiang , Peter West , Ximing Lu , Youngjae Yu , Pei Zhou , Ronan Le Bras , Malihe Alikhani , Gunhee Kim , Maarten Sap , Yejin Choi

Text summarization models are approaching human levels of fidelity. Existing benchmarking corpora provide concordant pairs of full and abridged versions of Web, news or, professional content. To date, all summarization datasets operate…

Computation and Language · Computer Science 2022-06-01 Seyed Ali Bahrainian , Sheridan Feucht , Carsten Eickhoff

As audio machine learning outcomes are deployed in societally impactful applications, it is important to have a sense of the quality and origins of the data used. Noticing that being explicit about this sense is not trivially rewarded in…

Sound · Computer Science 2024-10-08 Cynthia C. S. Liem , Doğa Taşcılar , Andrew M. Demetriou

Systems for story generation are asked to produce plausible and enjoyable stories given an input context. This task is underspecified, as a vast number of diverse stories can originate from a single input. The large output space makes it…

Computation and Language · Computer Science 2020-10-06 Nader Akoury , Shufan Wang , Josh Whiting , Stephen Hood , Nanyun Peng , Mohit Iyyer

The applications of conversational agents for scientific disciplines (as expert domains) are understudied due to the lack of dialogue data to train such agents. While most data collection frameworks, such as Amazon Mechanical Turk, foster…

Computation and Language · Computer Science 2022-10-14 Federico Ruggeri , Mohsen Mesgar , Iryna Gurevych

Recommender systems (RS) commonly retrieve potential candidate items for users from a massive number of items by modeling user interests based on historical interactions. However, historical interaction data is highly sparse, and most items…

Information Retrieval · Computer Science 2023-01-18 Ziwei Fan , Alice Wang , Zahra Nazari

The rapid expansion of scientific literature in computer science presents challenges in tracking research trends and extracting key insights. Existing datasets provide metadata but lack structured summaries that capture core contributions…

Information Retrieval · Computer Science 2025-03-03 Javin Liu , Aryan Vats , Zihao He

Speech data is notoriously difficult to work with due to a variety of codecs, lengths of recordings, and meta-data formats. We present Lhotse, a speech data representation library that draws upon lessons learned from Kaldi speech…

Sound · Computer Science 2021-10-26 Piotr Żelasko , Daniel Povey , Jan "Yenda" Trmal , Sanjeev Khudanpur

Recent advances in text-to-speech (TTS) have been driven by large, multi-domain speech corpora, yet the expressive potential of audiobook data remains underexamined. We argue that human-narrated audiobooks, particularly fictional works,…

Audio and Speech Processing · Electrical Eng. & Systems 2026-04-22 Gaspard Michel , Elena V. Epure , Christophe Cerisara
‹ Prev 1 8 9 10 Next ›