中文
相关论文

相关论文: IMDB Spoiler Dataset

200 篇论文

Recommender systems play a crucial role in modern life, including information retrieval, the pharmaceutical industry, retail, and entertainment. The entertainment sector, in particular, attracts significant attention and generates…

机器学习 · 计算机科学 2025-02-11 Amirhossein Dadashzadeh Taromi , Sina Heydari , Mohsen Hooshmand , Majid Ramezani

Like other social media websites, YouTube is not immune from the attention of spammers. In particular, evidence can be found of attempts to attract users to malicious third-party websites. As this type of spam is often associated with…

社会与信息网络 · 计算机科学 2012-02-24 Derek O'Callaghan , Martin Harrigan , Joe Carthy , Pádraig Cunningham

With the large chunks of social media data being created daily and the parallel rise of realistic multimedia tampering methods, detecting and localising tampering in images and videos has become essential. This survey focusses on approaches…

计算机视觉与模式识别 · 计算机科学 2024-01-17 Ankit Yadav , Dinesh Kumar Vishwakarma

This paper proposes a movie genre-prediction based on multinomial probability model. To the best of our knowledge, this problem has not been addressed yet in the field of recommender system. The prediction of a movie genre has many…

信息检索 · 计算机科学 2016-03-28 Eric Makita , Artem Lenskiy

Recent years have seen remarkable advances in visual understanding. However, how to understand a story-based long video with artistic styles, e.g. movie, remains challenging. In this paper, we introduce MovieNet -- a holistic dataset for…

计算机视觉与模式识别 · 计算机科学 2020-07-22 Qingqiu Huang , Yu Xiong , Anyi Rao , Jiaze Wang , Dahua Lin

Information-seeking dialogues span a wide range of questions, from simple factoid to complex queries that require exploring multiple facets and viewpoints. When performing exploratory searches in unfamiliar domains, users may lack…

信息检索 · 计算机科学 2024-10-30 Weronika Łajewska , Krisztian Balog , Damiano Spina , Johanne Trippas

Over the years, explosive growth in the number of items in the catalog of e-commerce businesses, such as Amazon, Netflix, Pandora, etc., have warranted the development of recommender systems to guide consumers towards their desired products…

信息检索 · 计算机科学 2019-09-30 Mojdeh Saadati , Syed Shihab , Mohammed Shaiqur Rahman

Automatic unreliable news detection is a research problem with great potential impact. Recently, several papers have shown promising results on large-scale news datasets with models that only use the article itself without resorting to any…

计算与语言 · 计算机科学 2021-04-21 Xiang Zhou , Heba Elfardy , Christos Christodoulopoulos , Thomas Butler , Mohit Bansal

Recommender systems often struggle to strike a balance between matching users' tastes and providing unexpected recommendations. When recommendations are too narrow and fail to cover the full range of users' preferences, the system is…

人机交互 · 计算机科学 2023-10-10 Ruixuan Sun , Avinash Akella , Ruoyan Kong , Moyan Zhou , Joseph A. Konstan

The lack of large realistic datasets presents a bottleneck in online deception detection studies. In this paper, we apply a data collection method based on social network analysis to quickly identify high-quality deceptive and truthful…

计算与语言 · 计算机科学 2017-08-01 Wenlin Yao , Zeyu Dai , Ruihong Huang , James Caverlee

Online reviews play a crucial role in deciding the quality before purchasing any product. Unfortunately, spammers often take advantage of online review forums by writing fraud reviews to promote/demote certain products. It may turn out to…

社会与信息网络 · 计算机科学 2019-07-30 Sarthika Dhawan , Siva Charan Reddy Gangireddy , Shiv Kumar , Tanmoy Chakraborty

Movie ratings play an important role both in determining the likelihood of a potential viewer to watch the movie and in reflecting the current viewer satisfaction with the movie. They are available in several sources like the television…

信息检索 · 计算机科学 2016-05-02 Eric Makita , Artem Lenskiy

The increasing reliance on digital information necessitates advancements in conversational search systems, particularly in terms of information transparency. While prior research in conversational information-seeking has concentrated on…

信息检索 · 计算机科学 2024-05-07 Weronika Łajewska , Damiano Spina , Johanne Trippas , Krisztian Balog

Idioms are figurative expressions whose meanings often cannot be inferred from their individual words, making them difficult to process computationally and posing challenges for human experimental studies. This survey reviews datasets…

计算与语言 · 计算机科学 2025-08-19 Michael Flor , Xinyi Liu , Anna Feldman

Research publication requires public datasets. In recommender systems, some datasets are largely used to compare algorithms against a --supposedly-- common benchmark. Problem: for various reasons, these datasets are heavily preprocessed,…

信息检索 · 计算机科学 2019-09-30 Anne-Marie Tousch

We explore story generation: creative systems that can build coherent and fluent passages of text about a topic. We collect a large dataset of 300K human-written stories paired with writing prompts from an online forum. Our dataset enables…

计算与语言 · 计算机科学 2018-05-15 Angela Fan , Mike Lewis , Yann Dauphin

To address the problem of narrow recommendation ranges caused by an emphasis on prediction accuracy, serendipitous recommendations, which consider both usefulness and unexpectedness, have attracted attention. However, realizing…

信息检索 · 计算机科学 2025-04-10 Zhelin Xu , Atsushi Matsumura

With recent advances in computer vision and graphics, it is now possible to generate videos with extremely realistic synthetic faces, even in real time. Countless applications are possible, some of which raise a legitimate alarm, calling…

计算机视觉与模式识别 · 计算机科学 2018-03-28 Andreas Rössler , Davide Cozzolino , Luisa Verdoliva , Christian Riess , Justus Thies , Matthias Nießner

Decision-making is a cognitively intensive task that requires synthesizing relevant information from multiple unstructured sources, weighing competing factors, and incorporating subjective user preferences. Existing methods, including large…

计算与语言 · 计算机科学 2026-04-21 Akriti Jain , Anish Mulay , Divyansh Verma , Aishani Pandey , Pritika Ramu , Aparna Garimella

The curation of hate speech datasets involves complex design decisions that balance competing priorities. This paper critically examines these methodological choices in a diverse range of datasets, highlighting common themes and practices,…

计算与语言 · 计算机科学 2025-06-23 Luna Wang , Andrew Caines , Alice Hutchings