中文
相关论文

相关论文: HOTVCOM: Generating Buzzworthy Comments for Videos

200 篇论文

Multimodal large language models (MLLMs) are flourishing, but mainly focus on images with less attention than videos, especially in sub-fields such as prompt engineering, video chain-of-thought (CoT), and instruction tuning on videos.…

计算机视觉与模式识别 · 计算机科学 2024-07-09 Yan Wang , Yawen Zeng , Jingsheng Zheng , Xiaofen Xing , Jin Xu , Xiangmin Xu

Video question answering benefits from the rich information in videos, enabling various applications. However, the large volume of tokens generated from long videos presents challenges to memory efficiency and model performance. To…

计算机视觉与模式识别 · 计算机科学 2025-09-16 Yumeng Shi , Quanyu Long , Wenya Wang

As short-form funny videos on social networks are gaining popularity, it becomes demanding for AI models to understand them for better communication with humans. Unfortunately, previous video humor datasets target specific domains, such as…

计算与语言 · 计算机科学 2024-04-02 Dayoon Ko , Sangho Lee , Gunhee Kim

The medical conversational question answering (CQA) system aims at providing a series of professional medical services to improve the efficiency of medical care. Despite the success of large language models (LLMs) in complex reasoning tasks…

计算与语言 · 计算机科学 2023-05-11 Yixuan Weng , Bin Li , Fei Xia , Minjun Zhu , Bin Sun , Shizhu He , Kang Liu , Jun Zhao

Complaining is a speech act that expresses a negative inconsistency between reality and human expectations. While prior studies mostly focus on identifying the existence or the type of complaints, in this work, we present the first study in…

计算与语言 · 计算机科学 2022-04-21 Ming Fang , Shi Zong , Jing Li , Xinyu Dai , Shujian Huang , Jiajun Chen

Learning tasks through videos is a dynamic way to acquire skills by witnessing entire processes. However, compared to in-person demonstrations, videos may omit tacit knowledge, including subtle details and contextual nuances. Users' unique…

人机交互 · 计算机科学 2026-03-16 Nayoung Kim , Yotam Sechayk , Zhongyi Zhou , Takeo Igarashi

Much research in recent years has focused on automatic article commenting. However, few of previous studies focus on the controllable generation of comments. Besides, they tend to generate dull and commonplace comments, which further limits…

计算与语言 · 计算机科学 2021-07-27 Linhao Zhang , Houfeng Wang

Recent methods for visual question answering rely on large-scale annotated datasets. Manual annotation of questions and answers for videos, however, is tedious, expensive and prevents scalability. In this work, we propose to avoid manual…

计算机视觉与模式识别 · 计算机科学 2021-08-13 Antoine Yang , Antoine Miech , Josef Sivic , Ivan Laptev , Cordelia Schmid

Argumentation mining aims at automatically extracting the premises-claim discourse structures in natural language texts. There is a great demand for argumentation corpora for customer reviews. However, due to the controversial nature of the…

计算与语言 · 计算机科学 2017-05-08 Mengxue Li , Shiqiang Geng , Yang Gao , Haijing Liu , Hao Wang

The recent years have witnessed great advances in video generation. However, the development of automatic video metrics is lagging significantly behind. None of the existing metric is able to provide reliable scores over generated videos.…

In recent years, with the rapid development of large language models, serval models such as GPT-4o have demonstrated extraordinary capabilities, surpassing human performance in various language tasks. As a result, many researchers have…

计算与语言 · 计算机科学 2024-09-30 Yi Ren , Tianyi Zhang , Weibin Li , DuoMu Zhou , Chenhao Qin , FangCheng Dong

Existing cyberbullying detection benchmarks were organized by the polarity of speech, such as "offensive" and "non-offensive", which were essentially hate speech detection. However, in the real world, cyberbullying often attracted…

计算与语言 · 计算机科学 2026-05-12 Yi Zhu , Xin Zou , Xindong Wu

Dialogue contradiction is a critical issue in open-domain dialogue systems. The contextualization nature of conversations makes dialogue contradiction detection rather challenging. In this work, we propose a benchmark for Contradiction…

计算与语言 · 计算机科学 2022-10-18 Chujie Zheng , Jinfeng Zhou , Yinhe Zheng , Libiao Peng , Zhen Guo , Wenquan Wu , Zhengyu Niu , Hua Wu , Minlie Huang

Micro-videos are six-second videos popular on social media networks with several unique properties. Firstly, because of the authoring process, they contain significantly more diversity and narrative structure than existing collections of…

计算机视觉与模式识别 · 计算机科学 2016-04-04 Phuc Xuan Nguyen , Gregory Rogez , Charless Fowlkes , Deva Ramanan

Micro-videos have recently gained immense popularity, sparking critical research in micro-video recommendation with significant implications for the entertainment, advertising, and e-commerce industries. However, the lack of large-scale…

信息检索 · 计算机科学 2023-09-28 Yongxin Ni , Yu Cheng , Xiangyan Liu , Junchen Fu , Youhua Li , Xiangnan He , Yongfeng Zhang , Fajie Yuan

Incorporating multi-modal contexts in conversation is important for developing more engaging dialogue systems. In this work, we explore this direction by introducing MMChat: a large-scale Chinese multi-modal dialogue corpus (32.4M raw…

计算与语言 · 计算机科学 2022-05-03 Yinhe Zheng , Guanyi Chen , Xin Liu , Jian Sun

YouTube, a widely popular online platform, has transformed the dynamics of con-tent consumption and interaction for users worldwide. With its extensive range of content crea-tors and viewers, YouTube serves as a hub for video sharing,…

社会与信息网络 · 计算机科学 2023-11-13 Shadi Shajari , Mustafa Alassad , Nitin Agarwal

Recently, multimodal sentiment analysis has seen remarkable advance and a lot of datasets are proposed for its development. In general, current multimodal sentiment analysis datasets usually follow the traditional system of…

计算与语言 · 计算机科学 2021-09-20 Hongxuan Tang , Hao Liu , Xinyan Xiao , Hua Wu

Chinese stand-up comedy generation goes beyond plain text generation, requiring culturally grounded humor, precise timing, stage-performance cues, and implicit multi-step reasoning. Moreover, commonly used Chinese humor datasets are often…

人工智能 · 计算机科学 2026-01-14 Yuyang Wu , Hanzhong Cao , Jianhao Chen , Yufei Li

Automatically identifying harmful content in video is an important task with a wide range of applications. However, there is a lack of professionally labeled open datasets available. In this work VidHarm, an open dataset of 3589 video clips…

计算机视觉与模式识别 · 计算机科学 2022-09-07 Johan Edstedt , Amanda Berg , Michael Felsberg , Johan Karlsson , Francisca Benavente , Anette Novak , Gustav Grund Pihlgren