English
Related papers

Related papers: MEANT: Multimodal Encoder for Antecedent Informati…

200 papers

With the increasing prevalence of multimodal content on social media, sentiment analysis faces significant challenges in effectively processing heterogeneous data and recognizing multi-label emotions. Existing methods often lack effective…

Computation and Language · Computer Science 2025-08-26 Xilai Xu , Zilin Zhao , Chengye Song , Zining Wang , Jinhe Qiang , Jiongrui Yan , Yuhuai Lin

In a classification task, dealing with text snippets and metadata usually requires dealing with multimodal approaches. When those metadata are textual, it is tempting to use them intrinsically with a pre-trained transformer, in order to…

Computation and Language · Computer Science 2021-11-09 Barriere Valentin , Jacquet Guillaume

One of the emerging techniques in node classification in heterogeneous graphs is to restrict message aggregation to pre-defined, semantically meaningful structures called metapaths. This work is the first attempt to incorporate attention…

Machine Learning · Computer Science 2024-12-31 Calder Katyal

Temporal prediction is critical for making intelligent and robust decisions in complex dynamic environments. Motion prediction needs to model the inherently uncertain future which often contains multiple potential outcomes, due to…

Machine Learning · Computer Science 2019-12-10 Yichuan Charlie Tang , Ruslan Salakhutdinov

Learning specific hands-on skills such as cooking, car maintenance, and home repairs increasingly happens via instructional videos. The user experience with such videos is known to be improved by meta-information such as time-stamped…

Computer Vision and Pattern Recognition · Computer Science 2020-11-25 Gabriel Huang , Bo Pang , Zhenhai Zhu , Clara Rivera , Radu Soricut

Volume prediction is one of the fundamental objectives in the Fintech area, which is helpful for many downstream tasks, e.g., algorithmic trading. Previous methods mostly learn a universal model for different stocks. However, this kind of…

Trading and Market Microstructure · Quantitative Finance 2022-11-04 Ruibo Chen , Wei Li , Zhiyuan Zhang , Ruihan Bao , Keiko Harimoto , Xu Sun

Most approaches for semantic segmentation use only information from color cameras to parse the scenes, yet recent advancements show that using depth data allows to further improve performances. In this work, we focus on transformer-based…

Computer Vision and Pattern Recognition · Computer Science 2023-03-28 Francesco Barbato , Giulia Rizzoli , Pietro Zanuttigh

Users from the online environment can create different ways of expressing their thoughts, opinions, or conception of amusement. Internet memes were created specifically for these situations. Their main purpose is to transmit ideas by using…

Computation and Language · Computer Science 2020-11-11 George-Alexandru Vlad , George-Eduard Zaharia , Dumitru-Clementin Cercel , Costin-Gabriel Chiru , Stefan Trausan-Matu

Recent research has explored methods for updating and modifying factual knowledge in large language models, often focusing on specific multi-layer perceptron blocks. This study expands on this work by examining the effectiveness of existing…

Computation and Language · Computer Science 2025-02-05 Daniel Tamayo , Aitor Gonzalez-Agirre , Javier Hernando , Marta Villegas

The stock market is characterized by a complex relationship between companies and the market. This study combines a sequential graph structure with attention mechanisms to learn global and local information within temporal time.…

Statistical Finance · Quantitative Finance 2023-01-25 Tzu-Ya Lai , Wen Jung Cheng , Jun-En Ding

Stock trend forecasting, which forecasts stock prices' future trends, plays an essential role in investment. The stocks in a market can share information so that their stock prices are highly correlated. Several methods were recently…

Statistical Finance · Quantitative Finance 2022-01-21 Wentao Xu , Weiqing Liu , Lewen Wang , Yingce Xia , Jiang Bian , Jian Yin , Tie-Yan Liu

In recent years, there has been a significant increase in applications of multimodal signal processing and analysis, largely driven by the increased availability of multimodal datasets and the rapid progress in multimodal learning systems.…

Image and Video Processing · Electrical Eng. & Systems 2024-05-22 Hadi Hadizadeh , S. Faegheh Yeganli , Bahador Rashidi , Ivan V. Bajić

Tabular Foundation Models have recently established the state of the art in supervised tabular learning, by leveraging pretraining to learn generalizable representations of numerical and categorical structured data. However, they lack…

Image captioning model is a cross-modality knowledge discovery task, which targets at automatically describing an image with an informative and coherent sentence. To generate the captions, the previous encoder-decoder frameworks directly…

Computer Vision and Pattern Recognition · Computer Science 2021-02-24 Ziwei Wang , Yadan Luo , Zi Huang

In this dissertation, I present my work towards exploring temporal information for better video understanding. Specifically, I have worked on two problems: action recognition and semantic segmentation. For action recognition, I have…

Computer Vision and Pattern Recognition · Computer Science 2019-05-28 Yi Zhu

Researchers and financial professionals require robust computerized tools that allow users to rapidly operationalize and assess the semantic textual content in financial news. However, existing methods commonly work at the document-level…

Information Retrieval · Computer Science 2019-01-03 Bernhard Lutz , Nicolas Pröllochs , Dirk Neumann

This paper considers the problem of multi-modal future trajectory forecast with ranking. Here, multi-modality and ranking refer to the multiple plausible path predictions and the confidence in those predictions, respectively. We propose…

Computer Vision and Pattern Recognition · Computer Science 2021-03-26 Srikanth Malla , Chiho Choi , Behzad Dariush

With the rapid development of mobile Internet and big data, a huge amount of data is generated in the network, but the data that users are really interested in a very small portion. To extract the information that users are interested in…

Information Retrieval · Computer Science 2022-05-09 Bo Liu

Recent Transformer-based contextual word representations, including BERT and XLNet, have shown state-of-the-art performance in multiple disciplines within NLP. Fine-tuning the trained contextual models on task-specific datasets has been the…

Machine Learning · Computer Science 2020-11-24 Wasifur Rahman , Md. Kamrul Hasan , Sangwu Lee , Amir Zadeh , Chengfeng Mao , Louis-Philippe Morency , Ehsan Hoque

Semantic communication shifts the focus from bit-level accuracy to task-relevant semantic delivery, enabling efficient and intelligent communication for next-generation networks. However, existing multi-modal solutions often process all…

Information Theory · Computer Science 2026-01-01 Yujie Zhou , Cheng Peng , Rulong Wang , Yong Xiao , Yingyu Li , Guangming Shi , Ping Zhang
‹ Prev 1 8 9 10 Next ›