中文
相关论文

相关论文: Online Domain-aware LLM Decoding for Continual Dom…

200 篇论文

Detecting deceptive conversations on dynamic platforms is increasingly difficult due to evolving language patterns and Concept Drift (CD)-i.e., semantic or topical shifts that alter the context or intent of interactions over time. These…

计算与语言 · 计算机科学 2026-05-27 Ali Şenol , Garima Agrawal , Huan Liu

Unsupervised Domain Adaptation (UDA) aims at reducing the domain gap between training and testing data and is, in most cases, carried out in offline manner. However, domain changes may occur continuously and unpredictably during deployment…

计算机视觉与模式识别 · 计算机科学 2022-07-22 Theodoros Panagiotakopoulos , Pier Luigi Dovesi , Linus Härenstam-Nielsen , Matteo Poggi

Off-dynamics offline reinforcement learning (RL) aims to learn a policy for a target domain using limited target data and abundant source data collected under different transition dynamics. Existing methods typically address dynamics…

机器学习 · 计算机科学 2026-02-25 Zhangjie Xia , Yu Yang , Pan Xu

Real-world data sets often exhibit temporal dynamics characterized by evolving data distributions. Disregarding this phenomenon, commonly referred to as concept drift, can significantly diminish a model's predictive accuracy. Furthermore,…

机器学习 · 计算机科学 2025-12-16 Mohammad Abu Shaira , Yunhe Feng , Heng Fan , Weishi Shi

Reliable and accurate estimation of the error of an ML model in unseen test domains is an important problem for safe intelligent systems. Prior work uses disagreement discrepancy (DIS^2) to derive practical error bounds under distribution…

机器学习 · 计算机科学 2025-06-19 Aayush Mishra , Anqi Liu

Large Language Models (LLMs) have demonstrated unprecedented prowess across various natural language processing tasks in various application domains. Recent studies show that LLMs can be leveraged to perform lexical semantic tasks, such as…

计算与语言 · 计算机科学 2024-07-30 Huu Tan Mai , Cuong Xuan Chu , Heiko Paulheim

We propose a novel inference-time out-of-domain (OOD) detection algorithm for specialized large language models (LLMs). Despite achieving state-of-the-art performance on in-domain tasks through fine-tuning, specialized LLMs remain…

计算与语言 · 计算机科学 2025-09-17 Ayush Gupta , Ramneet Kaur , Anirban Roy , Adam D. Cobb , Rama Chellappa , Susmit Jha

Log analysis represents a critical sub-domain within AI applications that facilitates automatic approaches to fault and error management of large-scaled software systems, saving labors of traditional manual methods. While existing solutions…

Real-world datasets frequently exhibit evolving data distributions, reflecting temporal variations and underlying shifts. Overlooking this phenomenon, known as concept drift, can substantially degrade the predictive performance of the…

机器学习 · 计算机科学 2025-12-16 Mohammad Abu-Shaira , Weishi Shi

Large language models (LLMs) have showcased their capability with few-shot inference known as in-context learning. However, in-domain demonstrations are not always readily available in real scenarios, leading to cross-domain in-context…

计算与语言 · 计算机科学 2023-11-21 Quanyu Long , Wenya Wang , Sinno Jialin Pan

Knowledge transfer across several streaming processes remain challenging problem not only because of different distributions of each stream but also because of rapidly changing and never-ending environments of data streams. Albeit growing…

机器学习 · 计算机科学 2021-09-14 Renchunzi Xie , Mahardhika Pratama

LLMs operating in dynamic real-world contexts often encounter knowledge that evolves continuously or emerges incrementally. To remain accurate and effective, models must adapt to newly arriving information on the fly. We introduce Online…

计算与语言 · 计算机科学 2026-03-10 Jiyeon Kim , Hyunji Lee , Dylan Zhou , Sue Hyun Park , Seunghyun Yoon , Trung Bui , Franck Dernoncourt , Sungmin Cha , Minjoon Seo

Detecting fake interactions in digital communication platforms remains a challenging and insufficiently addressed problem. These interactions may appear as harmless spam or escalate into sophisticated scam attempts, making it difficult to…

计算与语言 · 计算机科学 2025-05-14 Ali Senol , Garima Agrawal , Huan Liu

Multi-modal Large Language Models (MLLMs) frequently face challenges from concept drift when dealing with real-world streaming data, wherein distributions change unpredictably. This mainly includes gradual drift due to long-tailed data and…

计算机视觉与模式识别 · 计算机科学 2025-02-18 Xiaoyu Yang , Jie Lu , En Yu

Data streams in real-world industrial scenarios often contain transitional operating conditions that are uncovered during offline training, leading to significant distribution shifts. To bridge the gap between static offline models and…

系统与控制 · 电气工程与系统科学 2026-05-26 Hongshuo Zhao , Zeyi Liu , Xiao He

Online Continual Learning (OCL) empowers machine learning models to acquire new knowledge online across a sequence of tasks. However, OCL faces a significant challenge: catastrophic forgetting, wherein the model learned in previous tasks is…

机器学习 · 计算机科学 2024-05-16 Fan Lyu , Daofeng Liu , Linglan Zhao , Zhang Zhang , Fanhua Shang , Fuyuan Hu , Wei Feng , Liang Wang

The increasing use of synthetic data from the public Internet has enhanced data usage efficiency in large language model (LLM) training. However, the potential threat of model collapse remains insufficiently explored. Existing studies…

机器学习 · 计算机科学 2025-07-25 Tianyu Wang , Akira Horiguchi , Lingyou Pang , Carey E. Priebe

Online anomaly detection (OAD) plays a pivotal role in real-time analytics and decision-making for evolving data streams. However, existing methods often rely on costly retraining and rigid decision boundaries, limiting their ability to…

机器学习 · 计算机科学 2026-04-22 Jiaqi Zhu , Shaofeng Cai , Jie Chen , Fang Deng , Beng Chin Ooi , Wenqiao Zhang

Federated Learning (FL) is an emerging domain in the broader context of artificial intelligence research. Methodologies pertaining to FL assume distributed model training, consisting of a collection of clients and a server, with the main…

机器学习 · 计算机科学 2023-05-09 Bhargav Ganguly , Vaneet Aggarwal

Unsupervised domain adaptation leverages abundant labeled data from various source domains to generalize onto unlabeled target data. Prior research has primarily focused on learning domain-invariant features across the source and target…

计算与语言 · 计算机科学 2025-03-10 Jie He , Wendi Zhou , Xiang Lorraine Li , Jeff Z. Pan
‹ 上一页 1 2 3 10 下一页 ›