中文
相关论文

相关论文: LEGO: LoRA-Enabled Generator-Oriented Framework fo…

200 篇论文

The field of artificial intelligence is built on object detection techniques. YOU ONLY LOOK ONCE (YOLO) algorithm and it's more evolved versions are briefly described in this research survey. This survey is all about YOLO and convolution…

计算机视觉与模式识别 · 计算机科学 2022-08-02 Viswanatha V , Chandana R K , Ramachandra A. C.

RL (reinforcement learning) methods (e.g., GRPO) for MLLM (Multimodal LLM) perception ability has attracted wide research interest owing to its remarkable generalization ability. Nevertheless, existing reinforcement learning methods still…

计算机视觉与模式识别 · 计算机科学 2025-11-25 Qihan Huang , Haofei Zhang , Rong Wei , Yi Wang , Rui Tang , Mingli Song , Jie Song

Low-Rank Adaptation (LoRA) has emerged as a leading technique for efficiently fine-tuning text-to-image diffusion models, and its widespread adoption on open-source platforms has fostered a vibrant culture of model sharing and…

计算机视觉与模式识别 · 计算机科学 2026-04-27 Liangwei Lyu , Jiaqi Xu , Jianwei Ding , Qiyao Deng

Deep reinforcement learning (DRL) methods have recently shown promise in path planning tasks. However, when dealing with global planning tasks, these methods face serious challenges such as poor convergence and generalization. To this end,…

机器学习 · 计算机科学 2024-01-10 Guoming Huang , Mingxin Hou , Xiaofang Yuan , Shuqiao Huang , Yaonan Wang

Recent advances in large-scale text-to-image diffusion models have heightened concerns about their potential misuse, especially in generating harmful or misleading content. This underscores the urgent need for effective machine unlearning,…

计算机视觉与模式识别 · 计算机科学 2025-08-11 Agnieszka Polowczyk , Alicja Polowczyk , Dawid Malarz , Artur Kasymov , Marcin Mazur , Jacek Tabor , Przemysław Spurek

At this moment, GAN-based image generation methods are still imperfect, whose upsampling design has limitations in leaving some certain artifact patterns in the synthesized image. Such artifact patterns can be easily exploited (by recent…

计算机视觉与模式识别 · 计算机科学 2020-08-18 Yihao Huang , Felix Juefei-Xu , Run Wang , Qing Guo , Lei Ma , Xiaofei Xie , Jianwen Li , Weikai Miao , Yang Liu , Geguang Pu

The rapid evolution of deep generative models poses a critical challenge to deepfake detection, as detectors trained on forgery-specific artifacts often suffer significant performance degradation when encountering unseen forgeries. While…

计算机视觉与模式识别 · 计算机科学 2025-04-25 Mengyu Qiao , Runze Tian , Yang Wang

Evaluating large language models (LLMs) on comprehensive benchmarks is a cornerstone of their development, yet it's often computationally and financially prohibitive. While Item Response Theory (IRT) offers a promising path toward…

人工智能 · 计算机科学 2025-10-07 Lele Liao , Qile Zhang , Ruofan Wu , Guanhua Fang

Low-rank adaptation (LoRA) is a popular method for fine-tuning large-scale pre-trained models in downstream tasks by learning low-rank incremental matrices. Though LoRA and its variants effectively reduce the number of trainable parameters…

机器学习 · 计算机科学 2024-03-21 Rushi Qiang , Ruiyi Zhang , Pengtao Xie

The rapid advancement of generative models has made the detection of AI-generated images a critical challenge for both research and society. Recent works have shown that most state-of-the-art fake image detection methods overfit to their…

计算机视觉与模式识别 · 计算机科学 2026-02-25 Aayush Dhakal , Subash Khanal , Srikumar Sastry , Jacob Arndt , Philipe Ambrozio Dias , Dalton Lunga , Nathan Jacobs

LoRA has become a universal Parameter-Efficient Fine-Tuning (PEFT) technique that equips Large Language Models (LLMs) to adapt quickly to new tasks. However, when these models are scaled up, even the latest LoRA variants still introduce…

计算与语言 · 计算机科学 2026-02-25 Xindian Ma , Rundong Kong , Peng Zhang , Ruoxiang Huang , Yongyu Jiang

Structural understanding of complex visual objects is an important unsolved component of artificial intelligence. To study this, we develop a new technique for the recently proposed Break-and-Make problem in LTRON where an agent must learn…

人工智能 · 计算机科学 2024-10-03 Aaron Walsman , Muru Zhang , Adam Fishman , Ali Farhadi , Dieter Fox

Advances in Generative AI have made video-level deepfake detection increasingly challenging, exposing the limitations of current detection techniques. In this paper, we present HOLA, our solution to the Video-Level Deepfake Detection track…

计算机视觉与模式识别 · 计算机科学 2025-07-31 Xuecheng Wu , Danlei Huang , Heli Sun , Xinyi Yin , Yifan Wang , Hao Wang , Jia Zhang , Fei Wang , Peihao Guo , Suyu Xing , Junxiao Xue , Liang He

Large Language Models (LLMs) excel at reasoning and generation but are inherently limited by static pretraining data, resulting in factual inaccuracies and weak adaptability to new information. Retrieval-Augmented Generation (RAG) addresses…

计算与语言 · 计算机科学 2025-11-03 Qi Luo , Xiaonan Li , Yuxin Wang , Tingshuo Fan , Yuan Li , Xinchi Chen , Xipeng Qiu

Generative models excel at mimicking real scenes, suggesting they might inherently encode important intrinsic scene properties. In this paper, we aim to explore the following key questions: (1) What intrinsic knowledge do generative models…

计算机视觉与模式识别 · 计算机科学 2024-10-17 Xiaodan Du , Nicholas Kolkin , Greg Shakhnarovich , Anand Bhattad

Generative adversarial networks (GANs) typically require ample data for training in order to synthesize high-fidelity images. Recent studies have shown that training GANs with limited data remains formidable due to discriminator…

计算机视觉与模式识别 · 计算机科学 2021-11-15 Liming Jiang , Bo Dai , Wayne Wu , Chen Change Loy

Beyond general recognition tasks, specialized domains and fine-grained settings often encounter data scarcity, especially for tail classes. To obtain less biased and more reliable models under such scarcity, practitioners leverage diffusion…

计算机视觉与模式识别 · 计算机科学 2026-03-11 Hoyoung Kim , Minwoo Jang , Jabin Koo , Sangdoo Yun , Jungseul Ok

The proliferation of fine-tuned language model experts for specific tasks and domains signals the need for efficient selection and combination methods. We propose LoRA-Augmented Generation (LAG) for leveraging large libraries of knowledge…

计算与语言 · 计算机科学 2025-08-19 William Fleshman , Benjamin Van Durme

Image Retrieval is a fundamental task of obtaining images similar to the query one from a database. A common image retrieval practice is to firstly retrieve candidate images via similarity search using global image features and then re-rank…

计算机视觉与模式识别 · 计算机科学 2021-08-12 Min Yang , Dongliang He , Miao Fan , Baorong Shi , Xuetong Xue , Fu Li , Errui Ding , Jizhou Huang

Finetuned LLMs often exhibit poor uncertainty quantification, manifesting as overconfidence, poor calibration, and unreliable prediction results on test data or out-of-distribution samples. One approach commonly used in vision for…

机器学习 · 计算机科学 2023-10-06 Xi Wang , Laurence Aitchison , Maja Rudolph