English
Related papers

Related papers: $\texttt{Droid}$: A Resource Suite for AI-Generate…

200 papers

The widespread adoption of Large Language Models (LLMs) has made the detection of AI-Generated text a pressing and complex challenge. Although many detection systems report high benchmark accuracy, their reliability in real-world settings…

Computation and Language · Computer Science 2026-04-23 Shushanta Pudasaini , Luis Miralles-Pechuán , David Lillis , Marisa Llorens Salvador

Code generation models are not robust to small perturbations, which often lead to incorrect generations and significantly degrade the performance of these models. Although improving the robustness of code generation models is crucial to…

Large Language Models (LLMs) are now capable of generating text that closely resembles human writing, making them powerful tools for content creation, but this growing ability has also made it harder to tell whether a piece of text was…

Computation and Language · Computer Science 2025-10-21 Muhammad Ammar , Hadiya Murad Hadi , Usman Majeed Butt

Detecting AI-involved text is essential for combating misinformation, plagiarism, and academic misconduct. However, AI text generation includes diverse collaborative processes (AI-written text edited by humans, human-written text edited by…

Computation and Language · Computer Science 2025-10-21 Yongxin He , Shan Zhang , Yixuan Cao , Lei Ma , Ping Luo

With the increasing popularity of LLM-based code completers, like GitHub Copilot, the interest in automatically detecting AI-generated code is also increasing-in particular in contexts where the use of LLMs to program is forbidden by policy…

Software Engineering · Computer Science 2024-12-20 Andrea Gurioli , Maurizio Gabbrielli , Stefano Zacchiroli

Large Language Models (LLMs) perform impressively well in various applications. However, the potential for misuse of these models in activities such as plagiarism, generating fake news, and spamming has raised concern about their…

Computation and Language · Computer Science 2025-01-20 Vinu Sankar Sadasivan , Aounon Kumar , Sriram Balasubramanian , Wenxiao Wang , Soheil Feizi

Large-scale content analysis is increasingly limited by the absence of observable ground truth or gold-standard labels, as creating such benchmarks through extensive human coding becomes impractical for massive datasets due to high time,…

Computation and Language · Computer Science 2026-03-09 Luis de-Marcos , Manuel Goyanes , Adrián Domínguez-Díaz

Agent-based coding tools have transformed software development practices. Unlike prompt-based approaches that require developers to manually integrate generated code, these agent-based tools autonomously interact with repositories to…

Software Engineering · Computer Science 2026-03-17 Suzuka Yoshimoto , Shun Fujita , Kosei Horikawa , Daniel Feitosa , Yutaro Kashiwa , Hajimu Iida

Developers often perform repetitive code editing activities for various reasons (e.g., code refactoring) during software development. Pre-trained code editing models have achieved the state-of-the-art (SOTA) results. Pre-trained models are…

Software Engineering · Computer Science 2023-09-08 Jia Li , Ge Li , Zhuo Li , Zhi Jin , Xing Hu , Kechi Zhang , Zhiyi Fu

AI modeling for source code understanding tasks has been making significant progress, and is being adopted in production development pipelines. However, reliability concerns, especially whether the models are actually learning task-related…

Software Engineering · Computer Science 2022-01-11 Sahil Suneja , Yufan Zhuang , Yunhui Zheng , Jim Laredo , Alessandro Morari

We present the shared task on artificial text detection in Russian, which is organized as a part of the Dialogue Evaluation initiative, held in 2022. The shared task dataset includes texts from 14 text generators, i.e., one human writer and…

Current code generation benchmarks focus primarily on functional correctness while overlooking two critical aspects of real-world programming: algorithmic efficiency and code quality. We introduce COMPASS (COdility's Multi-dimensional…

Software Engineering · Computer Science 2025-08-20 James Meaden , Michał Jarosz , Piotr Jodłowski , Grigori Melnik

Automatic code synthesis from natural language descriptions is a challenging task. We witness massive progress in developing code generation systems for domain-specific languages (DSLs) employing sequence-to-sequence deep learning…

Machine Learning · Computer Science 2021-06-23 Mrinal Anand , Pratik Kayal , Mayank Singh

Code synthesis, which requires a deep understanding of complex natural language problem descriptions, generation of code instructions for complex algorithms and data structures, and the successful execution of comprehensive unit tests,…

Computation and Language · Computer Science 2024-05-21 Md. Ashraful Islam , Mohammed Eunus Ali , Md Rizwan Parvez

Code data in large language model (LLM) pretraining is recognized crucial not only for code-related tasks but also for enhancing general intelligence of LLMs. Current open-source LLMs often heavily rely on human effort to produce their code…

Current techniques for detecting AI-generated text are largely confined to manual feature crafting and supervised binary classification paradigms. These methodologies typically lead to performance bottlenecks and unsatisfactory…

Computation and Language · Computer Science 2024-10-29 Xun Guo , Shan Zhang , Yongxin He , Ting Zhang , Wanquan Feng , Haibin Huang , Chongyang Ma

The aim of this study is to evaluate the performance of AI-assisted programming in actual mobile development teams that are focused on native mobile languages like Kotlin and Swift. The extensive case study involves 16 participants and 2…

Software Engineering · Computer Science 2023-09-26 Mircea-Serban Vasiliniuc , Adrian Groza

With the rise of generative language models, machine-generated text detection has become a critical challenge. A wide variety of models is available, but inconsistent datasets, evaluation metrics, and assessment strategies obscure…

Computation and Language · Computer Science 2026-04-23 Kevin Stowe , Kailash Patil

With the rapid growth of large language models for code generation, distinguishing between human-written and AI-generated code has become increasingly critical for academic integrity, hiring evaluations, and software security. We present…

Software Engineering · Computer Science 2026-05-01 Kargi Chauhan , Sadiba Nusrat Nur

The creation of large, diverse, high-quality robot manipulation datasets is an important stepping stone on the path toward more capable and robust robotic manipulation policies. However, creating such datasets is challenging: collecting…

Robotics · Computer Science 2025-04-23 Alexander Khazatsky , Karl Pertsch , Suraj Nair , Ashwin Balakrishna , Sudeep Dasari , Siddharth Karamcheti , Soroush Nasiriany , Mohan Kumar Srirama , Lawrence Yunliang Chen , Kirsty Ellis , Peter David Fagan , Joey Hejna , Masha Itkina , Marion Lepert , Yecheng Jason Ma , Patrick Tree Miller , Jimmy Wu , Suneel Belkhale , Shivin Dass , Huy Ha , Arhan Jain , Abraham Lee , Youngwoon Lee , Marius Memmel , Sungjae Park , Ilija Radosavovic , Kaiyuan Wang , Albert Zhan , Kevin Black , Cheng Chi , Kyle Beltran Hatch , Shan Lin , Jingpei Lu , Jean Mercat , Abdul Rehman , Pannag R Sanketi , Archit Sharma , Cody Simpson , Quan Vuong , Homer Rich Walke , Blake Wulfe , Ted Xiao , Jonathan Heewon Yang , Arefeh Yavary , Tony Z. Zhao , Christopher Agia , Rohan Baijal , Mateo Guaman Castro , Daphne Chen , Qiuyu Chen , Trinity Chung , Jaimyn Drake , Ethan Paul Foster , Jensen Gao , Vitor Guizilini , David Antonio Herrera , Minho Heo , Kyle Hsu , Jiaheng Hu , Muhammad Zubair Irshad , Donovon Jackson , Charlotte Le , Yunshuang Li , Kevin Lin , Roy Lin , Zehan Ma , Abhiram Maddukuri , Suvir Mirchandani , Daniel Morton , Tony Nguyen , Abigail O'Neill , Rosario Scalise , Derick Seale , Victor Son , Stephen Tian , Emi Tran , Andrew E. Wang , Yilin Wu , Annie Xie , Jingyun Yang , Patrick Yin , Yunchu Zhang , Osbert Bastani , Glen Berseth , Jeannette Bohg , Ken Goldberg , Abhinav Gupta , Abhishek Gupta , Dinesh Jayaraman , Joseph J Lim , Jitendra Malik , Roberto Martín-Martín , Subramanian Ramamoorthy , Dorsa Sadigh , Shuran Song , Jiajun Wu , Michael C. Yip , Yuke Zhu , Thomas Kollar , Sergey Levine , Chelsea Finn