English
Related papers

Related papers: Code Aesthetics with Agentic Reward Feedback

200 papers

Large Language Models (LLMs) have shown great success in code generation. LLMs take as the input a prompt and output the code. A key question is how to make prompts (i.e., Prompting Techniques). Existing prompting techniques are designed…

Software Engineering · Computer Science 2023-09-08 Jia Li , Yunfei Zhao , Yongmin Li , Ge Li , Zhi Jin

Qualitative analysis of textual contents unpacks rich and valuable information by assigning labels to the data. However, this process is often labor-intensive, particularly when working with large datasets. While recent AI-based tools…

Computation and Language · Computer Science 2023-04-24 Ziang Xiao , Xingdi Yuan , Q. Vera Liao , Rania Abdelghani , Pierre-Yves Oudeyer

Agentic AI systems built on large language models (LLMs) offer significant potential for automating complex workflows, from software development to customer support. However, LLM agents often underperform due to suboptimal configurations;…

Large language models (LLMs) have achieved impressive performance in code generation recently, offering programmers revolutionary assistance in software development. However, due to the auto-regressive nature of LLMs, they are susceptible…

Software Engineering · Computer Science 2025-03-25 Xue Jiang , Yihong Dong , Yongding Tao , Huanyu Liu , Zhi Jin , Wenpin Jiao , Ge Li

This paper provides a comprehensive review of the current methods and metrics used to evaluate the performance of Large Language Models (LLMs) in code generation tasks. With the rapid growth in demand for automated software development,…

Software Engineering · Computer Science 2025-03-05 Liguo Chen , Qi Guo , Hongrui Jia , Zhengran Zeng , Xin Wang , Yijiang Xu , Jian Wu , Yidong Wang , Qing Gao , Jindong Wang , Wei Ye , Shikun Zhang

This paper introduces Code-Vision, a benchmark designed to evaluate the logical understanding and code generation capabilities of Multimodal Large Language Models (MLLMs). It challenges MLLMs to generate a correct program that fulfills…

Computation and Language · Computer Science 2025-02-18 Hanbin Wang , Xiaoxuan Zhou , Zhipeng Xu , Keyuan Cheng , Yuxin Zuo , Kai Tian , Jingwei Song , Junting Lu , Wenhui Hu , Xueyang Liu

Large Language Models (LLMs) have seen great advance in both academia and industry, and their popularity results in numerous open-source frameworks and techniques in accelerating LLM pre-training, fine-tuning, and inference. Training and…

Performance · Computer Science 2023-12-04 Longteng Zhang , Xiang Liu , Zeyu Li , Xinglin Pan , Peijie Dong , Ruibo Fan , Rui Guo , Xin Wang , Qiong Luo , Shaohuai Shi , Xiaowen Chu

Large Language Model(LLM) inference demands massive compute and energy, making domain-specific tasks expensive and unsustainable. As foundation models keep scaling, we ask: Is bigger always better for hardware design? Our work tests this by…

This paper studies how AI-assisted programming and large language models (LLM) improve software developers' ability via AI tools (LLM agents) like Github Copilot and Amazon CodeWhisperer, while integrating human feedback to enhance…

Artificial Intelligence · Computer Science 2025-03-20 Man Fai Wong , Chee Wei Tan

Utilizing Large Language Models (LLMs) for complex tasks is challenging, often involving a time-consuming and uncontrollable prompt engineering process. This paper introduces a novel human-LLM interaction framework, Low-code LLM. It…

Computation and Language · Computer Science 2024-04-02 Yuzhe Cai , Shaoguang Mao , Wenshan Wu , Zehua Wang , Yaobo Liang , Tao Ge , Chenfei Wu , Wang You , Ting Song , Yan Xia , Jonathan Tien , Nan Duan , Furu Wei

Large Language Models have seen increasing use in various software development tasks, especially in code generation. The most advanced recent methods attempt to incorporate feedback from code execution into prompts to help guide LLMs in…

Software Engineering · Computer Science 2025-07-31 Hamed Taherkhani , Melika Sepindband , Hung Viet Pham , Song Wang , Hadi Hemmati

We present GLM-4.5, an open-source Mixture-of-Experts (MoE) large language model with 355B total parameters and 32B activated parameters, featuring a hybrid reasoning method that supports both thinking and direct response modes. Through…

Computation and Language · Computer Science 2025-08-11 5 Team , Aohan Zeng , Xin Lv , Qinkai Zheng , Zhenyu Hou , Bin Chen , Chengxing Xie , Cunxiang Wang , Da Yin , Hao Zeng , Jiajie Zhang , Kedong Wang , Lucen Zhong , Mingdao Liu , Rui Lu , Shulin Cao , Xiaohan Zhang , Xuancheng Huang , Yao Wei , Yean Cheng , Yifan An , Yilin Niu , Yuanhao Wen , Yushi Bai , Zhengxiao Du , Zihan Wang , Zilin Zhu , Bohan Zhang , Bosi Wen , Bowen Wu , Bowen Xu , Can Huang , Casey Zhao , Changpeng Cai , Chao Yu , Chen Li , Chendi Ge , Chenghua Huang , Chenhui Zhang , Chenxi Xu , Chenzheng Zhu , Chuang Li , Congfeng Yin , Daoyan Lin , Dayong Yang , Dazhi Jiang , Ding Ai , Erle Zhu , Fei Wang , Gengzheng Pan , Guo Wang , Hailong Sun , Haitao Li , Haiyang Li , Haiyi Hu , Hanyu Zhang , Hao Peng , Hao Tai , Haoke Zhang , Haoran Wang , Haoyu Yang , He Liu , He Zhao , Hongwei Liu , Hongxi Yan , Huan Liu , Huilong Chen , Ji Li , Jiajing Zhao , Jiamin Ren , Jian Jiao , Jiani Zhao , Jianyang Yan , Jiaqi Wang , Jiayi Gui , Jiayue Zhao , Jie Liu , Jijie Li , Jing Li , Jing Lu , Jingsen Wang , Jingwei Yuan , Jingxuan Li , Jingzhao Du , Jinhua Du , Jinxin Liu , Junkai Zhi , Junli Gao , Ke Wang , Lekang Yang , Liang Xu , Lin Fan , Lindong Wu , Lintao Ding , Lu Wang , Man Zhang , Minghao Li , Minghuan Xu , Mingming Zhao , Mingshu Zhai , Pengfan Du , Qian Dong , Shangde Lei , Shangqing Tu , Shangtong Yang , Shaoyou Lu , Shijie Li , Shuang Li , Shuang-Li , Shuxun Yang , Sibo Yi , Tianshu Yu , Wei Tian , Weihan Wang , Wenbo Yu , Weng Lam Tam , Wenjie Liang , Wentao Liu , Xiao Wang , Xiaohan Jia , Xiaotao Gu , Xiaoying Ling , Xin Wang , Xing Fan , Xingru Pan , Xinyuan Zhang , Xinze Zhang , Xiuqing Fu , Xunkai Zhang , Yabo Xu , Yandong Wu , Yida Lu , Yidong Wang , Yilin Zhou , Yiming Pan , Ying Zhang , Yingli Wang , Yingru Li , Yinpei Su , Yipeng Geng , Yitong Zhu , Yongkun Yang , Yuhang Li , Yuhao Wu , Yujiang Li , Yunan Liu , Yunqing Wang , Yuntao Li , Yuxuan Zhang , Zezhen Liu , Zhen Yang , Zhengda Zhou , Zhongpei Qiao , Zhuoer Feng , Zhuorui Liu , Zichen Zhang , Zihan Wang , Zijun Yao , Zikang Wang , Ziqiang Liu , Ziwei Chai , Zixuan Li , Zuodong Zhao , Wenguang Chen , Jidong Zhai , Bin Xu , Minlie Huang , Hongning Wang , Juanzi Li , Yuxiao Dong , Jie Tang

Large language models (LLMs) have shown remarkable abilities to generate code, however their ability to develop software for embedded systems, which requires cross-domain knowledge of hardware and software has not been studied. In this…

Software Engineering · Computer Science 2023-11-23 Zachary Englhardt , Richard Li , Dilini Nissanka , Zhihan Zhang , Girish Narayanswamy , Joseph Breda , Xin Liu , Shwetak Patel , Vikram Iyer

This review presents a comprehensive analysis of two emerging paradigms in AI-assisted software development: vibe coding and agentic coding. While both leverage large language models (LLMs), they differ fundamentally in autonomy,…

Software Engineering · Computer Science 2025-05-27 Ranjan Sapkota , Konstantinos I. Roumeliotis , Manoj Karkee

Large language models (LLMs) have recently demonstrated a remarkable ability to generate code from natural language (NL) prompts. However, in the real world, NL is often too ambiguous to capture the true intent behind programming problems,…

Machine Learning · Computer Science 2024-03-18 Yeming Wen , Pengcheng Yin , Kensen Shi , Henryk Michalewski , Swarat Chaudhuri , Alex Polozov

Open-source pre-trained Large Language Models (LLMs) exhibit strong language understanding and generation capabilities, making them highly successful in a variety of tasks. However, when used as agents for dealing with complex problems in…

Computation and Language · Computer Science 2024-04-01 Qinhao Zhou , Zihan Zhang , Xiang Xiang , Ke Wang , Yuchuan Wu , Yongbin Li

Generating performant executables from high level languages is critical to software performance across a wide range of domains. Modern compilers perform this task by passing code through a series of well-studied optimizations at…

Programming Languages · Computer Science 2026-04-07 Benjamin Mikek , Danylo Vashchilenko , Bryan Lu , Panpan Xu

Large language models (LLMs) have shown great potential in automating significant aspects of coding by producing natural code from informal natural language (NL) intent. However, when interacting with LLMs, users have no guarantees that the…

Large Language Models (LLMs) are one of the most promising developments in the field of artificial intelligence, and the software engineering community has readily noticed their potential role in the software development life-cycle.…

Software Engineering · Computer Science 2026-03-16 Greta Dolcetti , Vincenzo Arceri , Eleonora Iotti , Sergio Maffeis , Agostino Cortesi , Enea Zaffanella

Evaluating the alignment of large language models (LLMs) with user-defined coding preferences is a challenging endeavour that requires a deep assessment of LLMs' outputs. Existing methods and benchmarks rely primarily on automated metrics…

Software Engineering · Computer Science 2024-12-30 Martin Weyssow , Aton Kamanda , Xin Zhou , Houari Sahraoui