English
Related papers

Related papers: Deep Self-Evolving Reasoning

200 papers

Large Reasoning Models (LRMs) have the ability to self-correct even when they make mistakes in their reasoning paths. However, our study reveals that when the reasoning process starts with a short but poor beginning, it becomes difficult…

Computation and Language · Computer Science 2025-05-13 Tongxu Luo , Wenyu Du , Jiaxi Bi , Stephen Chung , Zhengyang Tang , Hao Yang , Min Zhang , Benyou Wang

Recent advances in large language models (LLMs) have enabled deep research systems that synthesize comprehensive, report-style answers to open-ended queries by combining retrieval, reasoning, and generation. Yet most frameworks rely on…

Computation and Language · Computer Science 2026-05-26 Lin Ai , Victor S. Bursztyn , Xiang Chen , Julia Hirschberg , Saayan Mitra

Large language models (LLMs) have notably progressed in multi-step and long-chain reasoning. However, extending their reasoning capabilities to encompass deep interactions with search remains a non-trivial challenge, as models often fail to…

Computation and Language · Computer Science 2025-06-05 Qingfei Zhao , Ruobing Wang , Dingling Xu , Daren Zha , Limin Liu

We introduce DeepSeek-Prover-V2, an open-source large language model designed for formal theorem proving in Lean 4, with initialization data collected through a recursive theorem proving pipeline powered by DeepSeek-V3. The cold-start…

We introduce Seed1.5-Thinking, capable of reasoning through thinking before responding, resulting in improved performance on a wide range of benchmarks. Seed1.5-Thinking achieves 86.7 on AIME 2024, 55.0 on Codeforces and 77.3 on GPQA,…

Computation and Language · Computer Science 2025-04-30 ByteDance Seed , : , Jiaze Chen , Tiantian Fan , Xin Liu , Lingjun Liu , Zhiqi Lin , Mingxuan Wang , Chengyi Wang , Xiangpeng Wei , Wenyuan Xu , Yufeng Yuan , Yu Yue , Lin Yan , Qiying Yu , Xiaochen Zuo , Chi Zhang , Ruofei Zhu , Zhecheng An , Zhihao Bai , Yu Bao , Xingyan Bin , Jiangjie Chen , Feng Chen , Hongmin Chen , Riwei Chen , Liangqiang Chen , Zixin Chen , Jinsong Chen , Siyan Chen , Kaiyuan Chen , Zhi Chen , Jin Chen , Jiecao Chen , Jinxin Chi , Weinan Dai , Ning Dai , Jiahui Dai , Shihan Dou , Yantao Du , Zhengyin Du , Jianhui Duan , Chen Dun , Ting-Han Fan , Jiazhan Feng , Junda Feng , Ziyuan Feng , Yuwei Fu , Wenqi Fu , Hanjie Fu , Hao Ge , Hongyi Guo , Mingji Han , Li Han , Wenhao Hao , Xintong Hao , Qianyu He , Jerry He , Feng He , Wen Heng , Zehua Hong , Qi Hou , Liang Hu , Shengding Hu , Nan Hu , Kai Hua , Qi Huang , Ziyue Huang , Hongzhi Huang , Zihao Huang , Ting Huang , Wenhao Huang , Wei Jia , Bin Jia , Xiaoying Jia , Yuhua Jiang , Haobin Jiang , Ziheng Jiang , Kaihua Jiang , Chengquan Jiang , Jianpeng Jiao , Xiaoran Jin , Xing Jin , Xunhao Lai , Zheng Li , Xiang Li , Liyi Li , Hongkai Li , Zheng Li , Shengxian Wan , Ya Wang , Yunshui Li , Chenggang Li , Niuniu Li , Siyu Li , Xi Li , Xiao Li , Aoyan Li , Yuntao Li , Nianning Liang , Xinnian Liang , Haibin Lin , Weijian Lin , Ye Lin , Zhicheng Liu , Guanlin Liu , Guanlin Liu , Chenxiao Liu , Yan Liu , Gaohong Liu , Juncai Liu , Chundian Liu , Deyi Liu , Kaibo Liu , Siyao Liu , Qi Liu , Yongfei Liu , Kang Liu , Gan Liu , Boyi Liu , Rui Long , Weiqiang Lou , Chenwei Lou , Xiang Luo , Yao Luo , Caiping Lv , Heyang Lv , Bole Ma , Qianli Ma , Hongzhi Ma , Yiyuan Ma , Jin Ma , Wenchang Ma , Tingting Ma , Chen Mao , Qiyang Min , Zhe Nan , Guanghan Ning , Jinxiang Ou , Haojie Pan , Renming Pang , Yanghua Peng , Tao Peng , Lihua Qian , Lihua Qian , Mu Qiao , Meng Qu , Cheng Ren , Hongbin Ren , Yong Shan , Wei Shen , Ke Shen , Kai Shen , Guangming Sheng , Jinlong Shi , Wenlei Shi , Guang Shi , Shuai Shuai Cao , Yuxin Song , Zuquan Song , Jing Su , Yifan Sun , Tao Sun , Zewei Sun , Borui Wan , Zihan Wang , Xiaohui Wang , Xi Wang , Shuguang Wang , Jun Wang , Qinlong Wang , Chenyuan Wang , Shuai Wang , Zihan Wang , Changbao Wang , Jiaqiang Wang , Shihang Wang , Xuwu Wang , Zaiyuan Wang , Yuxuan Wang , Wenqi Wang , Taiqing Wang , Chengzhi Wei , Houmin Wei , Ziyun Wei , Shufa Wei , Zheng Wu , Yonghui Wu , Yangjun Wu , Bohong Wu , Shuang Wu , Jingqiao Wu , Ning Wu , Shuangzhi Wu , Jianmin Wu , Chenguang Xi , Fan Xia , Yuqiao Xian , Liang Xiang , Boren Xiang , Bowen Xiao , Zhen Xiao , Xia Xiao , Yongsheng Xiao , Chao Xin , Shulin Xin , Yuwen Xiong , Jingjing Xu , Ziwen Xu , Chenyin Xu , Jiayi Xu , Yifan Xu , Wei Xu , Yufei Xu , Shikun Xu , Shipeng Yan , Shen Yan , Qingping Yang , Xi Yang , Tianhao Yang , Yuehang Yang , Yuan Yang , Ximing Yang , Zeyu Yang , Guang Yang , Yifan Yang , Xuesong Yao , Bairen Yi , Fan Yin , Jianian Yin , Ziqiang Ying , Xiangyu Yu , Hongli Yu , Song Yu , Menghan Yu , Huan Yu , Siyu Yuan , Jun Yuan , Yutao Zeng , Tianyang Zhan , Zheng Zhang , Yun Zhang , Mofan Zhang , Wang Zhang , Ru Zhang , Zhi Zhang , Tianqi Zhang , Xinyi Zhang , Zhexi Zhang , Sijun Zhang , Wenqiang Zhang , Xiangxiang Zhang , Yongtao Zhang , Yuyu Zhang , Ge Zhang , He Zhang , Yue Zhang , Renjie Zheng , Ningxin Zheng , Zhuolin Zheng , Yaowei Zheng , Chen Zheng , Xiaoyun Zhi , Wanjun Zhong , Cheng Zhong , Zheng Zhong , Baoquan Zhong , Xun Zhou , Na Zhou , Huan Zhou , Hang Zhu , Defa Zhu , Wenjia Zhu , Lei Zuo

Iterative algorithms solve problems by taking steps until a solution is reached. Models in the form of Deep Thinking (DT) networks have been demonstrated to learn iterative algorithms in a way that can scale to different sized problems at…

Machine Learning · Computer Science 2024-11-01 Jay Bear , Adam Prügel-Bennett , Jonathon Hare

Although large language models (LLMs) have demonstrated remarkable reasoning capabilities, they still face challenges in knowledge-intensive multi-hop reasoning. Recent work explores iterative retrieval to address complex problems. However,…

Computation and Language · Computer Science 2025-05-27 Zheng Chu , Huiming Fan , Jingchang Chen , Qianyu Wang , Mingda Yang , Jiafeng Liang , Zhongjie Wang , Hao Li , Guo Tang , Ming Liu , Bing Qin

Deep learning (DL) creates impactful advances following a virtuous recipe: model architecture search, creating large training data sets, and scaling computation. It is widely believed that growing training sets and models should improve…

Despite the success of test-time scaling, Large Reasoning Models (LRMs) frequently encounter repetitive loops that lead to computational waste and inference failure. In this paper, we identify a distinct failure mode termed Circular…

Artificial Intelligence · Computer Science 2026-01-12 Zenghao Duan , Liang Pang , Zihao Wei , Wenbin Duan , Yuxin Tian , Shicheng Xu , Jingcheng Deng , Zhiyi Yin , Xueqi Cheng

Large Reasoning Models (LRMs) solve complex tasks by generating long Chain-of-Thought (CoT) sequences; however, the emergent dynamics governing reasoning trajectories are not well understood and can lead to inconsistencies and reasoning…

Artificial Intelligence · Computer Science 2026-05-29 G M Shahariar , Erfan Shayegani , Ali Nazari , Nael Abu-Ghazaleh

Scientific idea generation is a cornerstone of autonomous knowledge discovery, yet the iterative evolution required to transform initial concepts into high-quality research proposals remains a formidable challenge for Large Language Models…

Artificial Intelligence · Computer Science 2026-03-24 Andreas Sauter , Yuyue Zhao , Jacopo Urbani , Wenxiang Hu , Zaiqiao Meng , Lun Zhou , Xiaohui Yan , Yougang Lyu

Reasoning segmentation is an emerging vision-language task that requires reasoning over intricate text queries to precisely segment objects. However, existing methods typically suffer from overthinking, generating verbose reasoning chains…

Computer Vision and Pattern Recognition · Computer Science 2026-02-03 Yulin He , Wei Chen , Zhikang Jian , Tianhang Guo , Wenjuan Zhou , Minglong Li , Shaowu Yang , Wenjing Yang

Visual reasoning refers to the task of solving questions about visual information. Current visual reasoning methods typically employ pre-trained vision-language model (VLM) strategies or deep neural network approaches. However, existing…

Computer Vision and Pattern Recognition · Computer Science 2025-09-03 Chao Wang , Chunbai Zhang , Yongxiao Tian , Yang Zhou , Yan Peng

Most efforts to improve the reasoning capabilities of large language models (LLMs) involve either scaling the number of parameters and the size of training data, or scaling inference computation by letting models generate complex chains of…

Machine Learning · Computer Science 2025-10-10 Yeskendir Koishekenov , Aldo Lipani , Nicola Cancedda

Large Language Models exhibit impressive reasoning capabilities across diverse tasks, motivating efforts to distill these capabilities into smaller models through generated reasoning data. However, direct training on such synthesized…

Computation and Language · Computer Science 2025-02-05 Shengmin Piao , Sanghyun Park

Reasoning about failures is crucial for building reliable and trustworthy robotic systems. Prior approaches either treat failure reasoning as a closed-set classification problem or assume access to ample human annotations. Failures in the…

Large Language Models (LLMs) struggle with complex reasoning due to limited diversity and inefficient search. We propose Soft Reasoning, an embedding-based search framework that optimises the embedding of the first token to guide…

Computation and Language · Computer Science 2025-09-16 Qinglin Zhu , Runcong Zhao , Hanqi Yan , Yulan He , Yudong Chen , Lin Gui

Chain-of-thought (CoT) prompting enables Large Language Models to solve complex problems, but deploying these models safely requires reliable confidence estimates, a capability where existing methods suffer from poor calibration and severe…

Artificial Intelligence · Computer Science 2025-11-11 Abhishek More , Anthony Zhang , Nicole Bonilla , Ashvik Vivekan , Kevin Zhu , Parham Sharafoleslami , Maheep Chaudhary

We present a novel framework that bridges the gap between the interpretability of decision trees and the advanced reasoning capabilities of large language models (LLMs) to predict startup success. Our approach leverages chain-of-thought…

Artificial Intelligence · Computer Science 2025-04-17 Jack Preuveneers , Joseph Ternasky , Fuat Alican , Yigit Ihlamur

Large reasoning language models are typically run with fixed inference budgets, which can waste computation or terminate reasoning prematurely. We introduce Certainty-Guided Reasoning (CGR), a model-agnostic adaptive inference procedure…

Artificial Intelligence · Computer Science 2026-02-10 João Paulo Nogueira , Wentao Sun , Alonso Silva , Laith Zumot