English
Related papers

Related papers: ComfyUI-R1: Exploring Reasoning Models for Workflo…

200 papers

Recently, there have been notable advancements in large language models (LLMs), demonstrating their growing abilities in complex reasoning. However, existing research largely overlooks a thorough and systematic comparison of these models'…

Computation and Language · Computer Science 2025-06-30 Junhao Liu , Zhenhao Xu , Yuxin Fang , Yichuan Chen , Zuobin Ying , Wenhan Chang

Large language models (LLMs) excel in speed and adaptability across various reasoning tasks, but they often struggle when strict logic or constraint enforcement is required. In contrast, Large Reasoning Models (LRMs) are specifically…

Large language models (LLMs) have recently shown strong reasoning abilities in domains like mathematics, coding, and scientific problem-solving, yet their potential for ranking tasks, where prime examples include retrieval, recommender…

Information Retrieval · Computer Science 2025-10-17 Tao Feng , Zhigang Hua , Zijie Lei , Yan Xie , Shuang Yang , Bo Long , Jiaxuan You

Chain-of-Thought (CoT) prompting has shown promise in enhancing the reasoning capabilities of large language models (LLMs) by generating natural language (NL) rationales that lead to the final answer. However, it struggles with numerical…

Artificial Intelligence · Computer Science 2025-02-13 Cheryl Li , Tianyuan Xu , Yiwen Guo

Recent studies have shown that long chain-of-thought (CoT) reasoning can significantly enhance the performance of large language models (LLMs) on complex tasks. However, this benefit is yet to be demonstrated in the domain of video…

Computer Vision and Pattern Recognition · Computer Science 2026-03-18 Yuanxin Liu , Kun Ouyang , Haoning Wu , Yi Liu , Lin Sui , Xinhao Li , Yan Zhong , Y. Charles , Xinyu Zhou , Xu Sun

Chain-of-thought (CoT) reasoning has enabled large language models (LLMs) to utilize additional computation through intermediate tokens to solve complex tasks. However, we posit that typical reasoning traces contain many redundant tokens,…

Computation and Language · Computer Science 2025-06-11 Tergel Munkhbat , Namgyu Ho , Seo Hyun Kim , Yongjin Yang , Yujin Kim , Se-Young Yun

DeepSeek-R1, known for its low training cost and exceptional reasoning capabilities, has achieved state-of-the-art performance on various benchmarks. However, detailed evaluations for DeepSeek Series models from the perspective of…

Chain-of-Thought (CoT) reasoning enhances Large Language Models (LLMs) by encouraging step-by-step reasoning in natural language. However, leveraging a latent continuous space for reasoning may offer benefits in terms of both efficiency and…

Computation and Language · Computer Science 2025-09-24 Zhenyi Shen , Hanqi Yan , Linhai Zhang , Zhanghao Hu , Yali Du , Yulan He

Medical Question-Answering (QA) encompasses a broad spectrum of tasks, including multiple choice questions (MCQ), open-ended text generation, and complex computational reasoning. Despite this variety, a unified framework for delivering…

Computation and Language · Computer Science 2025-06-23 Xiaotian Zhang , Yuan Wang , Zhaopeng Feng , Ruizhe Chen , Zhijie Zhou , Yan Zhang , Hongxia Xu , Jian Wu , Zuozhu Liu

Chain-of-Thought (CoT) prompting helps Large Language Models (LLMs) tackle complex reasoning by eliciting explicit step-by-step rationales. However, CoT's verbosity increases latency and memory usage and may propagate early errors across…

Computation and Language · Computer Science 2025-09-30 Hongyu Shan , Mingyang Song , Chang Dai , Di Liang , Han Chen

Large language models (LLMs) trained via reinforcement learning with verifiable reward (RLVR) have achieved breakthroughs on tasks with explicit, automatable verification, such as software programming and mathematical problems. Extending…

Recent reasoning models through test-time scaling have demonstrated that long chain-of-thoughts can unlock substantial performance boosts in hard reasoning tasks such as math and code. However, the benefit of such long thoughts for system-2…

Computer Vision and Pattern Recognition · Computer Science 2025-04-23 Yuan-Hong Liao , Sven Elflein , Liu He , Laura Leal-Taixé , Yejin Choi , Sanja Fidler , David Acuna

Long-context reasoning has significantly empowered large language models (LLMs) to tackle complex tasks, yet it introduces severe efficiency bottlenecks due to the computational complexity. Existing efficient approaches often rely on…

Computation and Language · Computer Science 2026-02-03 Yibo Wang , Yongcheng Jing , Shunyu Liu , Hao Guan , Rong-cheng Tu , Chengyu Wang , Jun Huang , Dacheng Tao

Large language models (LLMs) exhibit strong generative capabilities and have shown great potential in code generation. Existing chain-of-thought (CoT) prompting methods enhance model reasoning by eliciting intermediate steps, but suffer…

Artificial Intelligence · Computer Science 2025-12-17 Shen Li , Li Huang , Shaoxiong Zhan , Weifeng Sun , Tao Yin , Zhongxin Liu , Meng Yan

The o1 model series is trained with large-scale reinforcement learning to reason using chain of thought. These advanced reasoning capabilities provide new avenues for improving the safety and robustness of our models. In particular, our…

Artificial Intelligence · Computer Science 2026-05-01 OpenAI , : , Aaron Jaech , Adam Kalai , Adam Lerer , Adam Richardson , Ahmed El-Kishky , Aiden Low , Alec Helyar , Aleksander Madry , Alex Beutel , Alex Carney , Alex Iftimie , Alex Karpenko , Alex Tachard Passos , Alexander Neitz , Alexander Prokofiev , Alexander Wei , Allison Tam , Ally Bennett , Ananya Kumar , Andre Saraiva , Andrea Vallone , Andrew Duberstein , Andrew Kondrich , Andrey Mishchenko , Andy Applebaum , Angela Jiang , Ashvin Nair , Barret Zoph , Behrooz Ghorbani , Bohan Zhang , Ben Rossen , Benjamin Sokolowsky , Boaz Barak , Bob McGrew , Borys Minaiev , Botao Hao , Bowen Baker , Brandon Houghton , Brandon McKinzie , Brydon Eastman , Camillo Lugaresi , Cary Bassin , Cary Hudson , Chak Ming Li , Charles de Bourcy , Chelsea Voss , Chen Shen , Chong Zhang , Chris Koch , Chris Orsinger , Christopher Hesse , Claudia Fischer , Clive Chan , Dan Roberts , Daniel Kappler , Daniel Levy , Daniel Selsam , David Dohan , David Farhi , David Mely , David Robinson , Dimitris Tsipras , Doug Li , Dragos Oprica , Eben Freeman , Eddie Zhang , Edmund Wong , Elizabeth Proehl , Enoch Cheung , Eric Mitchell , Eric Wallace , Erik Ritter , Evan Mays , Fan Wang , Felipe Petroski Such , Filippo Raso , Florencia Leoni , Foivos Tsimpourlas , Francis Song , Fred von Lohmann , Freddie Sulit , Geoff Salmon , Giambattista Parascandolo , Gildas Chabot , Grace Zhao , Greg Brockman , Guillaume Leclerc , Hadi Salman , Haiming Bao , Hao Sheng , Hart Andrin , Hessam Bagherinezhad , Hongyu Ren , Hunter Lightman , Hyung Won Chung , Ian Kivlichan , Ian O'Connell , Ian Osband , Ignasi Clavera Gilaberte , Ilge Akkaya , Ilya Kostrikov , Ilya Sutskever , Irina Kofman , Jakub Pachocki , James Lennon , Jason Wei , Jean Harb , Jerry Twore , Jiacheng Feng , Jiahui Yu , Jiayi Weng , Jie Tang , Jieqi Yu , Joaquin Quiñonero Candela , Joe Palermo , Joel Parish , Johannes Heidecke , John Hallman , John Rizzo , Jonathan Gordon , Jonathan Uesato , Jonathan Ward , Joost Huizinga , Julie Wang , Kai Chen , Kai Xiao , Karan Singhal , Karina Nguyen , Karl Cobbe , Katy Shi , Kayla Wood , Kendra Rimbach , Keren Gu-Lemberg , Kevin Liu , Kevin Lu , Kevin Stone , Kevin Yu , Lama Ahmad , Lauren Yang , Leo Liu , Leon Maksin , Leyton Ho , Liam Fedus , Lilian Weng , Linden Li , Lindsay McCallum , Lindsey Held , Lorenz Kuhn , Lukas Kondraciuk , Lukasz Kaiser , Luke Metz , Madelaine Boyd , Maja Trebacz , Manas Joglekar , Mark Chen , Marko Tintor , Mason Meyer , Matt Jones , Matt Kaufer , Max Schwarzer , Meghan Shah , Mehmet Yatbaz , Melody Y. Guan , Mengyuan Xu , Mengyuan Yan , Mia Glaese , Mianna Chen , Michael Lampe , Michael Malek , Michele Wang , Michelle Fradin , Mike McClay , Mikhail Pavlov , Miles Wang , Mingxuan Wang , Mira Murati , Mo Bavarian , Mostafa Rohaninejad , Nat McAleese , Neil Chowdhury , Neil Chowdhury , Nick Ryder , Nikolas Tezak , Noam Brown , Ofir Nachum , Oleg Boiko , Oleg Murk , Olivia Watkins , Patrick Chao , Paul Ashbourne , Pavel Izmailov , Peter Zhokhov , Rachel Dias , Rahul Arora , Randall Lin , Rapha Gontijo Lopes , Raz Gaon , Reah Miyara , Reimar Leike , Renny Hwang , Rhythm Garg , Robin Brown , Roshan James , Rui Shu , Ryan Cheu , Ryan Greene , Saachi Jain , Sam Altman , Sam Toizer , Sam Toyer , Samuel Miserendino , Sandhini Agarwal , Santiago Hernandez , Sasha Baker , Scott McKinney , Scottie Yan , Shengjia Zhao , Shengli Hu , Shibani Santurkar , Shraman Ray Chaudhuri , Shuyuan Zhang , Siyuan Fu , Spencer Papay , Steph Lin , Suchir Balaji , Suvansh Sanjeev , Szymon Sidor , Tal Broda , Aidan Clark , Tao Wang , Taylor Gordon , Ted Sanders , Tejal Patwardhan , Thibault Sottiaux , Thomas Degry , Thomas Dimson , Tianhao Zheng , Timur Garipov , Tom Stasi , Trapit Bansal , Trevor Creech , Troy Peterson , Tyna Eloundou , Valerie Qi , Vineet Kosaraju , Vinnie Monaco , Vitchyr Pong , Vlad Fomenko , Weiyi Zheng , Wenda Zhou , Wenting Zhan , Wes McCabe , Wojciech Zaremba , Yann Dubois , Yinghai Lu , Yining Chen , Young Cha , Yu Bai , Yuchen He , Yuchen Zhang , Yunyun Wang , Zheng Shao , Zhuohan Li

Reasoning-oriented Large Language Models (LLMs) often rely on generating explicit tokens step by step, and their effectiveness typically hinges on large-scale supervised fine-tuning or reinforcement learning. While Chain-of-Thought (CoT)…

Computation and Language · Computer Science 2025-09-30 Haoyu Zheng , Zhuonan Wang , Yuqian Yuan , Tianwei Lin , Wenqiao Zhang , Zheqi Lv , Juncheng Li , Siliang Tang , Yueting Zhuang , Hongyang He

The Chain-of-Thought (CoT) paradigm, while enhancing the interpretability of Large Language Models (LLMs), is constrained by the inefficiencies and expressive limits of natural language. Latent Chain-of-Thought (latent CoT) reasoning, which…

Computation and Language · Computer Science 2026-05-12 Xiaocheng Luo , Kang Wang , Zaifu Zhan , Yuechi Zhou , Xiangyu Duan

Chain-of-Thought (CoT) reasoning, which breaks down complex tasks into intermediate reasoning steps, has significantly enhanced the performance of large language models (LLMs) on challenging tasks. However, the detailed reasoning process in…

Computation and Language · Computer Science 2025-02-20 Yingqian Cui , Pengfei He , Jingying Zeng , Hui Liu , Xianfeng Tang , Zhenwei Dai , Yan Han , Chen Luo , Jing Huang , Zhen Li , Suhang Wang , Yue Xing , Jiliang Tang , Qi He

Large language models (LLMs) achieve strong performance on code generation, but the mechanisms by which Chain-of-Thought (CoT) prompting helps remain unclear. We present a systematic empirical and information-theoretic study of CoT…

Software Engineering · Computer Science 2025-12-11 Naizhu Jin , Zhong Li , Guang Yang , Tian Zhang , Qingkai Zeng

Recent advances in large language models (LLMs) have accelerated AI-assisted software development, yet practical deployment remains constrained by incomplete implementations, weak modularization, and inconsistent security practices. We…

Software Engineering · Computer Science 2026-03-13 Yen-Ku Liu , Yun-Cheng Tsai