English
Related papers

Related papers: Open-sci-ref-0.01: open and reproducible reference…

200 papers

The fast-growing demands in using Large Language Models (LLMs) to tackle complex multi-step data science tasks create an emergent need for accurate benchmarking. There are two major gaps in existing benchmarks: (i) the lack of standardized,…

Artificial Intelligence · Computer Science 2026-03-02 Fan Shu , Yite Wang , Ruofan Wu , Boyi Liu , Zhewei Yao , Yuxiong He , Feng Yan

In this work we explore recent advances in instruction-tuning language models on a range of open instruction-following datasets. Despite recent claims that open models can be on par with state-of-the-art proprietary models, these claims are…

DeepSeek-V3 and DeepSeek-R1 are leading open-source Large Language Models (LLMs) for general-purpose tasks and reasoning, achieving performance comparable to state-of-the-art closed-source models from companies like OpenAI and Anthropic --…

Machine Learning · Computer Science 2025-03-17 Chengen Wang , Murat Kantarcioglu

One problem with researching cognitive modeling and reinforcement learning (RL) is that researchers spend too much time on setting up an appropriate computational framework for their experiments. Many open source implementations of current…

Machine Learning · Computer Science 2024-01-29 Jan Dohmen , Frank Röder , Manfred Eppe

The Open Whisper-style Speech Model (OWSM) series was introduced to achieve full transparency in building advanced speech-to-text (S2T) foundation models. To this end, OWSM models are trained on 25 public speech datasets, which are…

Computation and Language · Computer Science 2024-06-14 Jinchuan Tian , Yifan Peng , William Chen , Kwanghee Choi , Karen Livescu , Shinji Watanabe

In this report, we present ChuXin, an entirely open-source language model with a size of 1.6 billion parameters. Unlike the majority of works that only open-sourced the model weights and architecture, we have made everything needed to train…

Computation and Language · Computer Science 2024-05-09 Xiaomin Zhuang , Yufan Jiang , Qiaozhi He , Zhihua Wu

Accurate parsing of citations is necessary for machine-readable scholarly infrastructure. But, despite sustained interest in this problem, existing evaluation techniques are often not generalizable, based on synthetic data, or not publicly…

Digital Libraries · Computer Science 2026-03-27 Parth Sarin , Juan Pablo Alperin , Adam Buttrick , Dione Mentis

Large, high-capacity models trained on diverse datasets have shown remarkable successes on efficiently tackling downstream applications. In domains from NLP to Computer Vision, this has led to a consolidation of pretrained models, with…

Robotics · Computer Science 2025-05-15 Embodiment Collaboration , Abby O'Neill , Abdul Rehman , Abhinav Gupta , Abhiram Maddukuri , Abhishek Gupta , Abhishek Padalkar , Abraham Lee , Acorn Pooley , Agrim Gupta , Ajay Mandlekar , Ajinkya Jain , Albert Tung , Alex Bewley , Alex Herzog , Alex Irpan , Alexander Khazatsky , Anant Rai , Anchit Gupta , Andrew Wang , Andrey Kolobov , Anikait Singh , Animesh Garg , Aniruddha Kembhavi , Annie Xie , Anthony Brohan , Antonin Raffin , Archit Sharma , Arefeh Yavary , Arhan Jain , Ashwin Balakrishna , Ayzaan Wahid , Ben Burgess-Limerick , Beomjoon Kim , Bernhard Schölkopf , Blake Wulfe , Brian Ichter , Cewu Lu , Charles Xu , Charlotte Le , Chelsea Finn , Chen Wang , Chenfeng Xu , Cheng Chi , Chenguang Huang , Christine Chan , Christopher Agia , Chuer Pan , Chuyuan Fu , Coline Devin , Danfei Xu , Daniel Morton , Danny Driess , Daphne Chen , Deepak Pathak , Dhruv Shah , Dieter Büchler , Dinesh Jayaraman , Dmitry Kalashnikov , Dorsa Sadigh , Edward Johns , Ethan Foster , Fangchen Liu , Federico Ceola , Fei Xia , Feiyu Zhao , Felipe Vieira Frujeri , Freek Stulp , Gaoyue Zhou , Gaurav S. Sukhatme , Gautam Salhotra , Ge Yan , Gilbert Feng , Giulio Schiavi , Glen Berseth , Gregory Kahn , Guangwen Yang , Guanzhi Wang , Hao Su , Hao-Shu Fang , Haochen Shi , Henghui Bao , Heni Ben Amor , Henrik I Christensen , Hiroki Furuta , Homanga Bharadhwaj , Homer Walke , Hongjie Fang , Huy Ha , Igor Mordatch , Ilija Radosavovic , Isabel Leal , Jacky Liang , Jad Abou-Chakra , Jaehyung Kim , Jaimyn Drake , Jan Peters , Jan Schneider , Jasmine Hsu , Jay Vakil , Jeannette Bohg , Jeffrey Bingham , Jeffrey Wu , Jensen Gao , Jiaheng Hu , Jiajun Wu , Jialin Wu , Jiankai Sun , Jianlan Luo , Jiayuan Gu , Jie Tan , Jihoon Oh , Jimmy Wu , Jingpei Lu , Jingyun Yang , Jitendra Malik , João Silvério , Joey Hejna , Jonathan Booher , Jonathan Tompson , Jonathan Yang , Jordi Salvador , Joseph J. Lim , Junhyek Han , Kaiyuan Wang , Kanishka Rao , Karl Pertsch , Karol Hausman , Keegan Go , Keerthana Gopalakrishnan , Ken Goldberg , Kendra Byrne , Kenneth Oslund , Kento Kawaharazuka , Kevin Black , Kevin Lin , Kevin Zhang , Kiana Ehsani , Kiran Lekkala , Kirsty Ellis , Krishan Rana , Krishnan Srinivasan , Kuan Fang , Kunal Pratap Singh , Kuo-Hao Zeng , Kyle Hatch , Kyle Hsu , Laurent Itti , Lawrence Yunliang Chen , Lerrel Pinto , Li Fei-Fei , Liam Tan , Linxi "Jim" Fan , Lionel Ott , Lisa Lee , Luca Weihs , Magnum Chen , Marion Lepert , Marius Memmel , Masayoshi Tomizuka , Masha Itkina , Mateo Guaman Castro , Max Spero , Maximilian Du , Michael Ahn , Michael C. Yip , Mingtong Zhang , Mingyu Ding , Minho Heo , Mohan Kumar Srirama , Mohit Sharma , Moo Jin Kim , Muhammad Zubair Irshad , Naoaki Kanazawa , Nicklas Hansen , Nicolas Heess , Nikhil J Joshi , Niko Suenderhauf , Ning Liu , Norman Di Palo , Nur Muhammad Mahi Shafiullah , Oier Mees , Oliver Kroemer , Osbert Bastani , Pannag R Sanketi , Patrick "Tree" Miller , Patrick Yin , Paul Wohlhart , Peng Xu , Peter David Fagan , Peter Mitrano , Pierre Sermanet , Pieter Abbeel , Priya Sundaresan , Qiuyu Chen , Quan Vuong , Rafael Rafailov , Ran Tian , Ria Doshi , Roberto Martín-Martín , Rohan Baijal , Rosario Scalise , Rose Hendrix , Roy Lin , Runjia Qian , Ruohan Zhang , Russell Mendonca , Rutav Shah , Ryan Hoque , Ryan Julian , Samuel Bustamante , Sean Kirmani , Sergey Levine , Shan Lin , Sherry Moore , Shikhar Bahl , Shivin Dass , Shubham Sonawani , Shubham Tulsiani , Shuran Song , Sichun Xu , Siddhant Haldar , Siddharth Karamcheti , Simeon Adebola , Simon Guist , Soroush Nasiriany , Stefan Schaal , Stefan Welker , Stephen Tian , Subramanian Ramamoorthy , Sudeep Dasari , Suneel Belkhale , Sungjae Park , Suraj Nair , Suvir Mirchandani , Takayuki Osa , Tanmay Gupta , Tatsuya Harada , Tatsuya Matsushima , Ted Xiao , Thomas Kollar , Tianhe Yu , Tianli Ding , Todor Davchev , Tony Z. Zhao , Travis Armstrong , Trevor Darrell , Trinity Chung , Vidhi Jain , Vikash Kumar , Vincent Vanhoucke , Vitor Guizilini , Wei Zhan , Wenxuan Zhou , Wolfram Burgard , Xi Chen , Xiangyu Chen , Xiaolong Wang , Xinghao Zhu , Xinyang Geng , Xiyuan Liu , Xu Liangwei , Xuanlin Li , Yansong Pang , Yao Lu , Yecheng Jason Ma , Yejin Kim , Yevgen Chebotar , Yifan Zhou , Yifeng Zhu , Yilin Wu , Ying Xu , Yixuan Wang , Yonatan Bisk , Yongqiang Dou , Yoonyoung Cho , Youngwoon Lee , Yuchen Cui , Yue Cao , Yueh-Hua Wu , Yujin Tang , Yuke Zhu , Yunchu Zhang , Yunfan Jiang , Yunshuang Li , Yunzhu Li , Yusuke Iwasawa , Yutaka Matsuo , Zehan Ma , Zhuo Xu , Zichen Jeff Cui , Zichen Zhang , Zipeng Fu , Zipeng Lin

Grounding-DINO is a state-of-the-art open-set detection model that tackles multiple vision tasks including Open-Vocabulary Detection (OVD), Phrase Grounding (PG), and Referring Expression Comprehension (REC). Its effectiveness has led to…

Computer Vision and Pattern Recognition · Computer Science 2024-01-08 Xiangyu Zhao , Yicheng Chen , Shilin Xu , Xiangtai Li , Xinjiang Wang , Yining Li , Haian Huang

Deep research systems represent an emerging class of agentic information retrieval methods that generate comprehensive and well-supported reports to complex queries. However, most existing frameworks rely on dynamic commercial search APIs,…

This paper presents a configuration-first framework for evaluating cross-backend compatibility in deep learning systems deployed on CPU, GPU, and compiled runtimes. The framework decouples experiments from code using YAML, supports both…

Machine Learning · Computer Science 2025-09-10 Zehua Li

Large foundational models, through upstream pre-training and downstream fine-tuning, have achieved immense success in the broad AI community due to improved model performance and significant reductions in repetitive engineering. By…

Information Retrieval · Computer Science 2024-03-19 Jiaqi Zhang , Yu Cheng , Yongxin Ni , Yunzhu Pan , Zheng Yuan , Junchen Fu , Youhua Li , Jie Wang , Fajie Yuan

The flourishing blossom of deep learning has witnessed the rapid development of text recognition in recent years. However, the existing text recognition methods are mainly proposed for English texts. As another widely-spoken language,…

Computer Vision and Pattern Recognition · Computer Science 2022-11-28 Haiyang Yu , Jingye Chen , Bin Li , Jianqi Ma , Mengnan Guan , Xixi Xu , Xiaocong Wang , Shaobo Qu , Xiangyang Xue

We present RoboManipBaselines, an open-source software framework for imitation learning research in robotic manipulation. The framework supports the entire imitation learning pipeline, including data collection, policy training, and…

Reinforcement learning (RL) with large language models shows promise in complex reasoning. However, its progress is hindered by the lack of large-scale training data that is sufficiently challenging, contamination-free and verifiable. To…

Continual learning (CL) in large language models (LLMs) is an evolving domain that focuses on developing efficient and sustainable training strategies to adapt models to emerging knowledge and achieve robustness in dynamic environments. Our…

Computation and Language · Computer Science 2025-02-13 Çağatay Yıldız , Nishaanth Kanna Ravichandran , Nitin Sharma , Matthias Bethge , Beyza Ermis

Large language models (LLMs) are increasingly expected to go beyond simple factual queries toward Deep Research-tasks that require decomposing questions into sub-problems, coordinating multi-step reasoning, and synthesizing evidence from…

Computation and Language · Computer Science 2025-09-03 Ziyi Xia , Kun Luo , Hongjin Qian , Zheng Liu

Evaluating large language models (LLMs) for medical applications remains challenging due to benchmark saturation, limited data accessibility, and insufficient coverage of relevant tasks. Existing suites have either saturated, heavily depend…

Due to limited supervised training data, large language models (LLMs) are typically pre-trained via a self-supervised "predict the next word" objective on a vast amount of unstructured text data. To make the resulting model useful to users,…

Computation and Language · Computer Science 2026-01-30 Ajay Patel , Colin Raffel , Chris Callison-Burch

Optical data center networks (DCNs) are emerging as a promising design for cloud infrastructure. However, existing optical DCN architectures operate as closed ecosystems, tying software solutions to specific optical hardware. We introduce…

Networking and Internet Architecture · Computer Science 2025-08-07 Yiming Lei , Federico De Marchi , Jialong Li , Raj Joshi , Balakrishnan Chandrasekaran , Yiting Xia
‹ Prev 1 8 9 10 Next ›