English
Related papers

Related papers: Seed2Scale: A Self-Evolving Data Engine for Embodi…

200 papers

We propose IR2Vec, a Concise and Scalable encoding infrastructure to represent programs as a distributed embedding in continuous space. This distributed embedding is obtained by combining representation learning methods with flow…

Programming Languages · Computer Science 2020-12-25 S. VenkataKeerthy , Rohit Aggarwal , Shalini Jain , Maunendra Sankar Desarkar , Ramakrishna Upadrasta , Y. N. Srikant

Multidimensional scaling of gene sequence data has long played a vital role in analysing gene sequence data to identify clusters and patterns. However the computation complexities and memory requirements of state-of-the-art dimensional…

Artificial Intelligence · Computer Science 2021-04-20 Pulasthi Wickramasinghe , Geoffrey Fox

Precise and rapid delineation of sharp boundaries and robust semantics is essential for numerous downstream robotic tasks, such as robot grasping and manipulation, real-time semantic mapping, and online sensor calibration performed on edge…

Computer Vision and Pattern Recognition · Computer Science 2024-03-12 Youqi Liao , Shuhao Kang , Jianping Li , Yang Liu , Yun Liu , Zhen Dong , Bisheng Yang , Xieyuanli Chen

The rapid evolution of multimodal foundation model has demonstrated significant progresses in vision-language understanding and generation, e.g., our previous work SEED-LLaMA. However, there remains a gap between its capability and the…

Computer Vision and Pattern Recognition · Computer Science 2025-03-04 Yuying Ge , Sijie Zhao , Jinguo Zhu , Yixiao Ge , Kun Yi , Lin Song , Chen Li , Xiaohan Ding , Ying Shan

Advancements in foundation models have catalyzed research in Embodied AI to develop interactive agents capable of environmental reasoning and interaction. Developing such agents requires diverse, large-scale datasets. Prior frameworks…

Robotics · Computer Science 2026-02-10 Siddharth Singh , Ifrah Idrees , Abraham Dauhajre

Simulation is increasingly being used for generating large labelled datasets in many machine learning problems. Recent methods have focused on adjusting simulator parameters with the goal of maximising accuracy on a validation task, usually…

Computer Vision and Pattern Recognition · Computer Science 2020-08-20 Harkirat Singh Behl , Atılım Güneş Baydin , Ran Gal , Philip H. S. Torr , Vibhav Vineet

With the rapid advancement of low-altitude remote sensing and Vision-Language Models (VLMs), Embodied Agents based on Unmanned Aerial Vehicles (UAVs) have shown significant potential in autonomous tasks. However, current evaluation methods…

Robotics · Computer Science 2025-12-09 Mingning Guo , Mengwei Wu , Jiarun He , Shaoxian Li , Haifeng Li , Chao Tao

Embodied AI requires agents to understand goals, plan actions, and execute tasks in simulated environments. We present a comprehensive evaluation of Large Language Models (LLMs) on the VirtualHome benchmark using the Embodied Agent…

Artificial Intelligence · Computer Science 2026-02-04 Jiaqi Xu , Tao Huang , Kai Zhang

As new data and updates are constantly arriving, the results of data mining applications become stale and obsolete over time. Incremental processing is a promising approach to refreshing mining results. It utilizes previously saved states…

Distributed, Parallel, and Cluster Computing · Computer Science 2015-01-21 Yanfeng Zhang , Shimin Chen , Qiang Wang , Ge Yu

We introduce Seed-TTS, a family of large-scale autoregressive text-to-speech (TTS) models capable of generating speech that is virtually indistinguishable from human speech. Seed-TTS serves as a foundation model for speech generation and…

In an era increasingly shaped by decentralized knowledge ecosystems and pervasive AI technologies, fostering sustainable learner agency has become a critical educational imperative. This study introduces a novel conceptual framework…

Computers and Society · Computer Science 2025-04-30 Qianrun Mao

Recent strides in video generation have paved the way for unified audio-visual generation. In this work, we present Seedance 1.5 pro, a foundational model engineered specifically for native, joint audio-video generation. Leveraging a…

Computer Vision and Pattern Recognition · Computer Science 2025-12-24 Team Seedance , Heyi Chen , Siyan Chen , Xin Chen , Yanfei Chen , Ying Chen , Zhuo Chen , Feng Cheng , Tianheng Cheng , Xinqi Cheng , Xuyan Chi , Jian Cong , Jing Cui , Qinpeng Cui , Qide Dong , Junliang Fan , Jing Fang , Zetao Fang , Chengjian Feng , Han Feng , Mingyuan Gao , Yu Gao , Dong Guo , Qiushan Guo , Boyang Hao , Qingkai Hao , Bibo He , Qian He , Tuyen Hoang , Ruoqing Hu , Xi Hu , Weilin Huang , Zhaoyang Huang , Zhongyi Huang , Donglei Ji , Siqi Jiang , Wei Jiang , Yunpu Jiang , Zhuo Jiang , Ashley Kim , Jianan Kong , Zhichao Lai , Shanshan Lao , Yichong Leng , Ai Li , Feiya Li , Gen Li , Huixia Li , JiaShi Li , Liang Li , Ming Li , Shanshan Li , Tao Li , Xian Li , Xiaojie Li , Xiaoyang Li , Xingxing Li , Yameng Li , Yifu Li , Yiying Li , Chao Liang , Han Liang , Jianzhong Liang , Ying Liang , Zhiqiang Liang , Wang Liao , Yalin Liao , Heng Lin , Kengyu Lin , Shanchuan Lin , Xi Lin , Zhijie Lin , Feng Ling , Fangfang Liu , Gaohong Liu , Jiawei Liu , Jie Liu , Jihao Liu , Shouda Liu , Shu Liu , Sichao Liu , Songwei Liu , Xin Liu , Xue Liu , Yibo Liu , Zikun Liu , Zuxi Liu , Junlin Lyu , Lecheng Lyu , Qian Lyu , Han Mu , Xiaonan Nie , Jingzhe Ning , Xitong Pan , Yanghua Peng , Lianke Qin , Xueqiong Qu , Yuxi Ren , Kai Shen , Guang Shi , Lei Shi , Yan Song , Yinglong Song , Fan Sun , Li Sun , Renfei Sun , Yan Sun , Zeyu Sun , Wenjing Tang , Yaxue Tang , Zirui Tao , Feng Wang , Furui Wang , Jinran Wang , Junkai Wang , Ke Wang , Kexin Wang , Qingyi Wang , Rui Wang , Sen Wang , Shuai Wang , Tingru Wang , Weichen Wang , Xin Wang , Yanhui Wang , Yue Wang , Yuping Wang , Yuxuan Wang , Ziyu Wang , Guoqiang Wei , Wanru Wei , Di Wu , Guohong Wu , Hanjie Wu , Jian Wu , Jie Wu , Ruolan Wu , Xinglong Wu , Yonghui Wu , Ruiqi Xia , Liang Xiang , Fei Xiao , XueFeng Xiao , Pan Xie , Shuangyi Xie , Shuang Xu , Jinlan Xue , Shen Yan , Bangbang Yang , Ceyuan Yang , Jiaqi Yang , Runkai Yang , Tao Yang , Yang Yang , Yihang Yang , ZhiXian Yang , Ziyan Yang , Songting Yao , Yifan Yao , Zilyu Ye , Bowen Yu , Jian Yu , Chujie Yuan , Linxiao Yuan , Sichun Zeng , Weihong Zeng , Xuejiao Zeng , Yan Zeng , Chuntao Zhang , Heng Zhang , Jingjie Zhang , Kuo Zhang , Liang Zhang , Liying Zhang , Manlin Zhang , Ting Zhang , Weida Zhang , Xiaohe Zhang , Xinyan Zhang , Yan Zhang , Yuan Zhang , Zixiang Zhang , Fengxuan Zhao , Huating Zhao , Yang Zhao , Hao Zheng , Jianbin Zheng , Xiaozheng Zheng , Yangyang Zheng , Yijie Zheng , Jiexin Zhou , Jiahui Zhu , Kuan Zhu , Shenhan Zhu , Wenjia Zhu , Benhui Zou , Feilong Zuo

The role of AI-generated synthetic data has recently been expanded to support realistic Monte Carlo simulations. However, guidance is limited on generating data with multilevel structures and designing simulations based on such data. This…

Methodology · Statistics 2026-05-08 Youmi Suk , Chenguang Pan , Weixuan Xiao

Autoregressive image modeling relies on visual tokenizers to compress images into compact latent representations. We design an end-to-end training pipeline that jointly optimizes reconstruction and generation, enabling direct supervision…

Computer Vision and Pattern Recognition · Computer Science 2026-05-05 Wenda Chu , Bingliang Zhang , Jiaqi Han , Yizhuo Li , Linjie Yang , Yisong Yue , Qiushan Guo

High-fidelity text-to-music generation typically relies on massive proprietary datasets and immense computational resources. Existing models often struggle to generate coherent pure musical accompaniments and lack precise, localized…

Audio and Speech Processing · Electrical Eng. & Systems 2026-05-19 Huakang Chen , Wenkai Cheng , Guobin Ma , Chunbo Hao , Yuxuan Xia , Mengqi Wei , Zhixian Zhao , Pengcheng Zhu , Hanbing Zhang , Lei Xie

The rapid advancement in self-supervised representation learning has highlighted its potential to leverage unlabeled data for learning rich visual representations. However, the existing techniques, particularly those employing different…

Computer Vision and Pattern Recognition · Computer Science 2024-12-18 Sana Ayromlou , Vahid Reza Khazaie , Fereshteh Forghani , Arash Afkanpour

Building models capable of generating structured output is a key challenge for AI and robotics. While generative models have been explored on many types of data, little work has been done on synthesizing lidar scans, which play a key role…

Computer Vision and Pattern Recognition · Computer Science 2019-12-04 Lucas Caccia , Herke van Hoof , Aaron Courville , Joelle Pineau

Gradient-based optimization has been critical to the success of machine learning, updating a single set of parameters to minimize a single loss. A growing number of applications rely on a generalization of this, where we have a bilevel or…

Machine Learning · Computer Science 2024-07-02 Jonathan Lorraine

Visual Auto-Regressive (VAR) models significantly reduce inference steps through the "next-scale" prediction paradigm. However, progressive multi-scale generation incurs substantial memory overhead due to cumulative KV caching, limiting…

Computer Vision and Pattern Recognition · Computer Science 2025-11-21 Xiaoyue Chen , Yuling Shi , Kaiyuan Li , Huandong Wang , Yong Li , Xiaodong Gu , Xinlei Chen , Mingbao Lin

AI modeling for source code understanding tasks has been making significant progress, and is being adopted in production development pipelines. However, reliability concerns, especially whether the models are actually learning task-related…

Software Engineering · Computer Science 2022-01-11 Sahil Suneja , Yufan Zhuang , Yunhui Zheng , Jim Laredo , Alessandro Morari
‹ Prev 1 3 4 5 6 7 10 Next ›