中文
相关论文

相关论文: Human Preference-Based Learning for High-dimension…

200 篇论文

Recent advancements in Direct Preference Optimization (DPO) have significantly enhanced the alignment of Large Language Models (LLMs) with human preferences, owing to its simplicity and effectiveness. However, existing methods typically…

计算与语言 · 计算机科学 2024-10-28 Shilong Li , Yancheng He , Hui Huang , Xingyuan Bu , Jiaheng Liu , Hangyu Guo , Weixun Wang , Jihao Gu , Wenbo Su , Bo Zheng

This work presents an extended framework for learning-based bipedal locomotion that incorporates a heuristic step-planning strategy guided by desired torso velocity tracking. The framework enables precise interaction between a humanoid…

机器人学 · 计算机科学 2025-12-01 William Suliman , Ekaterina Chaikovskaia , Egor Davydenko , Roman Gorbachev

Human preference research is a significant domain in psychology and psychophysiology, with broad applications in psychiatric evaluation and daily life quality enhancement. This study explores the neural mechanisms of human preference…

神经元与认知 · 定量生物学 2025-05-27 Siyuan Li , Xiangze Meng , Yijian Yang , Yiwen Xu , Yunfei Wang , Chenghu Qiu , Hanyi Jiang , Pin Wu , Shegnbo Chen , Xiao Wei , Hao Wang , Lan Ni , Huiran Zhang

Gait phase estimation based on inertial measurement unit (IMU) signals facilitates precise adaptation of exoskeletons to individual gait variations. However, challenges remain in achieving high accuracy and robustness, particularly during…

机器人学 · 计算机科学 2025-06-19 Yuanlong Ji , Xingbang Yang , Ruoqi Zhao , Qihan Ye , Quan Zheng , Yubo Fan

To proactively navigate and traverse various terrains, active use of visual perception becomes indispensable. We aim to investigate the feasibility and performance of using sparse visual observations to achieve perceptual locomotion over a…

机器人学 · 计算机科学 2022-05-27 Fernando Acero , Kai Yuan , Zhibin Li

Human motion prediction is a stochastic process: Given an observed sequence of poses, multiple future motions are plausible. Existing approaches to modeling this stochasticity typically combine a random noise vector with information about…

Many kinds of lower-limb exoskeletons were developed for walking assistance. However, when controlling these exoskeletons, time-delay due to the computation time and the communication delays is still a general problem. In this research, we…

机器人学 · 计算机科学 2020-08-07 Ming Ding , Mikihisa Nagashima , Sung-Gwi Cho , Jun Takamatsu , Tsukasa Ogasawara

A mesoscopic approach to modeling pedestrian simulation with multiple exits is proposed in this paper. A floor field based on Qlearning Algorithm is used. Attractiveness of exits to pedestrian typically is based on shortest path. However,…

多智能体系统 · 计算机科学 2016-09-07 Allan Lao , Kardi Teknomo

Patterns of human motion in outdoor and indoor environments are substantially different due to the scope of the environment and the typical intentions of people therein. While outdoor trajectory forecasting has received significant…

计算机视觉与模式识别 · 计算机科学 2024-09-04 Luigi Capogrosso , Andrea Toaiari , Andrea Avogaro , Uzair Khan , Aditya Jivoji , Franco Fummi , Marco Cristani

Grasp User Interfaces (grasp UIs) enable dual-tasking in XR by allowing interaction with digital content while holding physical objects. However, current grasp UI design practices face a fundamental challenge: existing approaches either…

人机交互 · 计算机科学 2025-09-11 Arthur Caetano , Yunhao Luo , Adwait Sharma , Misha Sra

Human motion prediction is an essential component for enabling closer human-robot collaboration. The task of accurately predicting human motion is non-trivial. It is compounded by the variability of human motion, both at a skeletal level…

机器人学 · 计算机科学 2021-07-02 Mohammad Samin Yasar , Tariq Iqbal

Domain experts often possess valuable physical insights that are overlooked in fully automated decision-making processes such as Bayesian optimisation. In this article we apply high-throughput (batch) Bayesian optimisation alongside…

机器学习 · 计算机科学 2023-12-06 Tom Savage , Ehecatl Antonio del Rio Chanona

Gait recognition, a rapidly advancing vision technology for person identification from a distance, has made significant strides in indoor settings. However, evidence suggests that existing methods often yield unsatisfactory results when…

计算机视觉与模式识别 · 计算机科学 2025-06-03 Chao Fan , Saihui Hou , Junhao Liang , Chuanfu Shen , Jingzhe Ma , Dongyang Jin , Yongzhen Huang , Shiqi Yu

Given a sequence of sets, where each set has a timestamp and contains an arbitrary number of elements, temporal sets prediction aims to predict the elements in the subsequent set. Previous studies for temporal sets prediction mainly focus…

机器学习 · 计算机科学 2023-08-29 Le Yu , Zihang Liu , Leilei Sun , Bowen Du , Chuanren Liu , Weifeng Lv

Human gait has been shown to provide crucial motion cues for various applications. Recognizing patterns in human gait has been widely adopted in various application areas such as security, virtual reality gaming, medical rehabilitation, and…

人机交互 · 计算机科学 2024-02-16 Chinmay Prakash Swami

In this work, we demonstrate robust walking in the bipedal robot Digit on uneven terrains by just learning a single linear policy. In particular, we propose a new control pipeline, wherein the high-level trajectory modulator shapes the…

机器人学 · 计算机科学 2021-10-06 Lokesh Krishna , Guillermo A. Castillo , Utkarsh A. Mishra , Ayonga Hereid , Shishir Kolathaya

In applications such as autonomous driving, it is important to understand, infer, and anticipate the intention and future behavior of pedestrians. This ability allows vehicles to avoid collisions and improve ride safety and quality. This…

机器人学 · 计算机科学 2019-09-16 Xiaoxiao Du , Ram Vasudevan , Matthew Johnson-Roberson

Aligning language models with human preferences through reinforcement learning from human feedback is crucial for their safe and effective deployment. The human preference is typically represented through comparison where one response is…

机器学习 · 计算机科学 2025-07-15 Hoang Anh Just , Ming Jin , Anit Sahu , Huy Phan , Ruoxi Jia

This paper investigates simultaneous preference and metric learning from a crowd of respondents. A set of items represented by $d$-dimensional feature vectors and paired comparisons of the form ``item $i$ is preferable to item $j$'' made by…

机器学习 · 统计学 2022-07-11 Gregory Canal , Blake Mason , Ramya Korlakai Vinayak , Robert Nowak

Multi-objective reinforcement learning (MORL) aims to find a set of high-performing and diverse policies that address trade-offs between multiple conflicting objectives. However, in practice, decision makers (DMs) often deploy only one or a…

神经与进化计算 · 计算机科学 2024-01-05 Ke Li , Han Guo