中文
相关论文

相关论文: B-RIGHT: Benchmark Re-evaluation for Integrity in …

200 篇论文

The software of robotic assistants needs to be verified, to ensure its safety and functional correctness. Testing in simulation allows a high degree of realism in the verification. However, generating tests that cover both interesting…

机器人学 · 计算机科学 2016-03-03 Dejanira Araiza-Illan , Tony Pipe , Kerstin Eder

As robots become more prevalent, the importance of the field of human-robot interaction (HRI) grows accordingly. As such, we should endeavor to employ the best statistical practices. Likert scales are commonly used metrics in HRI to measure…

人机交互 · 计算机科学 2020-02-03 Mariah L. Schrum , Michael Johnson , Muyleng Ghuy , Matthew C. Gombolay

Human-Object Interaction (HOI) detection plays a crucial role in activity understanding. Though significant progress has been made, interactiveness learning remains a challenging problem in HOI detection: existing methods usually generate…

计算机视觉与模式识别 · 计算机科学 2022-10-05 Xiaoqian Wu , Yong-Lu Li , Xinpeng Liu , Junyi Zhang , Yuzhe Wu , Cewu Lu

Enabling humanoid robots to physically interact with humans is a critical frontier, but progress is hindered by the scarcity of high-quality Human-Humanoid Interaction (HHoI) data. While leveraging abundant Human-Human Interaction (HHI)…

机器人学 · 计算机科学 2026-01-15 Wei-Jin Huang , Yue-Yi Zhang , Yi-Lin Wei , Zhi-Wei Xia , Juantao Tan , Yuan-Ming Li , Zhilin Zhao , Wei-Shi Zheng

Artificial intelligence (AI) systems are deployed as collaborators in human decision-making. Yet, evaluation practices focus primarily on model accuracy rather than whether human-AI teams are prepared to collaborate safely and effectively.…

人机交互 · 计算机科学 2026-03-20 Min Hun Lee

Existing neural information retrieval (IR) models have often been studied in homogeneous and narrow settings, which has considerably limited insights into their out-of-distribution (OOD) generalization capabilities. To address this, and to…

信息检索 · 计算机科学 2021-10-22 Nandan Thakur , Nils Reimers , Andreas Rücklé , Abhishek Srivastava , Iryna Gurevych

The stunning qualitative improvement of recent text-to-image models has led to their widespread attention and adoption. However, we lack a comprehensive quantitative understanding of their capabilities and risks. To fill this gap, we…

Large Language Models (LLMs) are advancing rapidly, yet the benchmarks used to measure this progress are becoming increasingly unreliable. Score inflation and selective reporting have eroded the authority of standard benchmarks, leaving the…

人工智能 · 计算机科学 2026-02-13 Longyuan Zhu , Hairan Hua , Linlin Miao , Bing Zhao

Open-vocabulary human-object interaction (HOI) detection is a step towards building scalable systems that generalize to unseen interactions in real-world scenarios and support grounded multimodal systems that reason about human-object…

计算机视觉与模式识别 · 计算机科学 2026-04-03 Maja Noack , Qinqian Lei , Taipeng Tian , Bihan Dong , Robby T. Tan , Yixin Chen , John Young , Saijun Zhang , Bo Wang

A comprehensive understanding of human-object interaction (HOI) requires detecting not only a small portion of predefined HOI concepts (or categories) but also other reasonable HOI concepts, while current approaches usually fail to explore…

计算机视觉与模式识别 · 计算机科学 2022-07-26 Zhi Hou , Baosheng Yu , Dacheng Tao

Hindsight relabeling is a powerful tool for overcoming sparsity in goal-conditioned reinforcement learning (GCRL), especially in certain domains such as navigation and locomotion. However, hindsight relabeling can struggle in object-centric…

机器学习 · 计算机科学 2025-05-07 Caleb Chuck , Fan Feng , Carl Qi , Chang Shi , Siddhant Agarwal , Amy Zhang , Scott Niekum

Aligning AI with human values is a pressing unsolved problem. To address the lack of quantitative metrics for value alignment, we propose EigenBench: a black-box method for comparatively benchmarking language models' values. Given an…

人工智能 · 计算机科学 2026-03-03 Jonathn Chang , Leonhard Piff , Suvadip Sana , Jasmine X. Li , Lionel Levine

Posture control and balance are basic requirements for a humanoid robot performing motor tasks like walking and interacting with the environment. For this reason, posture control is one of the elements taken into account when evaluating the…

机器人学 · 计算机科学 2021-10-28 Vittorio Lippi , Christoph Maurer , Thomas Mergner

Recent failures such as Google Gemini generating people of color in Nazi-era uniforms illustrate how AI outputs can be factually plausible yet socially harmful. AI models are increasingly evaluated for "fairness," yet existing benchmarks…

计算与语言 · 计算机科学 2025-10-01 Jen-tse Huang , Yuhang Yan , Linqi Liu , Yixin Wan , Wenxuan Wang , Kai-Wei Chang , Michael R. Lyu

Research on fairness, accountability, transparency and ethics of AI-based interventions in society has gained much-needed momentum in recent years. However it lacks an explicit alignment with a set of normative values and principles that…

人工智能 · 计算机科学 2022-10-07 Vinodkumar Prabhakaran , Margaret Mitchell , Timnit Gebru , Iason Gabriel

Zero-shot human-object interaction (HOI) detection remains a challenging task, particularly in generalizing to unseen actions. Existing methods address this challenge by tapping Vision-Language Models (VLMs) to access knowledge beyond the…

计算机视觉与模式识别 · 计算机科学 2025-08-05 Qinqian Lei , Bo Wang , Robby T. Tan

Recent advances in 3D human-aware generation have made significant progress. However, existing methods still struggle with generating novel Human Object Interaction (HOI) from text, particularly for open-set objects. We identify three main…

计算机视觉与模式识别 · 计算机科学 2025-06-02 Jinlu Zhang , Yixin Chen , Zan Wang , Jie Yang , Yizhou Wang , Siyuan Huang

The meteoric rise of AI, with its rapidly expanding market capitalization, presents both transformative opportunities and critical challenges. Chief among these is the urgent need for a new, unified paradigm for trustworthy evaluation, as…

Egocentric Human-Object Interaction (EHOI) analysis is crucial for industrial safety, yet the development of robust models is hindered by the scarcity of annotated domain-specific data. We address this challenge by introducing a data…

计算机视觉与模式识别 · 计算机科学 2026-01-15 Alfio Spoto , Rosario Leonardi , Francesco Ragusa , Giovanni Maria Farinella

As large language models (LLMs) evolve into autonomous agents capable of acting in open-ended environments, ensuring behavioral alignment with human values becomes a critical safety concern. Existing benchmarks, focused on static,…

计算与语言 · 计算机科学 2026-03-10 Weixiang Zhao , Haozhen Li , Yanyan Zhao , xuda zhi , Yongbo Huang , Hao He , Bing Qin , Ting Liu