中文
相关论文

相关论文: The Algonauts Project 2023 Challenge: UARK-UAlbany…

200 篇论文

The SoccerNet 2023 challenges were the third annual video understanding challenges organized by the SoccerNet team. For this third edition, the challenges were composed of seven vision-based tasks split into three main themes. The first…

计算机视觉与模式识别 · 计算机科学 2025-02-18 Anthony Cioppa , Silvio Giancola , Vladimir Somers , Floriane Magera , Xin Zhou , Hassan Mkhallati , Adrien Deliège , Jan Held , Carlos Hinojosa , Amir M. Mansourian , Pierre Miralles , Olivier Barnich , Christophe De Vleeschouwer , Alexandre Alahi , Bernard Ghanem , Marc Van Droogenbroeck , Abdullah Kamal , Adrien Maglo , Albert Clapés , Amr Abdelaziz , Artur Xarles , Astrid Orcesi , Atom Scott , Bin Liu , Byoungkwon Lim , Chen Chen , Fabian Deuser , Feng Yan , Fufu Yu , Gal Shitrit , Guanshuo Wang , Gyusik Choi , Hankyul Kim , Hao Guo , Hasby Fahrudin , Hidenari Koguchi , Håkan Ardö , Ibrahim Salah , Ido Yerushalmy , Iftikar Muhammad , Ikuma Uchida , Ishay Be'ery , Jaonary Rabarisoa , Jeongae Lee , Jiajun Fu , Jianqin Yin , Jinghang Xu , Jongho Nang , Julien Denize , Junjie Li , Junpei Zhang , Juntae Kim , Kamil Synowiec , Kenji Kobayashi , Kexin Zhang , Konrad Habel , Kota Nakajima , Licheng Jiao , Lin Ma , Lizhi Wang , Luping Wang , Menglong Li , Mengying Zhou , Mohamed Nasr , Mohamed Abdelwahed , Mykola Liashuha , Nikolay Falaleev , Norbert Oswald , Qiong Jia , Quoc-Cuong Pham , Ran Song , Romain Hérault , Rui Peng , Ruilong Chen , Ruixuan Liu , Ruslan Baikulov , Ryuto Fukushima , Sergio Escalera , Seungcheon Lee , Shimin Chen , Shouhong Ding , Taiga Someya , Thomas B. Moeslund , Tianjiao Li , Wei Shen , Wei Zhang , Wei Li , Wei Dai , Weixin Luo , Wending Zhao , Wenjie Zhang , Xinquan Yang , Yanbiao Ma , Yeeun Joo , Yingsen Zeng , Yiyang Gan , Yongqiang Zhu , Yujie Zhong , Zheng Ruan , Zhiheng Li , Zhijian Huang , Ziyu Meng

In this paper, we address the novel, highly challenging problem of estimating the layout of a complex urban driving scenario. Given a single color image captured from a driving platform, we aim to predict the bird's-eye view layout of the…

计算机视觉与模式识别 · 计算机科学 2020-02-21 Kaustubh Mani , Swapnil Daga , Shubhika Garg , N. Sai Shankar , Krishna Murthy Jatavallabhula , K. Madhava Krishna

Decoding human visual neural representations is a challenging task with great scientific significance in revealing vision-processing mechanisms and developing brain-like intelligent machines. Most existing methods are difficult to…

计算机视觉与模式识别 · 计算机科学 2023-03-31 Changde Du , Kaicheng Fu , Jinpeng Li , Huiguang He

In this report, we describe the technical details of our approach for the Ego4D Long-Term Action Anticipation Challenge 2023. The aim of this task is to predict a sequence of future actions that will take place at an arbitrary time or…

计算机视觉与模式识别 · 计算机科学 2023-07-06 Tatsuya Ishibashi , Kosuke Ono , Noriyuki Kugo , Yuji Sato

We revisit the planning problem in the blocks world, and we implement a known heuristic for this task. Importantly, our implementation is biologically plausible, in the sense that it is carried out exclusively through the spiking of…

In cognitive decoding, researchers aim to characterize a brain region's representations by identifying the cognitive states (e.g., accepting/rejecting a gamble) that can be identified from the region's activity. Deep learning (DL) methods…

机器学习 · 计算机科学 2021-08-17 Armin W. Thomas , Christopher Ré , Russell A. Poldrack

This work presents a novel deep-learning-based pipeline for the inverse problem of image deblurring, leveraging augmentation and pre-training with synthetic data. Our results build on our winning submission to the recent Helsinki Deblur…

计算机视觉与模式识别 · 计算机科学 2023-08-08 Theophil Trippe , Martin Genzel , Jan Macdonald , Maximilian März

Despite the impressive progress on understanding and generating images shown by the recent unified architectures, the integration of 3D tasks remains challenging and largely unexplored. In this paper, we introduce UniUGG, the first unified…

计算机视觉与模式识别 · 计算机科学 2026-03-10 Yueming Xu , Jiahui Zhang , Ze Huang , Yurui Chen , Yanpeng Zhou , Zhenyu Chen , Yu-Jie Yuan , Pengxiang Xia , Guowei Huang , Xinyue Cai , Zhongang Qi , Xingyue Quan , Jianye Hao , Hang Xu , Li Zhang

Aerial Vision-and-Language Navigation (VLN) aims to enable unmanned aerial vehicles (UAVs) to interpret natural language instructions and navigate complex urban environments using onboard visual observation. This task holds promise for…

计算机视觉与模式识别 · 计算机科学 2026-04-16 Huilin Xu , Zhuoyang Liu , Yixiang Luomei , Feng Xu

Each year, thousands of people learn new visual categorization tasks -- radiologists learn to recognize tumors, birdwatchers learn to distinguish similar species, and crowd workers learn how to annotate valuable data for applications like…

计算机视觉与模式识别 · 计算机科学 2022-07-25 Neehar Kondapaneni , Pietro Perona , Oisin Mac Aodha

In this paper, we investigated how to build a high-performance vision encoding model to predict brain activity as part of our participation in the Algonauts Project 2023 Challenge. The challenge provided brain activity recorded by…

神经元与认知 · 定量生物学 2023-08-02 Takuya Matsuyama , Kota S Sasaki , Shinji Nishimoto

Human parsing aims to partition humans in image or video into multiple pixel-level semantic parts. In the last decade, it has gained significantly increased interest in the computer vision community and has been utilized in a broad range of…

计算机视觉与模式识别 · 计算机科学 2024-03-15 Lu Yang , Wenhe Jia , Shan Li , Qing Song

Reconstructing seeing images from fMRI recordings is an absorbing research area in neuroscience and provides a potential brain-reading technology. The challenge lies in that visual encoding in brain is highly complex and not fully revealed.…

神经与进化计算 · 计算机科学 2021-01-29 Tao Fang , Yu Qi , Gang Pan

Deep learning is leading to major advances in the realm of brain decoding from functional Magnetic Resonance Imaging (fMRI). However, the large inter-subject variability in brain characteristics has limited most studies to train models on…

Image-to-fMRI encoding is important for both neuroscience research and practical applications. However, such "Brain-Encoders" have been typically trained per-subject and per fMRI-dataset, thus restricted to very limited training data. In…

计算机视觉与模式识别 · 计算机科学 2026-05-26 Roman Beliy , Navve Wasserman , Amit Zalcher , Michal Irani

Decoding visual information from time-resolved brain recordings, such as EEG and MEG, plays a pivotal role in real-time brain-computer interfaces. However, existing approaches primarily focus on direct brain-image feature alignment and are…

人机交互 · 计算机科学 2025-11-12 Chengjian Xu , Yonghao Song , Zelin Liao , Haochuan Zhang , Qiong Wang , Qingqing Zheng

In dyadic interactions, humans communicate their intentions and state of mind using verbal and non-verbal cues, where multiple different facial reactions might be appropriate in response to a specific speaker behaviour. Then, how to develop…

Modern video understanding systems excel at tasks such as scene classification, object detection, and short video retrieval. However, as video analysis becomes increasingly central to real-world applications, there is a growing need for…

人工智能 · 计算机科学 2025-05-21 Sahil Shah , Harsh Goel , Sai Shankar Narasimhan , Minkyu Choi , S P Sharan , Oguzhan Akcin , Sandeep Chinchali

Recent work has demonstrated that complex visual stimuli can be decoded from human brain activity using deep generative models, offering new ways to probe how the brain represents real-world scenes. However, many existing approaches first…

计算机视觉与模式识别 · 计算机科学 2026-03-03 Pinyuan Feng , Hossein Adeli , Wenxuan Guo , Fan Cheng , Ethan Hwang , Nikolaus Kriegeskorte

In recent years, numerous tasks have been proposed to encourage model to develop specified capability in understanding audio-visual scene, primarily categorized into temporal localization, spatial localization, spatio-temporal reasoning,…

计算机视觉与模式识别 · 计算机科学 2025-03-18 Henghui Du , Guangyao Li , Chang Zhou , Chunjie Zhang , Alan Zhao , Di Hu