English
Related papers

Related papers: The Advantage of a Multi-User Mode

200 papers

Transformers are powerful visual learners, in large part due to their conspicuous lack of manually-specified priors. This flexibility can be problematic in tasks that involve multiple-view geometry, due to the near-infinite possible…

Computer Vision and Pattern Recognition · Computer Science 2023-04-04 Yash Bhalgat , Joao F. Henriques , Andrew Zisserman

Inter-operator spectrum sharing in millimeter-wave bands has the potential of substantially increasing the spectrum utilization and providing a larger bandwidth to individual user equipment at the expense of increasing inter-operator…

Information Theory · Computer Science 2020-03-20 Hossein S. Ghadikolaei , Hadi Ghauch , Gabor Fodor , Mikael Skoglund , Carlo Fischione

Vision-language models have been widely explored across a wide range of tasks and achieve satisfactory performance. However, it's under-explored how to consolidate entity understanding through a varying number of images and to align it with…

Computer Vision and Pattern Recognition · Computer Science 2023-12-29 Wenyi Wu , Qi Li , Wenliang Zhong , Junzhou Huang

Fonts are ubiquitous across documents and come in a variety of styles. They are either represented in a native vector format or rasterized to produce fixed resolution images. In the first case, the non-standard representation prevents…

Computer Vision and Pattern Recognition · Computer Science 2022-01-11 Pradyumna Reddy , Zhifei Zhang , Matthew Fisher , Hailin Jin , Zhaowen Wang , Niloy J. Mitra

An observer-based Hamiltonian identification algorithm for quantum systems is proposed. For the 2-level case an exponential convergence result based on averaging arguments and some relevant transformations is provided. The convergence for…

Mathematical Physics · Physics 2007-05-23 Mazyar Mirrahimi , Pierre Rouchon

The calculus of variations applied to the image processing requires some numerical models able to perform the variations of images and the extremization of appropriate actions. To produce the variations of images, there are several…

Computer Vision and Pattern Recognition · Computer Science 2012-01-18 Amelia Carolina Sparavigna

Human-robot interaction benefits greatly from multimodal sensor inputs as they enable increased robustness and generalization accuracy. Despite this observation, few HRI methods are capable of efficiently performing inference for multimodal…

Robotics · Computer Science 2019-08-15 Joseph Campbell , Simon Stepputtis , Heni Ben Amor

Humans possess multimodal literacy, allowing them to actively integrate information from various modalities to form reasoning. Faced with challenges like lexical ambiguity in text, we supplement this with other modalities, such as thumbnail…

Computer Vision and Pattern Recognition · Computer Science 2024-10-24 Jiwan Chung , Seungwon Lim , Jaehyun Jeon , Seungbeen Lee , Youngjae Yu

Interaction with the physical environment and different users is essential to foster a collaborative experience. For this, we propose an interaction based on a central point represented by an Augmented Reality marker in which several users…

Human-Computer Interaction · Computer Science 2023-01-09 Bianca Marques , Rui Nóbrega , Carmen Morgado

We give an alternative derivation for the explicit formula of the effective Hamiltonian describing the evolution of the quantum state of any number of photons entering a linear optics multiport. The description is based on the effective…

Quantum Physics · Physics 2025-01-17 Juan Carlos Garcia-Escartin , Vicent Gimeno , Julio José Moyano-Fernández

This paper is a compact overview of the heuristic approach to the recently elaborated octonionic binocular mobilevision.

adap-org · Physics 2008-02-03 Denis V. Juriev

Various heuristic objectives for modeling hand-object interaction have been proposed in past work. However, due to the lack of a cohesive framework, these objectives often possess a narrow scope of applicability and are limited by their…

Computer Vision and Pattern Recognition · Computer Science 2023-12-27 Shutong Zhang , Yi-Ling Qiao , Guanglei Zhu , Eric Heiden , Dylan Turpin , Jingzhou Liu , Ming Lin , Miles Macklin , Animesh Garg

There is a multitude of interpretations of quantum mechanics, but foundational principles are lacking. Relational quantum mechanics views the observer as a physical system, which allows for an unambiguous interpretation as all axioms are…

Quantum Physics · Physics 2025-06-16 Dorian Daimer , Susanne Still

The characteristics of feature selection, nonlinear combination and multi-task auxiliary learning mechanism of the human visual perception system play an important role in real-world scenarios, but the research of image fusion theory based…

Computer Vision and Pattern Recognition · Computer Science 2020-06-23 Aiqing Fang , Xinbo Zhao , Jiaqi Yang , Yanning Zhang

We study the achievable rates by a single user in multibeam satellite scenarios. We show alternatives to the conventional symbol-by-symbol detection applied at user terminals. Single user detection is known to suffer from strong degradation…

Information Theory · Computer Science 2015-03-16 Giulio Colavolpe , Andrea Modenini , Amina Piemontese , Alessandro Ugolini

Understanding user interface (UI) functionality is a useful yet challenging task for both machines and people. In this paper, we investigate a machine learning approach for screen correspondence, which allows reasoning about UIs by mapping…

Human-Computer Interaction · Computer Science 2023-01-23 Jason Wu , Amanda Swearngin , Xiaoyi Zhang , Jeffrey Nichols , Jeffrey P. Bigham

Textbooks in applied mathematics often use graphs to explain the meaning of formulae, even though their benefit is still not fully explored. To test processes underlying this assumed multimedia effect we collected performance scores, eye…

Physics Education · Physics 2017-06-14 M. Ogren , M. Nystrom , H. Jarodzka

The recent popularity of text-to-image diffusion models (DM) can largely be attributed to the intuitive interface they provide to users. The intended generation can be expressed in natural language, with the model producing faithful…

Estimating 3D human poses from monocular videos is a challenging task due to depth ambiguity and self-occlusion. Most existing works attempt to solve both issues by exploiting spatial and temporal relationships. However, those works ignore…

Computer Vision and Pattern Recognition · Computer Science 2022-06-29 Wenhao Li , Hong Liu , Hao Tang , Pichao Wang , Luc Van Gool

We propose a scheme allowing a conditional implementation of suitably truncated general single- or multi-mode operators acting on states of traveling optical signal modes. The scheme solely relies on single-photon and coherent states and…

Quantum Physics · Physics 2016-04-19 J. Clausen , L. Knoell , D. -G. Welsch
‹ Prev 1 8 9 10 Next ›