English
Related papers

Related papers: A Conceptual Model of Intelligent Multimedia Data …

200 papers

This paper presents techniques to display 3D illuminations using Flying Light Specks, FLSs. Each FLS is a miniature (hundreds of micrometers) sized drone with one or more light sources to generate different colors and textures with…

Graphics · Computer Science 2022-07-19 Shahram Ghandeharizadeh

Unmanned Aerial Vehicles (UAVs) have moved beyond a platform for hobbyists to enable environmental monitoring, journalism, film industry, search and rescue, package delivery, and entertainment. This paper describes 3D displays using swarms…

Human-Computer Interaction · Computer Science 2021-11-09 Shahram Ghandeharizadeh

This study evaluates the accuracy of three different types of time-of-flight sensors to measure distance. We envision the possible use of these sensors to localize swarms of flying light specks (FLSs) to illuminate objects and avatars of a…

Multimedia · Computer Science 2023-08-22 Trung Phan , Hamed Alimohammadzadeh , Heather Culbertson , Shahram Ghandeharizadeh

Swarm-Merging, SwarMer, is a decentralized framework to localize Flying Light Specks (FLSs) to render 2D and 3D shapes. An FLS is a miniature sized drone equipped with one or more light sources to generate different colors and textures with…

Robotics · Computer Science 2023-12-11 Hamed Alimohammadzadeh , Shahram Ghandeharizadeh

Today's robotic laboratories for drones are housed in a large room. At times, they are the size of a warehouse. These spaces are typically equipped with permanent devices to localize the drones, e.g., Vicon Infrared cameras. Significant…

Foundation models (FM) have shown immense human-like capabilities for generating digital media. However, foundation models that can freely sense, interact, and actuate the physical domain is far from being realized. This is due to 1)…

Robotics · Computer Science 2025-03-07 Minghui Zhao , Junxi Xia , Kaiyuan Hou , Yanchen Liu , Stephen Xia , Xiaofan Jiang

I present a new FITS viewer designed to explore 3D spectral line data (in particular HI) and assist with visual source extraction and analysis. Using the artistic software Blender, FRELLED can visualise even large (~600^3 voxels) data sets…

Instrumentation and Methods for Astrophysics · Physics 2015-10-14 Rhys Taylor

We release two artificial datasets, Simulated Flying Shapes and Simulated Planar Manipulator that allow to test the learning ability of video processing systems. In particular, the dataset is meant as a tool which allows to easily assess…

Computer Vision and Pattern Recognition · Computer Science 2018-07-03 Fabio Ferreira , Jonas Rothfuss , Eren Erdal Aksoy , You Zhou , Tamim Asfour

Understanding how networks of neurons process information is one of the key challenges in modern neuroscience. A necessary step to achieve this goal is to be able to observe the dynamics of large populations of neurons over a large area of…

Image and Video Processing · Electrical Eng. & Systems 2022-03-09 Pingfan Song , Herman Verinaz Jadan , Carmel L. Howe , Amanda J. Foust , Pier Luigi Dragotti

Large Language Models (LLMs) have emerged as foundation models for IoT applications such as human activity recognition (HAR). However, directly applying high-frequency and multi-dimensional sensor data, such as eye-tracking data, leads to…

Human-Computer Interaction · Computer Science 2026-04-14 Jae Young Choi , Seon Gyeom Kim , Hyungjun Yoon , Taeckyung Lee , Donggun Lee , Jaeryung Chung , Jihyung Kil , Ryan Rossi , Sung-Ju Lee , Tak Yeon Lee

Despite a big leap forward in capability, multimodal large language models (MLLMs) tend to behave like a sloth in practical use, i.e., slow response and large latency. Recent efforts are devoted to building tiny MLLMs for better efficiency,…

Computer Vision and Pattern Recognition · Computer Science 2024-12-06 Bo Tong , Bokai Lai , Yiyi Zhou , Gen Luo , Yunhang Shen , Ke Li , Xiaoshuai Sun , Rongrong Ji

Modern datasets in neuroscience enable unprecedented inquiries into the relationship between complex behaviors and the activity of many simultaneously recorded neurons. While latent variable models can successfully extract low-dimensional…

Neurons and Cognition · Quantitative Biology 2024-12-03 Jaivardhan Kapoor , Auguste Schulz , Julius Vetter , Felix Pei , Richard Gao , Jakob H. Macke

The NASA Planetary Data System (PDS) hosts millions of images of planets, moons, and other bodies collected throughout many missions. The ever-expanding nature of data and user engagement demands an interpretable content classification…

Computer Vision and Pattern Recognition · Computer Science 2026-05-13 Bhavan Vasu , Steven Lu , Emily Dunkel , Kiri L. Wagstaff , Kevin Grimes , Michael McAuley

Limitations in processing capabilities and memory of today's computers make spiking neuron-based (human) whole-brain simulations inevitably characterized by a compromise between bio-plausibility and computational cost. It translates into…

Neurons and Cognition · Quantitative Biology 2020-07-17 Gianluca Susi , Pilar Garces , Alessandro Cristini , Emanuele Paracone , Mario Salerno , Fernando Maestu , Ernesto Pereda

This paper introduces Scene-LLM, a 3D-visual-language model that enhances embodied agents' abilities in interactive 3D indoor environments by integrating the reasoning strengths of Large Language Models (LLMs). Scene-LLM adopts a hybrid 3D…

Computer Vision and Pattern Recognition · Computer Science 2024-03-26 Rao Fu , Jingyu Liu , Xilun Chen , Yixin Nie , Wenhan Xiong

I present version 5.0 of FRELLED, the FITS Realtime Explorer of Low Latency in Every Dimension. This is a 3D data visualisation package for the popular Blender art software, designed to allow inspection of astronomical volumetric data sets…

Instrumentation and Methods for Astrophysics · Physics 2025-01-07 Rhys Taylor

The proliferation of resourceful mobile devices that store rich, multidimensional and privacy-sensitive user data motivate the design of federated learning (FL), a machine-learning (ML) paradigm that enables mobile devices to produce an ML…

Networking and Internet Architecture · Computer Science 2021-01-07 Christodoulos Pappas , Dimitris Chatzopoulos , Spyros Lalis , Manolis Vavalis

Multimodal Large Language Models (MLLMs) perform strong vision-language reasoning under standard conditions but fail in extreme illumination, where RGB inputs lose irrevocable structure and semantics. We propose Event-MLLM, an…

Computer Vision and Pattern Recognition · Computer Science 2026-03-31 Baoheng Zhang , Jiahui Liu , Gui Zhao , Weizhou Zhang , Yixuan Ma , Jun Jiang , Yingxian Chen , Wilton W. T. Fok , Xiaojuan Qi , Hayden Kwok-Hay So

The advantage of having a high-fidelity instrument simulation tool developed in tandem with novel instrumentation is having the ability to investigate, in isolation and in combination, the wide parameter space set by the instrument design.…

Instrumentation and Methods for Astrophysics · Physics 2021-10-29 Zackery Briesemeister , Steph Sallum , Andrew Skemer , Natasha Batalha

Few-shot semantic segmentation (FSS) aims to enable models to segment novel/unseen object classes using only a limited number of labeled examples. However, current FSS methods frequently struggle with generalization due to incomplete and…

Computer Vision and Pattern Recognition · Computer Science 2025-03-07 Amin Karimi , Charalambos Poullis
‹ Prev 1 2 3 10 Next ›