中文
相关论文

相关论文: Framework for 2D Ad placements in LinearTV

200 篇论文

Robotic manipulation requires understanding both the 3D spatial structure of the environment and its temporal evolution, yet most existing policies overlook one or both. They typically rely on 2D visual observations and backbones pretrained…

Reality TV shows that follow people in their day-to-day lives are not a new concept. However, the traditional methods used in the industry require a lot of manual labour and need the presence of at least one physical camera man. Because of…

计算机视觉与模式识别 · 计算机科学 2020-07-10 Timothy Callemein , Wiebe Van Ranst , Toon Goedemé

A major focus of current research on place recognition is visual localization for autonomous driving. In this scenario, as cameras will be operating continuously, it is realistic to expect videos as an input to visual localization…

计算机视觉与模式识别 · 计算机科学 2020-11-05 Anh-Dzung Doan , Yasir Latif , Tat-Jun Chin , Yu Liu , Shin-Fang Ch'ng , Thanh-Toan Do , Ian Reid

Images incorporate a wealth of information from a robot's surroundings. With the widespread availability of compact cameras, visual information has become increasingly popular for addressing the localisation problem, which is then termed as…

计算机视觉与模式识别 · 计算机科学 2023-05-11 Mihnea-Alexandru Tomita , Bruno Ferrarini , Michael Milford , Klaus McDonald-Maier , Shoaib Ehsan

In this paper, a new type of 3D bin packing problem (BPP) is proposed, in which a number of cuboid-shaped items must be put into a bin one by one orthogonally. The objective is to find a way to place these items that can minimize the…

人工智能 · 计算机科学 2017-08-22 Haoyuan Hu , Xiaodong Zhang , Xiaowei Yan , Longfei Wang , Yinghui Xu

We propose a generative framework which takes on the video frame interpolation problem. Our framework, which we call Deep Locally Linear Embedding (DeepLLE), is powered by a deep convolutional neural network (CNN) while it can be used…

计算机视觉与模式识别 · 计算机科学 2018-07-05 Anh-Duc Nguyen , Woojae Kim , Jongyoo Kim , Sanghoon Lee

Although Multimodal Large Language Models (MLLMs) excel at various image-related tasks, they encounter challenges in precisely aligning coordinates with spatial information within images, particularly in position-aware tasks such as visual…

计算机视觉与模式识别 · 计算机科学 2025-07-17 Wei Tang , Yanpeng Sun , Qinying Gu , Zechao Li

E-commerce platforms usually present an ordered list, mixed with several organic items and an advertisement, in response to each user's page view request. This list, the outcome of ad auction and allocation processes, directly impacts the…

计算机科学与博弈论 · 计算机科学 2024-04-12 Xuejian Li , Ze Wang , Bingqi Zhu , Fei He , Yongkang Wang , Xingxing Wang

Real-Time Bidding (RTB) is an important paradigm in display advertising, where advertisers utilize extended information and algorithms served by Demand Side Platforms (DSPs) to improve advertising performance. A common problem for DSPs is…

计算机科学与博弈论 · 计算机科学 2019-05-30 Xun Yang , Yasong Li , Hao Wang , Di Wu , Qing Tan , Jian Xu , Kun Gai

In this work we study indoor scene object placement. Given a 3D indoor scene and an object, the task is to predict placement locations within the scene. Empirical observations of data-driven approaches to the problem show their tendency to…

图形学 · 计算机科学 2026-05-05 Adrian Chang , Kai Wang , Yuanbo Li , Manolis Savva , Angel X. Chang , Daniel Ritchie

With the advent of faster internet services and growth of multimedia content, we observe a massive growth in the number of online videos. The users generate these video contents at an unprecedented rate, owing to the use of smart-phones and…

计算机视觉与模式识别 · 计算机科学 2019-04-30 Soumyabrata Dev , Murhaf Hossari , Matthew Nicholson , Killian McCabe , Atul Nautiyal , Clare Conran , Jian Tang , Wei Xu , François Pitié

We address the bin packing problem (BPP), which aims to maximize bin utilization when packing a variety of items. The offline problem, where the complete information about the item set and their sizes is known in advance, is proven to be…

机器人学 · 计算机科学 2025-10-16 Beomjoon Lee , Changjoo Nam

Photopolymerization-based additive manufacturing enables cost-effective, high-speed fabrication of complex 3D structures but is constrained by a trade-off between resolution and printing speed. Single-photon polymerization ensures rapid…

光学 · 物理学 2026-01-21 Buse Unlu , Felix Wechsler , Ye Pu , Christophe Moser

We propose an auction for online advertising where each ad occupies either one square or two horizontally-adjacent squares of a grid of squares. Our primary application are ads for products shown on retail websites such as Instacart or…

计算机科学与博弈论 · 计算机科学 2022-07-12 Jonathan Gu , David Pal , Kevin Ryan

Vision Transformers (ViTs) have demonstrated state-ofthe-art performance in several benchmarks, yet their high computational costs hinders their practical deployment. Patch Pruning offers significant savings, but existing approaches…

计算机视觉与模式识别 · 计算机科学 2026-04-02 Patrick Glandorf , Thomas Norrenbrock , Bodo Rosenhahn

With expansion of the video advertising market, research to predict the effects of video advertising is getting more attention. Although effect prediction of image advertising has been explored a lot, prediction for video advertising is…

计算机视觉与模式识别 · 计算机科学 2020-12-23 Jun Ikeda , Hiroyuki Seshime , Xueting Wang , Toshihiko Yamasaki

Video inpainting aims to fill spatio-temporal holes with plausible content in a video. Despite tremendous progress of deep neural networks for image inpainting, it is challenging to extend these methods to the video domain due to the…

计算机视觉与模式识别 · 计算机科学 2019-05-07 Dahun Kim , Sanghyun Woo , Joon-Young Lee , In So Kweon

Localizing an object accurately with respect to a robot is a key step for autonomous robotic manipulation. In this work, we propose to tackle this task knowing only 3D models of the robot and object in the particular case where the scene is…

计算机视觉与模式识别 · 计算机科学 2019-02-08 Vianney Loing , Renaud Marlet , Mathieu Aubry

In this communication revolution era, visible light communication (VLC) is the optimum efficacious answer to the increased request for high-speed data transmission with reduced cost, besides the illumination. This technology is considered…

信号处理 · 电气工程与系统科学 2022-09-16 Vailet Hikmat Faraj Al Khattat , Siti Barirah Ahmad Anas , Abdu Saif

Ad-hoc Video Search (AVS) involves using a textual query to search for multiple relevant videos in a large collection of unlabeled short videos. The main challenge of AVS is the visual diversity of relevant videos. A simple query such as…

计算机视觉与模式识别 · 计算机科学 2025-08-05 Fan Hu , Zijie Xin , Xirong Li