English
Related papers

Related papers: The 3D-DCT transform: didactic experiment and poss…

200 papers

Recent years have witnessed the increasing popularity of learning based methods to enhance the color and tone of photos. However, many existing photo enhancement methods either deliver unsatisfactory results or consume too much…

Image and Video Processing · Electrical Eng. & Systems 2020-10-01 Hui Zeng , Jianrui Cai , Lida Li , Zisheng Cao , Lei Zhang

3D image segmentation is a recent and crucial step in many medical analysis and recognition schemes. In fact, it represents a relevant research subject and a fundamental challenge due to its importance and influence. This paper provides a…

Image and Video Processing · Electrical Eng. & Systems 2022-07-22 Omar Boudraa

This study aims to improve photon counting CT (PCCT) image resolution using denoising diffusion probabilistic models (DDPM). Although DDPMs have shown superior performance when applied to various computer vision tasks, their effectiveness…

Computer Vision and Pattern Recognition · Computer Science 2024-08-29 Chuang Niu , Christopher Wiedeman , Mengzhou Li , Jonathan S Maltz , Ge Wang

We study the problem of attribute compression for large-scale unstructured 3D point clouds. Through an in-depth exploration of the relationships between different encoding steps and different attribute channels, we introduce a deep…

Computer Vision and Pattern Recognition · Computer Science 2022-04-12 Guangchi Fang , Qingyong Hu , Hanyun Wang , Yiling Xu , Yulan Guo

Deep learning methods, in particular, trained Convolutional Neural Networks (CNN) have recently been shown to produce compelling results for single image Super-Resolution (SR). Invariably, a CNN is learned to map the Low Resolution (LR)…

Computer Vision and Pattern Recognition · Computer Science 2019-09-04 Tiantong Guo , Hojjat S. Mousavi , Vishal Monga

The discrete cosine transform is a valuable tool in analysis of data on undirected rectangular grids, like images. In this paper it is shown how one can define an analogue of the discrete cosine transform on triangles. This is done by…

Numerical Analysis · Mathematics 2018-11-12 Bastian Seifert , Knut Hüper

With the thriving of deep learning, 3D Convolutional Neural Networks have become a popular choice in volumetric image analysis due to their impressive 3D contexts mining ability. However, the 3D convolutional kernels will introduce a…

Computer Vision and Pattern Recognition · Computer Science 2019-05-22 Lei Qu , Changfeng Wu , Liang Zou

The unification of understanding and generation within a single multi-modal large model (MLLM) remains one significant challenge, largely due to the dichotomy between continuous and discrete visual tokenizations. Continuous tokenizer (CT)…

Computer Vision and Pattern Recognition · Computer Science 2025-11-04 Yizhu Chen , Chen Ju , Zhicheng Wang , Shuai Xiao , Xu Chen , Jinsong Lan , Xiaoyong Zhu , Ying Chen

We present JointDiT, a diffusion transformer that models the joint distribution of RGB and depth. By leveraging the architectural benefit and outstanding image prior of the state-of-the-art diffusion transformer, JointDiT not only generates…

Computer Vision and Pattern Recognition · Computer Science 2025-08-06 Kwon Byung-Ki , Qi Dai , Lee Hyoseok , Chong Luo , Tae-Hyun Oh

Recent Diffusion Transformers (DiTs) have shown impressive capabilities in generating high-quality single-modality content, including images, videos, and audio. However, it is still under-explored whether the transformer-based diffuser can…

Computer Vision and Pattern Recognition · Computer Science 2024-06-13 Kai Wang , Shijian Deng , Jing Shi , Dimitrios Hatzinakos , Yapeng Tian

4D video control is essential in video generation as it enables the use of sophisticated lens techniques, such as multi-camera shooting and dolly zoom, which are currently unsupported by existing methods. Training a video Diffusion…

Computer Vision and Pattern Recognition · Computer Science 2025-01-07 Weikang Bian , Zhaoyang Huang , Xiaoyu Shi , Yijin Li , Fu-Yun Wang , Hongsheng Li

Multimodal transformer exhibits high capacity and flexibility to align image and text for visual grounding. However, the existing encoder-only grounding framework (e.g., TransVG) suffers from heavy computation due to the self-attention…

Computer Vision and Pattern Recognition · Computer Science 2023-10-27 Fengyuan Shi , Ruopeng Gao , Weilin Huang , Limin Wang

Since 2016, deep learning (DL) has advanced tomographic imaging with remarkable successes, especially in low-dose computed tomography (LDCT) imaging. Despite being driven by big data, the LDCT denoising and pure end-to-end reconstruction…

Image and Video Processing · Electrical Eng. & Systems 2023-03-28 Wenjun Xia , Hongming Shan , Ge Wang , Yi Zhang

Intrinsic image decomposition is fundamental for visual understanding, as RGB images entangle material properties, illumination, and view-dependent effects. Recent diffusion-based methods have achieved strong results for single-view…

Computer Vision and Pattern Recognition · Computer Science 2026-01-01 Kang Du , Yirui Guan , Zeyu Wang

Transformers are very powerful tools for a variety of tasks across domains, from text generation to image captioning. However, transformers require substantial amounts of training data, which is often a challenge in biomedical settings,…

Computer Vision and Pattern Recognition · Computer Science 2023-07-04 Andrew Kean Gao

3D dense captioning aims to describe individual objects by natural language in 3D scenes, where 3D scenes are usually represented as RGB-D scans or point clouds. However, only exploiting single modal information, e.g., point cloud, previous…

Computer Vision and Pattern Recognition · Computer Science 2022-04-07 Zhihao Yuan , Xu Yan , Yinghong Liao , Yao Guo , Guanbin Li , Zhen Li , Shuguang Cui

We propose the deep progressive image compression using trit-planes (DPICT) algorithm, which is the first learning-based codec supporting fine granular scalability (FGS). First, we transform an image into a latent tensor using an analysis…

Image and Video Processing · Electrical Eng. & Systems 2022-05-09 Jae-Han Lee , Seungmin Jeon , Kwang Pyo Choi , Youngo Park , Chang-Su Kim

JPEG is a popular image compression method widely used by individuals, data center, cloud storage and network filesystems. However, most recent progress on image compression mainly focuses on uncompressed images while ignoring trillions of…

Image and Video Processing · Electrical Eng. & Systems 2022-03-31 Lina Guo , Xinjie Shi , Dailan He , Yuanyuan Wang , Rui Ma , Hongwei Qin , Yan Wang

Image-text retrieval is a central problem for understanding the semantic relationship between vision and language, and serves as the basis for various visual and language tasks. Most previous works either simply learn coarse-grained…

Computer Vision and Pattern Recognition · Computer Science 2023-07-19 Chong Liu , Yuqi Zhang , Hongsong Wang , Weihua Chen , Fan Wang , Yan Huang , Yi-Dong Shen , Liang Wang

Applications of the three-dimensional transformation for rotating coordinate systems to quantum mechanics, general theory relativity and optics are considered.

General Physics · Physics 2019-01-08 B. V. Gisin
‹ Prev 1 8 9 10 Next ›