English
Related papers

Related papers: Combinatorial Hodge Theory in Simplicial Signal Pr…

200 papers

Traditional sentiment analysis has long been a unimodal task, relying solely on text. This approach overlooks non-verbal cues such as vocal tone and prosody that are essential for capturing true emotional intent. We introduce Dynamic…

Computation and Language · Computer Science 2025-09-30 Sadia Abdulhalim , Muaz Albaghdadi , Moshiur Farazi

These are lecture notes from my talks at the "Current Developments in Mathematics" conference (Harvard, 2006). They cover a variety of topics involving symplectic cohomology. In particular, a discussion of (algorithmic) classification…

Symplectic Geometry · Mathematics 2010-02-15 Paul Seidel

Combining complementary information from multiple modalities is intuitively appealing for improving the performance of learning-based approaches. However, it is challenging to fully leverage different modalities due to practical challenges…

Machine Learning · Statistics 2018-05-31 Kuan Liu , Yanen Li , Ning Xu , Prem Natarajan

This is an informal discussion on one of the basic problems in the theory of empirical processes, addressed in our preprint "Combinatorics of random processes and sections of convex bodies", which is available at ArXiV and from our web…

Functional Analysis · Mathematics 2007-05-23 Mark Rudelson , Roman Vershynin

Omnimodal large language models (OmniLLMs) jointly process audio and visual streams, but the resulting long multimodal token sequences make inference prohibitively expensive. Existing compression methods typically rely on fixed window…

Multimedia · Computer Science 2026-03-18 Bingzhou Li , Tao Huang

Federated learning (FL) and federated distillation (FD) are distributed learning paradigms that train UE models with enhanced privacy, each offering different trade-offs between noise robustness and learning speed. To mitigate their…

Machine Learning · Computer Science 2026-01-09 Yongjun Kim , Hyeongjun Park , Hwanjin Kim , Junil Choi

Digital audio effects are widely used by audio engineers to alter the acoustic and temporal qualities of audio data. However, these effects can have a large number of parameters which can make them difficult to learn for beginners and…

Machine Learning · Computer Science 2023-10-02 Kieran Grant

The adaptive estimation of coexisting temporal vertex (node) and edge signals on graphs is a critical task when a change in edge signals influences the temporal dynamics of the vertex signals. However, the current Graph Signal Processing…

Social and Information Networks · Computer Science 2024-10-24 Yi Yan , Tian Xie , Ercan E. Kuruoglu

Machine Listening, as usually formalized, attempts to perform a task that is, from our perspective, fundamentally human-performable, and performed by humans. Current automated models of Machine Listening vary from purely data-driven…

Audio and Speech Processing · Electrical Eng. & Systems 2023-02-27 Laurie M. Heller , Benjamin Elizalde , Bhiksha Raj , Soham Deshmukh

The Hodge Conjecture, posits a profound connection between the topology and algebraic geometry of complex algebraic varieties. It asserts that Hodge cycles, specific elements in the cohomology of a K\"ahler variety with rational properties,…

Algebraic Geometry · Mathematics 2025-08-05 Bita Hajebi , Pooya Hajebi

Some beautiful identities involving hook contents of Young diagrams have been found in the field of quantum information processing, along with a combinatorial proof. We here give a representation theoretic proof of these identities and a…

High Energy Physics - Theory · Physics 2019-05-02 Sanjaye Ramgoolam , Michal Sedlák

Dialogue state tracking plays a crucial role in extracting information in task-oriented dialogue systems. However, preceding research are limited to textual modalities, primarily due to the shortage of authentic human audio datasets. We…

Sound · Computer Science 2023-12-05 Jihyun Lee , Yejin Jeon , Wonjun Lee , Yunsu Kim , Gary Geunbae Lee

The paramount challenge in audio-driven One-shot Talking Head Animation (ADOS-THA) lies in capturing subtle imperceptible changes between adjacent video frames. Inherently, the temporal relationship of adjacent audio clips is highly…

Computer Vision and Pattern Recognition · Computer Science 2025-04-09 Zhihua Xu , Tianshui Chen , Zhijing Yang , Siyuan Peng , Keze Wang , Liang Lin

This paper explores second-order optimization methods in Federated Learning (FL), addressing the critical challenges of slow convergence and the excessive communication rounds required to achieve optimal performance from the global model.…

Machine Learning · Computer Science 2025-05-30 Mrinmay Sen , Sidhant R Nair , C Krishna Mohan

This paper reviews some recent applications of the theory of the compensated convex transforms or of the proximity hull as developed by the authors to image processing and shape interrogation with special attention given to the Hausdorff…

Image and Video Processing · Electrical Eng. & Systems 2020-10-13 Antonio Orlando , Elaine Crooks , Kewei Zhang

Speech signals are inherently complex as they encompass both global acoustic characteristics and local semantic information. However, in the task of target speech extraction, certain elements of global and local semantic information in the…

Sound · Computer Science 2024-08-27 Zhaoxi Mu , Xinyu Yang , Sining Sun , Qing Yang

Sound event detection (SED) is the task of identifying sound events along with their onset and offset times. A recent, convolutional neural networks based SED method, proposed the usage of depthwise separable (DWS) and time-dilated…

Sound · Computer Science 2020-07-13 Konstantinos Drossos , Stylianos I. Mimilakis , Tuomas Virtanen

In this letter, we propose a novel computationally efficient coupled dictionary learning method that enforces pairwise correlation between the atoms of dictionaries learned to represent the underlying feature spaces of two different…

Machine Learning · Computer Science 2021-09-29 F. G. Veshki , S. A. Vorobyov

Compared to the error diffusion, dot diffusion provides an additional pixel-level parallelism for digital halftoning. However, even though its periodic and blocking artifacts had been eased by previous works, it was still far from…

Multimedia · Computer Science 2015-10-28 Yun-Fu Liu , Jing-Ming Guo

In federated distributed learning, the goal is to optimize a global training objective defined over distributed devices, where the data shard at each device is sampled from a possibly different distribution (a.k.a., heterogeneous or non…

Machine Learning · Computer Science 2019-12-10 Farzin Haddadpour , Mehrdad Mahdavi
‹ Prev 1 4 5 6 7 8 10 Next ›