中文
相关论文

相关论文: Apollo: An Interactive Environment for Generating …

200 篇论文

Despite the rapid progress of text-to-image (T2I) models, generating images that accurately reflect complex compositional prompts (covering attribute bindings, object relationships, counting) still remains challenging. To address this, we…

计算机视觉与模式识别 · 计算机科学 2026-05-28 Zhuohan Liu , Wujian Peng , Yitong Chen , Zuxuan Wu

Conceptual blending is a powerful tool for computational creativity where, for example, the properties of two harmonic spaces may be combined in a consistent manner to produce a novel harmonic space. However, deciding about the importance…

In this paper, we discuss the formalized approach for generating and estimating symbols (and alphabets), which can be communicated by the wide range of non-verbal means based on specific user requirements (medium, priorities, type of…

人机交互 · 计算机科学 2017-12-13 Serhii Hamotskyi , Sergii Stirenko , Yuri Gordienko , Anis Rojbi

This paper introduces the ACCompanion, an expressive accompaniment system. Similarly to a musician who accompanies a soloist playing a given musical piece, our system can produce a human-like rendition of the accompaniment part that follows…

We present JASCO, a temporally controlled text-to-music generation model utilizing both symbolic and audio-based conditions. JASCO can generate high-quality music samples conditioned on global text descriptions along with fine-grained local…

声音 · 计算机科学 2024-06-18 Or Tal , Alon Ziv , Itai Gat , Felix Kreuk , Yossi Adi

The creation of long melody sequences requires effective expression of coherent musical structure. However, there is no clear representation of musical structure. Recent works on music generation have suggested various approaches to deal…

声音 · 计算机科学 2021-11-04 Yi Zou , Pei Zou , Yi Zhao , Kaixiang Zhang , Ran Zhang , Xiaorui Wang

Music evokes emotion in many people. We introduce a novel way to manipulate the emotional content of a song using AI tools. Our goal is to achieve the desired emotion while leaving the original melody as intact as possible. For this, we…

声音 · 计算机科学 2024-06-14 Adel N. Abdalla , Jared Osborne , Razvan Andonie

Time-frequency (TF) representations in audio synthesis have been increasingly modeled with real-valued networks. However, overlooking the complex-valued nature of TF representations can result in suboptimal performance and require…

音频与语音处理 · 电气工程与系统科学 2022-06-22 Yongtao Wu , Grigorios G Chrysos , Volkan Cevher

By capturing statistical patterns in large corpora, machine learning has enabled significant advances in natural language processing, including in machine translation, question answering, and sentiment analysis. However, for agents to…

人工智能 · 计算机科学 2018-07-25 Igor Mordatch , Pieter Abbeel

Recent advancements in large language models (LLMs) have enabled a wide range of natural language processing (NLP) tasks to be performed through simple prompt-based interactions. Consequently, several approaches have been proposed to…

计算与语言 · 计算机科学 2025-08-14 Artem Chernodub , Aman Saini , Yejin Huh , Vivek Kulkarni , Vipul Raheja

We present Speakerly, a new real-time voice-based writing assistance system that helps users with text composition across various use cases such as emails, instant messages, and notes. The user can interact with the system through…

In this paper, a system to build music in an intuitive and accessible way, with Lego bricks, is presented. The system makes use of the new powerful and cheap possibilities that technology offers for making old things in a new way. The…

人机交互 · 计算机科学 2024-11-21 Ana M. Barbancho , Lorenzo J. Tardon , Isabel Barbancho

Producing original and arranging existing musical outcomes is an art that takes years of learning and practice to master. Yet, despite the constant advances in the field of AI-powered musical creativity, production of quality musical…

声音 · 计算机科学 2023-04-04 Ivan S. Maksymov

In addressing the challenge of interpretability and generalizability of artificial music intelligence, this paper introduces a novel symbolic representation that amalgamates both explicit and implicit musical information across diverse…

声音 · 计算机科学 2024-01-08 Yikai Qian , Tianle Wang , Xinyi Tong , Xin Jin , Duo Xu , Bo Zheng , Tiezheng Ge , Feng Yu , Song-Chun Zhu

Recent years have witnessed a growing interest in research related to the detection of piano pedals from audio signals in the music information retrieval community. However, to our best knowledge, recent generative models for symbolic music…

声音 · 计算机科学 2021-11-03 Joann Ching , Yi-Hsuan Yang

As increasingly more software services have been published onto the Internet, it remains a significant challenge to recommend suitable services to facilitate scientific workflow composition. This paper proposes a novel NLP-inspired approach…

软件工程 · 计算机科学 2022-05-25 Xihao Xie , Jia Zhang , Rahul Ramachandran , Tsengdar J. Lee , Seungwon Lee

Drawing inspiration from the notion of cognitive incongruence associated with Stroop's famous experiment, from musical principles, and from the observation that music consumption on an individual basis is becoming increasingly ubiquitous,…

音频与语音处理 · 电气工程与系统科学 2018-11-19 Ishwarya Ananthabhotla , Joseph A. Paradiso

We describe Verse by Verse, our experiment in augmenting the creative process of writing poetry with an AI. We have created a group of AI poets, styled after various American classic poets, that are able to offer as suggestions generated…

计算与语言 · 计算机科学 2022-05-12 David Uthus , Maria Voitovich , R. J. Mical

This paper investigates the emerging text-to-audio paradigm in artificial intelligence (AI), examining its transformative implications for musical creation, interpretation, and cognition. I explore the complex semantic and semiotic…

声音 · 计算机科学 2025-11-24 Guilherme Coelho

We propose a new task in the area of computational creativity: acrostic poem generation in English. Acrostic poems are poems that contain a hidden message; typically, the first letter of each line spells out a word or short phrase. We…

计算与语言 · 计算机科学 2020-10-07 Rajat Agarwal , Katharina Kann