Search Results for author: Jinhua Liang

Found 7 papers, 4 papers with code

Mind the Domain Gap: a Systematic Analysis on Bioacoustic Sound Event Detection

1 code implementation • 27 Mar 2024 • Jinhua Liang, Ines Nolasco, Burooj Ghani, Huy Phan, Emmanouil Benetos, Dan Stowell

A recent development in the field is the introduction of the task known as few-shot bioacoustic sound event detection, which aims to train a versatile animal sound detector using only a small set of audio samples.

Data Augmentation Domain Adaptation +3

Paper
Code

WavCraft: Audio Editing and Generation with Large Language Models

1 code implementation • 14 Mar 2024 • Jinhua Liang, huan zhang, Haohe Liu, Yin Cao, Qiuqiang Kong, Xubo Liu, Wenwu Wang, Mark D. Plumbley, Huy Phan, Emmanouil Benetos

We introduce WavCraft, a collective system that leverages large language models (LLMs) to connect diverse task-specific models for audio content creation and editing.

In-Context Learning

640

Paper
Code

Acoustic Prompt Tuning: Empowering Large Language Models with Audition Capabilities

1 code implementation • 30 Nov 2023 • Jinhua Liang, Xubo Liu, Wenwu Wang, Mark D. Plumbley, Huy Phan, Emmanouil Benetos

Moreover, we improve the framework of audio language model by using interleaved audio-text embeddings as the input sequence.

Audio Classification Few-Shot Audio Classification +2

Paper
Code

WavJourney: Compositional Audio Creation with Large Language Models

1 code implementation • 26 Jul 2023 • Xubo Liu, Zhongkai Zhu, Haohe Liu, Yi Yuan, Meng Cui, Qiushi Huang, Jinhua Liang, Yin Cao, Qiuqiang Kong, Mark D. Plumbley, Wenwu Wang

Subjective evaluations demonstrate the potential of WavJourney in crafting engaging storytelling audio content from text.

Audio Generation

509

Paper
Code

Adapting Language-Audio Models as Few-Shot Audio Learners

no code implementations • 28 May 2023 • Jinhua Liang, Xubo Liu, Haohe Liu, Huy Phan, Emmanouil Benetos, Mark D. Plumbley, Wenwu Wang

We presented the Treff adapter, a training-efficient adapter for CLAP, to boost zero-shot classification performance by making use of a small set of labelled data.

Audio Classification Few-Shot Learning +1

Paper
Add Code

Leveraging Pre-trained AudioLDM for Text to Sound Generation: A Benchmark Study

no code implementations • 7 Mar 2023 • Yi Yuan, Haohe Liu, Jinhua Liang, Xubo Liu, Mark D. Plumbley, Wenwu Wang

Deep neural networks have recently achieved breakthroughs in sound generation with text prompts.

Audio Generation Benchmarking +1

Paper
Add Code

Channel Compression: Rethinking Information Redundancy among Channels in CNN Architecture

no code implementations • 2 Jul 2020 • Jinhua Liang, Tao Zhang, Guoqing Feng

Aiming at channel compression, a novel convolutional construction named compact convolution is proposed to embrace the progress in spatial convolution, channel grouping and pooling operation.

Acoustic Scene Classification Event Detection +4

Paper
Add Code

Cannot find the paper you are looking for? You can Submit a new open access paper.