TriA Pipeline: A Large-Scale Automatic Audio Annotation Pipeline For Audio Classification In Specific Scenarios

๐Ÿ“… 2026-07-07
๐Ÿ“ˆ Citations: 0
โœจ Influential: 0
๐Ÿ“„ PDF
๐Ÿค– AI Summary
This study addresses the scarcity of high-quality labeled data for audio classification in domain-specific scenarios such as domestic environments. To this end, the authors propose TriA Pipeline, the first large-scale automated audio annotation framework tailored for such settings, which efficiently generates the TriA dataset comprising 2,130 hours of audio across 431 event classes. Furthermore, they introduce a prior knowledgeโ€“guided data filtering mechanism to construct a refined subset, TriA_GK. Experimental results on three household audio classification tasks demonstrate that models trained on TriA_GK achieve relative improvements of 3.97% in average accuracy and 3.35% in Macro-F1 score over baseline methods, highlighting the effectiveness of the proposed approach.
๐Ÿ“ Abstract
There are some datasets of varying scales for audio classification (AC) applied to different tasks. However, annotated data is limited for most scenarios, such as domestic environments. To address this challenge, we propose an $\textbf{A}$utomatic $\textbf{A}$udio $\textbf{A}$nnotation Pipeline--TriA Pipeline, which can efficiently convert audio from various scenarios into high-quality training data with audio event annotations. A TriA dataset was constructed with the TriA Pipeline, over 2130 hours of audio covering 431 audio classes. Furthermore, we partitioned a prior-knowledge-guided subset (TriA$_{\mathrm{GK}}$) from TriA and conduct comparative experiments on three domestic AC tasks. Comparing the result on manually annotated data only and that on manually annotated data combines TriA$_{\mathrm{GK}}$, TriA$_{\mathrm{GK}}$ could achieve average relative gains of 3.97% in accuracy and 3.35% in Macro-F1, validating the effectiveness of TriA$_{\mathrm{GK}}$ and the TriA Pipeline.
Problem

Research questions and friction points this paper is trying to address.

audio classification
data annotation
limited labeled data
domestic environments
audio datasets
Innovation

Methods, ideas, or system contributions that make the work stand out.

Automatic Audio Annotation
Audio Classification
TriA Pipeline
Prior-Knowledge-Guided
Large-Scale Dataset
๐Ÿ”Ž Similar Papers
No similar papers found.
๐Ÿ’ผ Related Jobs
No related jobs found.
H
Hong Lyu
School of Electronic and Information Engineering, South China University of Technology, Guangzhou, China
M
Mingru Yang
School of Electronic and Information Engineering, South China University of Technology, Guangzhou, China
Q
Qianhua He
School of Electronic and Information Engineering, South China University of Technology, Guangzhou, China
Y
Yanxiong Li
School of Electronic and Information Engineering, South China University of Technology, Guangzhou, China
J
Jinxin Huang
School of Electronic and Information Engineering, South China University of Technology, Guangzhou, China
Z
Zhengyu Pei
School of Electronic and Information Engineering, South China University of Technology, Guangzhou, China