TAME: Temporal Audio-based Mamba for Enhanced Drone Trajectory Estimation and Classification

๐Ÿ“… 2024-12-17
๐Ÿ“ˆ Citations: 3
โœจ Influential: 0
๐Ÿ“„ PDF
๐Ÿค– AI Summary
The proliferation of small unmanned aerial vehicles (UAVs) poses escalating public safety risks, yet existing detection systems suffer from large physical footprints, high costs, and limited accuracy in trajectory estimation and UAV-type classification. To address these challenges, this paper proposes a lightweight, audio-driven counter-UAV detection framework. Methodologically, it introduces a novel parallel selective state-space model (Mamba) architecture that fuses timeโ€“frequency acoustic features; incorporates a residual cross-attention mechanism to enhance temporal modeling; and jointly performs 3D acoustic source localization and fine-grained UAV-type classification. Evaluated on the MMUAD benchmark, the method achieves state-of-the-art performance: trajectory estimation error is reduced by 21.3%, and UAV-type classification accuracy reaches 94.7%. The source code and trained models are publicly released.

Technology Category

Computer Vision: Multi-modal VisionIntelligent Robots: State EstimationMachine Learning: Large Multimodal Models (LMMs)

Application Category

User Modeling, Personalization and Recommendation: Attacks and countermeasures in recommendation systemsSystems and Infrastructure for Web, Mobile and WoT: Applied ML and AI for Web-based mobile applicationsResponsible Web: Machine-in-the-loop, human agency and autonomy
๐Ÿ“ Abstract
The increasing prevalence of compact UAVs has introduced significant risks to public safety, while traditional drone detection systems are often bulky and costly. To address these challenges, we present TAME, the Temporal Audio-based Mamba for Enhanced Drone Trajectory Estimation and Classification. This innovative anti-UAV detection model leverages a parallel selective state-space model to simultaneously capture and learn both the temporal and spectral features of audio, effectively analyzing propagation of sound. To further enhance temporal features, we introduce a Temporal Feature Enhancement Module, which integrates spectral features into temporal data using residual cross-attention. This enhanced temporal information is then employed for precise 3D trajectory estimation and classification. Our model sets a new standard of performance on the MMUAD benchmarks, demonstrating superior accuracy and effectiveness. The code and trained models are publicly available on GitHub url{https://github.com/AmazingDay1/TAME}.
Problem

Research questions and friction points this paper is trying to address.

Small Unmanned Aerial Vehicles (SUAVs)
Public Safety
Cost-effective Drone Detection
Innovation

Methods, ideas, or system contributions that make the work stand out.

TAME System
Time-Frequency Analysis
3D Path Prediction
๐Ÿ”Ž Similar Papers
No similar papers found.
Nanjing University of Aeronautics and Astronautics
Z
Zhenyuan Xiao
College of Automation Engineering, Nanjing University of Aeronautics and Astronautics, China
H
Huanran Hu
College of Automation Engineering, Nanjing University of Aeronautics and Astronautics, China
G
Guili Xu
College of Automation Engineering, Nanjing University of Aeronautics and Astronautics, China
Junwei He
Junwei He
Institute of Computing Technology, Chinese Academy of Sciences
LLM ReasoningGraph Learning