← 피드로
[Submitted on 6 Jul 2026]
Abstract:While audio deepfake detection has advanced significantly, representative detectors show limited generalization to synthetic sound effects. Existing environmental audio datasets such as EnvSDD provide important initial resources, but remain limited in scale and generation provenance for studying isolated sound-effect deepfakes. To support this direction, we present SynSFX, a large-scale corpus of 43374 clips (26452 synthetic, 16922 real) spanning 7 popular text-to-audio models.
Submission history
From: Linxi Li [view email]
[v1]
Mon, 6 Jul 2026 09:19:03 UTC (154 KB)
추출 본문 · 출처: arxiv.org · https://arxiv.org/abs/2607.04848
답글 남기기