1. Opening Remarks
9:30 โ 9:40 AM ยท 10 min- Welcome and introduction
- Overview of the challenge/session
- Introduction of schedule and presentation format
Session Schedule
On Evaluation Metrics of Speech Deepfake Detection โ Lessons Learned from ASVSpoof5
| Paper ID | Paper Title |
|---|---|
| 223 | Overview of ESDD2: Environment-Aware Speech and Sound Deepfake Detection Challenge |
Certificate presentation and commemorative photos for each team; sponsor acknowledgement during the award presentation for the first-place team.
| Order | Paper ID | Paper Title |
|---|---|---|
| 1 | 208 | Component-Level Ensemble Fusion for Speech and Environmental Sound Deepfake Detection |
| 2 | 209 | MIF: Multi-view Interaction Framework for Speech and Environment Sound Deepfake Detection |
| 3 | 215 | Dual-Stream EAT-XLSR: A Cascade Framework for Component-Aware Audio Deepfake Detection |
| 4 | 217 | Deepfake Audio Detection Using Self-supervised Fusion Representations |
| 5 | 221 | Phoneme-Guided Fusion for Environment-Aware Speech and Sound Deepfake Detection |
| 6 | 224 | EnvTriCascade: An Environment-Aware Tri-Stage Cascaded Framework for ESDD2 2026 Challenge |
| 7 | 232 | From Signal Separation to Feature Decomposition: A Framework for Synthetic Speech Detection in Complex Environments |
Acknowledgement to participants, organizers, reviewers, sponsor, and ICME.
| Time | Event | Duration | Participants | Host |
|---|---|---|---|---|
| 9:30 โ 9:40 | Opening Remarks | 10 min | โ | Ming Li |
| 9:40 โ 10:40 | Invited Talk & Q&A | 60 min | Xin Wang | Ming Li |
| 10:40 โ 11:00 | Oral Session | 20 min | Summary paper | Ming Li |
| 11:00 โ 11:10 | Coffee Break | 10 min | โ | Ming Li |
| 11:10 โ 12:00 | Oral Session & Awards | 50 min | All papers | Ming Li |
| 12:00 โ 12:20 | Panel Discussion | 20 min | Ming Li, Xin Wang | Xueping Zhang |
| 12:20 โ 12:30 | Closing Remarks | 10 min | โ | Ming Li |