OpenSVBench Scenario-driven SV leaderboard
Speaker Verification Benchmark

OpenSVBench speaker verification leaderboard.

Global ranking compares reviewed full-core submissions. Scenario pages group related evaluation conditions. Trial-set pages show the raw ranked results behind each scenario.

18 ranked systems 11 scenarios 26 trial sets

Dataset names are shown as source metadata inside the scenario and trial-set views.

Current #1 model: wespeaker/w2vbert2 with 9.9604 scenario-macro EER.

How To Read
  • Global ranking: reviewed full-core submissions ranked by source-balanced scenario-macro EER.
  • Scenarios: high-level evaluation conditions.
  • Trial sets: raw ranked benchmark entries.
  • Datasets: source metadata shown inside scenario and trial-set views.
Ranked Systems
18
Reviewed submissions included in the public ranking.
Benchmark Scenarios
11
High-level scenarios such as short-duration robustness and aging robustness.
Trial Sets
26
The current trial-set entries used by the benchmark.
Source Datasets
16
Source datasets behind the ranked trial sets.
Global Leaderboard

Ranking

Every ranked row is maintainer-reviewed and recomputed against the current benchmark definition, so rankings stay comparable as the leaderboard is updated.

Overall #1

wespeaker/w2vbert2

vb2-w2vbert

Lowest scenario-macro EER across the current benchmark.

#1 7 scenarios ranked #1
Broadest Scenario Top-3 Coverage

wespeaker/w2vbert2

vb2-w2vbert

Highest count of top-three high-level scenario ranks on the current board.

#1 9 scenarios in top 3
Rank Model Global Score Higher-Ranked Scenarios Lower-Ranked Scenarios
#1 wespeaker/w2vbert2
vb2-w2vbert
9.9604
scenario-macro EER
0.360767 scenario-macro minDCF
#2 palabraai/redimnet2-b6
redimnet2
10.5277
scenario-macro EER
0.419432 scenario-macro minDCF
#3 wespeaker/samresnet100
vb2-sam100
10.5780
scenario-macro EER
0.390239 scenario-macro minDCF
#4 wespeaker/res152-voxceleb
wespeaker_resnet152
11.2275
scenario-macro EER
0.442748 scenario-macro minDCF
#5 wespeaker/res293-voxceleb
wespeaker_resnet293
11.3130
scenario-macro EER
0.439607 scenario-macro minDCF
#6 iic/eres2netv2-zh
eresv2_zh
11.5449
scenario-macro EER
0.484961 scenario-macro minDCF
#7 wespeaker/res34-voxceleb
wespeaker_resnet34_voxceleb
12.0516
scenario-macro EER
0.477054 scenario-macro minDCF
#8 iic/eres2net-en
eres_en
12.0955
scenario-macro EER
0.472802 scenario-macro minDCF
#9 wespeaker/campplus-voxceleb
campplus_wespeaker
13.4165
scenario-macro EER
0.520675 scenario-macro minDCF
#10 speechbrain/ecapa
ecapa
13.5461
scenario-macro EER
0.543660 scenario-macro minDCF
#11 wespeaker/ecapa1024-voxceleb
wespeaker_ecapa1024
14.0835
scenario-macro EER
0.534111 scenario-macro minDCF
#12 iic/eres2net-large-3dspeaker
eres2net_large_3dspeaker
14.1682
scenario-macro EER
0.593516 scenario-macro minDCF
#13 wespeaker/ecapa512-voxceleb
wespeaker_ecapa512
14.3120
scenario-macro EER
0.541141 scenario-macro minDCF
#14 wespeaker/res34-cnceleb
wespeaker_r34
14.4910
scenario-macro EER
0.607964 scenario-macro minDCF
#15 iic/campplus
campplus
15.1677
scenario-macro EER
0.590178 scenario-macro minDCF
#16 speechbrain/xvector
xvector
23.3539
scenario-macro EER
0.776078 scenario-macro minDCF
#17 microsoft/wavlm-base-plus-sv
wavlm_base
25.1577
scenario-macro EER
0.875090 scenario-macro minDCF
#18 microsoft/unispeech-sat-base-plus-sv
unispeech_sat_base_plus
25.6711
scenario-macro EER
0.890736 scenario-macro minDCF
Benchmark Scenarios

Scenarios covered by OpenSVBench

Each scenario describes a high-level evaluation condition, while the underlying trial sets and datasets remain visible.

Open Scenario Index
Accent/dialect 2 trial sets

Accent / Dialect Robustness

How reliably the model tracks identity across accent and dialect mismatch.

Accent and dialect variation GLOBE 3D-Speaker Dialect
View scenario ranking
Aging 2 trial sets

Aging Robustness

How stable identity representations remain across age-derived and longitudinal recording gaps.

Speaker aging and time gaps VoxKnesset VoxPopuli Aging
View scenario ranking
Channel/device 2 trial sets

Channel / Device Robustness

How resilient the model is to device and channel mismatch.

Device and channel variation 3D-Speaker Device FFSVC 2022 Cross-Channel
View scenario ranking
Cross-lingual 1 trial set

Cross-Lingual Robustness

How well speaker identity survives enrollment-test language mismatch.

Language mismatch TidyVoiceX2-ASV
View scenario ranking
Distance 4 trial sets

Distance Robustness

How much performance changes across meeting and domestic distance mismatch.

Distance mismatch 3D-Speaker Distance AliMeeting Near/Far +2 more
View scenario ranking
Genre shift 1 trial set

Genre-Shift Robustness

Whether performance holds when CN-Celeb enrollment and test speech come from different source genres.

Source-genre variation CN-Celeb Genre
View scenario ranking
In-the-wild 4 trial sets

In-The-Wild Robustness

How strong the model is on unconstrained celebrity and media speech across official CN-Celeb and VoxCeleb protocols.

Open-domain media speech CN-Celeb VoxCeleb1-O +2 more
View scenario ranking
Noise/reverb 1 trial set

Noise / Reverb Robustness

How reliably the model preserves identity when room acoustics, distractor noise, and microphone placement deviate from an easier in-corpus reference condition.

Noise and reverberation VOiCES Noise/Reverb
View scenario ranking
Overlap 2 trial sets

Overlap Robustness

How well speaker identity survives light, mid, and heavy overlap in both meeting and domestic recordings.

Overlapping speakers AliMeeting Overlap CHiME-6 Overlap
View scenario ranking
Short-duration 4 trial sets

Short-Duration Robustness

How much performance holds up when speech evidence is limited by duration.

Short-duration speech HI-MIA GSC Short +2 more
View scenario ranking
Speaking style 3 trial sets

Speaking-Style Robustness

Whether the model can preserve identity across emotion-driven change, whispered speech, and noise-induced Lombard speaking style.

Speaking style shift ESD Whisper40 Whisper +1 more
View scenario ranking