Loading the leaderboard…
Loading the leaderboard…
02 · Compare
Overall WER against CS-WER — errors inside code-switched spans — scoped to ViMedCSS, the hub’s first live dataset. Lower-left wins. Bootstrap 95% CIs; normalizer v0.1; WER is only comparable within a normalizer version.
02.1 · Head-to-head
Pick a baseline and up to two challengers — all measured on ViMedCSS. Deltas are challenger minus baseline — green improves on the baseline, red regresses. Models without measurements stay pickable and show their evidence gap.
02.2 · Efficiency
RTFx — audio seconds per wall second on platform hardware — against WER on ViMedCSS, with bubble area scaling by parameter count. Track E (sub-1B, CPU, under 8 GB VRAM) is config-ready.
2 of 17 systems measured — RTFx only ships with platform-verified runs; the rest of this plane is the evidence gap, not the field.