| Task Name | Dataset Name | SOTA Result | Trend | |
|---|---|---|---|---|
| Mispronunciation Detection | L2-ARCTIC (test) | F1 Score71.77 | 20 | |
| Accent Normalization | L2-ARCTIC (test) | NAT70.73 | 17 | |
| Phoneme Recognition | L2-ARCTIC (test) | Phoneme Error Rate (PER)13.72 | 14 | |
| Mispronunciation Diagnosis | L2-ARCTIC (test) | EDR19.98 | 14 | |
| Automatic Speech Recognition | L2-ARCTIC (test) | WER7.2 | 9 | |
| Speech Recognition | L2-Arctic | Arabic WER15.54 | 9 | |
| Phonetic Perception | L2-ARCTIC | PFER8 | 8 | |
| Suggestion Generation | L2-Arctic-plus (test) | BLEU-220.4 | 8 | |
| Mispronunciation Detection | L2-Arctic-plus (test) | Precision53.8 | 8 | |
| Automatic Speech Recognition | L2-Arctic Arabic and Vietnamese accents (test) | WER (ARB)20.3 | 8 | |
| Automatic Speech Recognition | L2-Arctic Chinese and Vietnamese accents (test) | WER (CHN)25.3 | 8 | |
| Automatic Speech Recognition | L2-Arctic Arabic and Chinese accents (test) | WER (Arabic)20.3 | 8 | |
| Automatic Speech Recognition | L2-ARCTIC created subset | Error Rate (ZH)11.67 | 8 | |
| Mispronunciation Detection and Diagnosis | L2-ARCTIC 6-speaker subset (test) | F1 Score69.6 | 7 | |
| Speech Quality Assessment | L2-ARCTIC Indian English speakers | NISQA-MOS4.84 | 6 | |
| Accent Neutralization | L2-ARCTIC Indian English speakers | CNA100 | 6 | |
| Text-to-Speech | L2-ARCTIC Non-native | WER1.44 | 5 | |
| Mispronunciation Diagnosis | L2-ARCTIC | FRR5.07 | 5 | |
| Mispronunciation Detection | L2-ARCTIC | Recall61.67 | 5 | |
| Automatic Speech Recognition | L2-ARCTIC 6-speaker subset (test) | PER12.55 | 5 | |
| Accent Conversion | L2 Arctic | UTMOS4.001 | 4 | |
| Automatic Speech Recognition | L2-ARCTIC Unseen-accent (UA) | WER7.41 | 4 | |
| Automatic Speech Recognition | L2-ARCTIC (Unseen-transcript (UT)) | WER9.14 | 4 | |
| Automatic Speech Recognition | L2-ARCTIC Korean English (YKWK) | WER0.123 | 4 | |
| Automatic Speech Recognition | L2-ARCTIC Korean English (YDCK) | WER11.6 | 4 |