Papers and publication summaries.
TokAN: Accent Normalization Using Self-Supervised Speech Tokens
TokAN is a token-based accent normalization framework built on self-supervised discrete speech tokens, jointly trained VQ tokenization, autoregressive token conversion, flow-matching synthesis, and GRPO post-training.