They need to show some examples. You beat all the top models on your private data set? Show at least a couple examples - audio and transcripts - from examples that your model got right and others got wrong.
after reading the article I still have no idea how their thing performs, or if I should care how it performs since a majority of voice benchmarks still don't map to human evals
ks2048 · · focus · HN ↗
ymaws · · focus · HN ↗