Skip to content

Kimi / Moonshot AI

Moonshot AI's Kimi models and products: the open K series, long-context work and how the product changes.

1selectedRelated topicsQwenDeepSeekMiniMax

Latest selected

1–1 of 1

Aug 21

Friday
  1. Measuring benchmark optimization in speech recognition

    Hugging Face proposed three tests to quantify benchmark optimisation in speech recognition: a consensus-disagreement probe, masked-entity retrieval and spelling switches. Evaluating 11 open-source ASR models, it found some high-scoring models reproduce errors in VoxPopuli and LibriSpeech reference transcripts even when contradicted by audio, when relevant words are muted or when both spellings fit the audio.