Take audit: GJCrQMikojWu53O8cWy_4

2026-08-22T09:20:00.921Z — Mika

0/13 judged · keyboard: Space play/pause, ↑/↓ focus, P pass, T Translation failure, F fail

Do a blind check (transcripts hidden)

Ground Truth

What the Tester actually said. Never pre-filled from an Engine — type what you hear.


Transcripts

mini_prod_baseline gpt-4o-mini-transcribe · degraded · en · bare — MER accuracy pending Ground Truth · 647 ms · verdict: unjudged

老師,Can you help me? 我需要幫忙

mini_auto_bare_deg gpt-4o-mini-transcribe · degraded · auto · bare — MER accuracy pending Ground Truth · 529 ms · verdict: unjudged

老師,Can you help me? 我需要幫忙

mini_auto_bare_clean gpt-4o-mini-transcribe · clean · auto · bare — MER accuracy pending Ground Truth · 773 ms · verdict: unjudged

老師,Can you help me? 我需要幫忙。

mini_auto_prompt_deg gpt-4o-mini-transcribe · degraded · auto · mixed prompt — MER accuracy pending Ground Truth · 667 ms · verdict: unjudged

老師,Can you help me? 我需要幫忙。

mini_auto_prompt_clean gpt-4o-mini-transcribe · clean · auto · mixed prompt — MER accuracy pending Ground Truth · 715 ms · verdict: unjudged

老師,can you help me? 我需要幫忙。

mini_auto_prompt_64k gpt-4o-mini-transcribe · 64k · auto · mixed prompt — MER accuracy pending Ground Truth · 702 ms · verdict: unjudged

老師,Can you help me? 我需要幫忙。

mini_zh_prompt_deg gpt-4o-mini-transcribe · degraded · zh · mixed prompt — MER accuracy pending Ground Truth · 1030 ms · verdict: unjudged

老師,Can you help me? 我需要幫忙。

mini_zh_prompt_clean gpt-4o-mini-transcribe · clean · zh · mixed prompt — MER accuracy pending Ground Truth · 842 ms · verdict: unjudged

老師,can you help me? 我需要幫忙。

gpt_tw_en_bare_deg gpt-transcribe · degraded · zh-tw + en · bare — MER accuracy pending Ground Truth · 718 ms · verdict: unjudged

老師,Can you help me? 我需要幫忙。

gpt_tw_en_prompt_deg gpt-transcribe · degraded · zh-tw + en · mixed prompt — MER accuracy pending Ground Truth · 2339 ms · verdict: unjudged

老師,Can you help me? 我需要幫忙。

gpt_tw_en_prompt_clean gpt-transcribe · clean · zh-tw + en · mixed prompt — MER accuracy pending Ground Truth · 667 ms · verdict: unjudged

老師,Can you help me? 我需要幫忙。

gpt_tw_en_prompt_64k gpt-transcribe · 64k · zh-tw + en · mixed prompt — MER accuracy pending Ground Truth · 584 ms · verdict: unjudged

老師,Can you help me? 我需要幫忙。

gpt_tw_only_prompt_deg gpt-transcribe · degraded · zh-tw · mixed prompt — MER accuracy pending Ground Truth · 687 ms · verdict: unjudged

老師,Can you help me? 我需要幫忙。

Winning configuration

Record the variant and rationale to hand to qa-rails after reviewing this Take.