# Smoke test utterances

**URL:** https://community.drivendata.org/t/smoke-test-utterances/11401
**Category:** Children’s Speech Recognition Challenge
**Created:** [March 6, 2026, 11:39pm UTC](https://community.drivendata.org/t/smoke-test-utterances/11401 "2026-03-06T23:39:30Z")
**Posts on this page:** 4
**Page:** 1

<div class="post-metadata">

### Author: ![jialuli](https://avatars.discourse-cdn.com/v4/letter/j/a87d85/32.png) [@jialuli](https://community.drivendata.org/u/jialuli)
#### Post date: [March 6, 2026, 11:39pm UTC](https://community.drivendata.org/t/smoke-test-utterances/11401/1 "2026-03-06T23:39:30Z")

</div>

Hi,

I was wondering whether you could share the exact utterances used in the smoke dataset, along with the corresponding scoring script.

I ran my model locally on the utterances in smoke\_test\_submission\_format and evaluated the results using metric/score.py from the provided GitHub repository, but the WER I obtained locally was dramatically different from the WER reported by the cloud smoke test.

Having access to the exact smoke test utterances and scoring setup would be very helpful for debugging whether the discrepancy is due to an environment mismatch or an issue with my own model.

Thank you very much for your help.

---

<div class="post-metadata">

### Author: ![cszc](https://avatars.discourse-cdn.com/v4/letter/c/c68b51/32.png) [@cszc](https://community.drivendata.org/u/cszc)
#### Post date: [March 7, 2026, 3:01am UTC](https://community.drivendata.org/t/smoke-test-utterances/11401/2 "2026-03-07T03:01:29Z")

</div>

Hi @jialuli - The exact utterance IDs are shared in the “Smoke test submission format” file on the data download pages. That and `metric/score.py` should give you everything you need to replicate the score locally. Good luck!

---

<div class="post-metadata">

### Author: ![oknaitik](https://avatars.discourse-cdn.com/v4/letter/o/cc9497/32.png) [@oknaitik](https://community.drivendata.org/u/oknaitik)
#### Post date: [March 7, 2026, 1:15pm UTC](https://community.drivendata.org/t/smoke-test-utterances/11401/3 "2026-03-07T13:15:24Z")

</div>

Pardon me but does the smoke test score is representative of the public LB test set to any degree? I believed it doesn’t since it’s mentioned that it’s fake data and only purpose is test the submission run E2E, right?

---

<div class="post-metadata">

### Author: ![cszc](https://avatars.discourse-cdn.com/v4/letter/c/c68b51/32.png) [@cszc](https://community.drivendata.org/u/cszc)
#### Post date: [March 7, 2026, 2:50pm UTC](https://community.drivendata.org/t/smoke-test-utterances/11401/4 "2026-03-07T14:50:38Z")

</div>

No, the smoke test does not impact the leaderboard score at all. It is drawn from the training data.
