Four pages is not a short paper; it is a different literary form. The Interspeech house style evolved to let a reader — who may be a phonetician, an ASR engineer, or an area chair skimming forty PDFs — extract task, method, corpus, metric, and delta in one pass. Everything below serves that one-pass property.
| Move | Before | After |
|---|---|---|
| Protocol one-liner | A paragraph on data prep | "We follow the ESPnet LibriSpeech recipe (commit X) except: …" |
| Diff-based method description | Re-derive the known architecture | "As in Conformer [n], with two changes: …" |
| Table consolidation | Separate tables per test set | One table, sets as column groups, best in bold |
| Citation compression | Three sentences of related work per item | Clause-level positioning: "unlike streaming approaches [4,5], we…" |
| Figure economy | Architecture diagram redrawing the standard | One figure that carries the finding (error breakdown, trade-off curve) |
The related-work section can shrink to a paragraph if positioning is woven into
the introduction — common and accepted at this venue (see
interspeech-related-work).
Never buy space with template tricks; the PDF checker and the format police are
real (see interspeech-submission).
Before (3 lines): We evaluate our proposed method on the LibriSpeech
dataset. Our method achieves significant improvements over the baseline.
Results are shown in Table 2, which demonstrates the effectiveness.
After (2 lines): On LibriSpeech, the proposed front end reduces WER from
5.6% to 4.9% on test-other (Table 2; paired bootstrap p<0.01), with the
largest gains on utterances under 3 s.
The after-version is shorter and carries corpus, metric, delta, significance, and an analysis hook.
| Section | Budget | Job |
|---|---|---|
| Abstract + index terms | 0.15 col | task→move→corpus→metric→delta |
| 1. Introduction | 0.75 col | gap, hypothesis, contribution — no history lesson |
| 2. Relation to prior work | 0.3–0.5 col | clause-level positioning (may merge into 1) |
| 3. Method | 1.2 col | the diff against cited standards, one figure max |
| 4. Experimental setup | 0.6 col | corpus, protocol one-liners, disclosure block |
| 5. Results + analysis | 1.5 col | decisive table, ablation, one error-pattern finding |
| 6. Conclusions | 0.2 col | finding restated, scoped future sentence |
Budgets flex ±30% by paper type (science papers grow §4–5, systems papers §3), but a draft violating the total by a full column will not compress overnight.
Interspeech's readership is the most linguistically diverse of the major ML-adjacent venues. Prefer plain constructions, keep sentences under ~25 words, and gloss language-specific phenomena (tone sandhi, pitch accent) in one clause — reviewers outside your language's community must still follow the claim.
[One-pass check] task/method/corpus/metric/delta extractable from p.1? y/n
[Opening] title, abstract-number, index-terms verdicts
[Density audit] paragraphs convertible to one-liners or diffs
[Diction] bare numbers, undefined acronyms, unverifiable verbs found
[Compression plan] ordered cuts to reach 4 pages without claim loss
[Escalation] fits in 4pp / needs Long Paper / needs journal
Format facts (4 content pages + references page; Long Paper track parameters)
follow the 2026 cycle as logged in resources/official-source-map.md
(2026-07-08); the paper kit of the current cycle is the sole format authority.