Use this when splitting an ACL paper between body, appendix, and archive. The governing ARR principle: reviewers are not required to consider material in appendices or supplements, so anything decision-critical that lives only there is effectively invisible.
[ content pages: 8 long / 4 short ] <- the reviewed argument lives here
[ Limitations (REQUIRED, unlimited) ] <- after conclusion, outside page count
[ Ethics statement (optional) ]
[ References (unlimited) ]
[ Appendices (unlimited, same PDF) ] <- optional reading for reviewers
+ separate .tgz/.zip archive <- software / data supplement
Missing Limitations is a desk-reject condition; treating it as one throwaway sentence is a review-stage penalty even when it passes the gate.
| Weak pattern | Stronger ACL pattern |
|---|---|
| "Results may not generalize" | Name the languages, domains, and model scales actually tested and the nearest untested regime |
| "LLMs can hallucinate" | State which conclusions depend on a specific model snapshot and API behavior |
| Silent on data | Note license constraints, demographic skew, or collection-window bias in the corpora used |
| Written last-minute | Mirrors the risks reviewers will find anyway, defusing them on your terms |
ACL's policy explicitly instructs reviewers not to punish honest limitations, which makes this section the cheapest goodwill in the whole submission.
acl-reproducibility).A long paper introduces a retrieval-augmented QA method with results on six benchmarks in three languages. Body: method figure, main table (six benchmarks averaged + per-language block), two-paragraph error analysis, one ablation that carries the mechanism claim. Appendix: full per-benchmark tables, prompts, retrieval index details, remaining ablations, annotation guidelines for the human study. Archive: code, prompts as files, and all model outputs. The test: a reviewer who never scrolls past the references can still reconstruct and believe every claim in the abstract.
Number tables and figures continuously with the body so the author response can cite "Table 9" unambiguously during the discussion phase.
Write it when the paper involves human subjects or annotators, scraped user-generated content, demographic inference, dual-use capability, or release of models/data with realistic misuse paths. Skip it when nothing applies — the Responsible NLP checklist already covers the routine cases, and a padded statement invites the very scrutiny it fails to answer. It shares the unlimited space after the conclusion with Limitations.
[Split status] sound / body-overloaded / appendix-dependent
[Limitations quality] substantive / ritual / missing
[Must-move-up] <decision-critical items currently below the fold>
[Archive check] <format/anonymity/clean-machine findings>
[Reviewer-blind spots] <claims visible only outside the body>