pldi-experiments
brycewang-stanford/Awesome-Journal-Skills
This guide provides best practices for designing and auditing academic evaluations of programming language features and compilers, following standards used in venues like PLDI. It covers selecting defensible benchmark suites, establishing strong baselines, performing targeted ablations, and measuring multiple performance metrics (runtime, compile time, memory) to ensure research claims are reproducible and scientifically sound.