A research pipeline for predicting Codeforces problem ratings from official API
metadata, solved statistics, exposure-aware features, and lightweight
statement-structure features. Historical paper artifacts remain available for
transparent comparison with the corrected evaluation pipeline.
Research status (updated 2026-07-12): Earlier paper results are retained
for transparency and should be interpreted as retrospective. A code audit
identified a split-configuration mismatch and test-informed model selection;
the corrected controls and scope are documented in the
public erratum.
No corrected rerun has replaced the legacy full-API headline metrics. A
separate locked historical statement-only backtest
is now reported below.
If you only want the main idea, key results, and how to read the project quickly, see QUICK_OVERVIEW.md.
Cold-start note: Cold-start here means no solved-count behavior, not necessarily strict pre-contest prediction. Tags, metadata, and statement availability may not exactly match a real pre-contest setting.