Side B changes the ranking algorithm itself from contributor-level comparisons to pairwise ranking of every eligible commit, adds commit-level evidence and rollup logic, updates UI/evidence pages to expose per-commit rankings, and fixes the short-circuit so multiple commits by one contributor are still evaluated. Side A is a broad DSL syntax migration (leading explanation blocks and title-first items) with parser rewrites and widespread fixture updates, but it primarily changes input format rather than adding comparable lasting system capability.
constitution · epochs · watch · epoch 3 · comparison · attempt
judgment
openai/gpt-chat-latest → B (5:1)
jud_366ab6419d29bf · raw event
Metadata
judgment_idjud_366ab6419d29bf02677294d33cbe288ee9540cd48fd0e17fc71be2b367e04033
model_idopenai/gpt-chat-latest
winnerB
ratio5:1
comparison_idcmp_a19d0f3074e472cf532fc6b629586849253e378f03a27f4988a3b55623a10f3e
attempt_idatt_be4ed79f49006f7587472f7b62748a41784869851988ec1e6d85b53836d352a4