Side B changes the core ranking behavior from contributor-level to commit-level by comparing every eligible commit, rolling scores back up to contributors, updating prompts, evidence, UI, and adding tests for same-contributor multi-commit cases and new ranking outputs. Side A fixes a real reducer bug by moving the zero-ratio guard before side effects to prevent ghost items and voted pairs, with an accompanying regression test, but its scope is much narrower than the architectural change in Side B.
constitution · epochs · watch · epoch 3 · comparison · attempt
judgment
openai/gpt-chat-latest → B (5:1)
jud_f5ce47efd29ffb · raw event
Metadata
judgment_idjud_f5ce47efd29ffbe6ed7644c6cf675636042393de35c9adff1800af1cfda66e41
model_idopenai/gpt-chat-latest
winnerB
ratio5:1
comparison_idcmp_7693a31cebcef7af4d830c91f178cce52559dcc8df39dbe4017b76d9388e4e14
attempt_idatt_a8616be8054c9b7ff2bf7056fe59f84c59db9b438dfa3cbfef750c890ad5c9dc