Side B fixes a correctness bug in the core ranking algorithm by changing the Rank Centrality transition matrix to use degree-based d_max instead of summed edge weights, preventing oscillation on star topologies and producing the expected stationary distribution. It also adds targeted Rust and end-to-end regression tests with ranking fixtures, whereas Side A is largely a refactor that removes the demo counter and introduces cached ranking/settlement infrastructure without an equally clear correctness improvement.
constitution · epochs · watch · epoch 3 · comparison · attempt
judgment
openai/gpt-chat-latest → B (3:2)
jud_c498683956683a · raw event
Metadata
judgment_idjud_c498683956683a78d33181ba8e7a8248ea12dccda41d87b94c58005e1d290f41
model_idopenai/gpt-chat-latest
winnerB
ratio3:2
comparison_idcmp_127f7e579c702215f224d5f8600902fda4764aa96a2112e7f6972def43cadc4c
attempt_idatt_3b7d3ac000500caac3d5a6e918efa45eb8732c08d35dcfd06c29e00981e9ddb9