constitution · epochs · watch · epoch 3 · comparison · attempt

judgment

~x-ai/grok-latest → B (2:3)

jud_36a0146c2710b8 · raw event

B fixes a core correctness bug in compute_scores_from_edges (weight-sum d_max → unweighted degree d_max per Negahban–Oh–Shah), so star topologies converge instead of oscillating to uniform scores, with Rust and Clojure fixture regression coverage. A is strong lasting design—demo-counter removal plus a batched SettlementClient and read-only ranked_items_cached path—but it mainly restructures how votes are applied and ranked, whereas B makes the ranking results themselves trustworthy.

Metadata
judgment_idjud_36a0146c2710b8cac773ecb73edf00141b487cadca2fa338b24bea61f4ca0e94
model_id~x-ai/grok-latest
winnerB
ratio2:3
comparison_idcmp_127f7e579c702215f224d5f8600902fda4764aa96a2112e7f6972def43cadc4c
attempt_idatt_16959d3260672b1d2d6022bbc8baceb966ef55c3608117f067958341f2ccde6a