Side A changes the ranking color logic from list-position-based gradients to score-based min–max normalization within each group, introduces a dedicated `score_gradient_t` helper, updates callers, and adds targeted tests covering normalization behavior and edge cases such as tied scores. Side B fixes a real configuration bug by defining `GITHUB_API_BASE_URL` with a default, but it is a small missing-variable fix with narrower impact than the broader, tested behavioral improvement in Side A.
constitution · epochs · watch · epoch 3 · comparison · attempt
judgment
openai/gpt-chat-latest → A (4:1)
jud_d1ffe7821f1c7a · raw event
Metadata
judgment_idjud_d1ffe7821f1c7ab5f0e21050dd54d5e562fc257360e032af4d4d244601477dca
model_idopenai/gpt-chat-latest
winnerA
ratio4:1
comparison_idcmp_66a3a76bc36f30ef88a4cb6ff7dcb72ee464a9a6ac6cc63cb6bcecd832d3ae3b
attempt_idatt_28edf91da7ababf5aebd0ddfc7fbf91179ec1be0d8f38490c11d96737f1be0e0