{"abstract":"A 7.3 mark is accepted, or a legitimate 10.0 is rejected.","category":"Sports scoring and tiebreakers","checks":9,"contract":"Diving dive score. scores are judge marks as strings from 0 to 10 in half points; any other mark returns \"invalid mark <s>\". A 5-judge panel drops the single highest and lowest marks, a 7-judge panel the two highest and two lowest; any other panel size returns \"invalid panel\". The three remaining marks are summed and multiplied by the degree of difficulty dd (a decimal string); return the exact result with two decimals.","contract_signature":"scores, dd","evaluation_group":"w2-sports-scoring-diving-judges-trim","failed_approach":"Adding the half-point check but tightening the range to below 10 rejects perfect marks.","family":"w2-sports-scoring-diving-judges-trim-mark-validation","id":"FA-84256","implementations":{"attempt":{"sha256":"6c4f6720c59f519780c51cbc9b270dfe6fd411bac48605fcc648a84884004035","source":"\"\"\"Failure Map reference implementation. Python standard library only.\"\"\"\nimport json\nfrom fractions import Fraction\nN = 1\nobservations = []\ndef solve(scores, dd):\n    marks = []\n    for s in scores:\n        v = Fraction(s)\n        if v < 0 or v >= 10 or (v * 2).denominator != 1:\n            return 'invalid mark ' + s\n        marks.append(v)\n    if len(marks) == 5:\n        drop = 1\n    elif len(marks) == 7:\n        drop = 2\n    else:\n        return 'invalid panel'\n    kept = sorted(marks)[drop:len(marks) - drop]\n    total = sum(kept) * Fraction(dd)\n    return '%.2f' % float(total)\ndef check(label, actual, expected):\n    observations.append({\"check\": label, \"actual\": actual, \"expected\": expected, \"passed\": actual == expected})\ndef run(args):\n    try:\n        return solve(*args)\n    except Exception as exc:\n        return 'raised ' + type(exc).__name__\ncases = [[('control five-judge panel', (['7.0', '7.5', '6.5', '8.0', '7.0'], '2.0'), '43.00'),\n  ('control seven-judge panel',\n   (['8.0', '8.5', '9.0', '7.5', '8.0', '8.5', '6.0'], '3.1'),\n   '75.95'),\n  ('boundary perfect tens', (['10.0', '10.0', '10.0', '10.0', '10.0'], '3.4'), '102.00'),\n  ('boundary non-half mark', (['7.3', '7.0', '7.0', '7.0', '7.0'], '2.0'), 'invalid mark 7.3'),\n  ('boundary six judges', (['7.0', '7.0', '7.0', '7.0', '7.0', '7.0'], '2.0'), 'invalid panel'),\n  ('regression: mark validation', (['3.5', '10.0', '4.0', '3.5', '7.5'], '2.0'), '30.00'),\n  ('regression: mark validation', (['7.3', '8.0', '3.5', '1.5', '0.5'], '2.0'), 'invalid mark 7.3'),\n  ('variant scenario 1', (['1.5', '8.0', '3.0', '7.0', '7.5'], '3.1'), '54.25'),\n  ('variant scenario 2',\n   (['-1.0', '2.5', '2.0', '3.5', '0.0', '7.5', '6.5'], '2.0'),\n   'invalid mark -1.0')],\n [('control five-judge panel', (['7.0', '7.5', '6.5', '8.0', '7.0'], '2.0'), '43.00'),\n  ('control seven-judge panel',\n   (['8.0', '8.5', '9.0', '7.5', '8.0', '8.5', '6.0'], '3.1'),\n   '75.95'),\n  ('boundary perfect tens', (['10.0', '10.0', '10.0', '10.0', '10.0'], '3.4'), '102.00'),\n  ('boundary non-half mark', (['7.3', '7.0', '7.0', '7.0', '7.0'], '2.0'), 'invalid mark 7.3'),\n  ('boundary six judges', (['7.0', '7.0', '7.0', '7.0', '7.0', '7.0'], '2.0'), 'invalid panel'),\n  ('regression: mark validation',\n   (['1.5', '8.0', '9.0', '1.5', '7.0', '10.0', '0.0'], '1.6'),\n   '26.40'),\n  ('regression: mark validation',\n   (['7.3', '1.5', '7.0', '1.5', '1.5', '0.0'], '2.8'),\n   'invalid mark 7.3'),\n  ('variant scenario 1', (['8.5', '3.0', '8.5', '2.0', '4.0', '4.0'], '1.6'), 'invalid panel'),\n  ('variant scenario 2', (['8.0', '3.0', '7.5', '3.0', '0.0', '7.5'], '2.0'), 'invalid panel')],\n [('control five-judge panel', (['7.0', '7.5', '6.5', '8.0', '7.0'], '2.0'), '43.00'),\n  ('control seven-judge panel',\n   (['8.0', '8.5', '9.0', '7.5', '8.0', '8.5', '6.0'], '3.1'),\n   '75.95'),\n  ('boundary perfect tens', (['10.0', '10.0', '10.0', '10.0', '10.0'], '3.4'), '102.00'),\n  ('boundary non-half mark', (['7.3', '7.0', '7.0', '7.0', '7.0'], '2.0'), 'invalid mark 7.3'),\n  ('boundary six judges', (['7.0', '7.0', '7.0', '7.0', '7.0', '7.0'], '2.0'), 'invalid panel'),\n  ('regression: mark validation', (['9.0', '4.5', '10.0', '9.5', '4.5'], '3.4'), '78.20'),\n  ('regression: mark validation',\n   (['7.3', '5.0', '4.5', '5.0', '3.5', '7.5', '0.5'], '1.6'),\n   'invalid mark 7.3'),\n  ('variant scenario 1', (['0.0', '6.5', '8.0', '7.5', '5.0'], '2.8'), '53.20'),\n  ('variant scenario 2', (['8.5', '4.5', '1.0', '6.5', '6.5'], '3.4'), '59.50')],\n [('control five-judge panel', (['7.0', '7.5', '6.5', '8.0', '7.0'], '2.0'), '43.00'),\n  ('control seven-judge panel',\n   (['8.0', '8.5', '9.0', '7.5', '8.0', '8.5', '6.0'], '3.1'),\n   '75.95'),\n  ('boundary perfect tens', (['10.0', '10.0', '10.0', '10.0', '10.0'], '3.4'), '102.00'),\n  ('boundary non-half mark', (['7.3', '7.0', '7.0', '7.0', '7.0'], '2.0'), 'invalid mark 7.3'),\n  ('boundary six judges', (['7.0', '7.0', '7.0', '7.0', '7.0', '7.0'], '2.0'), 'invalid panel'),\n  ('regression: mark validation', (['0.0', '1.0', '2.0', '10.0', '8.0'], '1.6'), '17.60'),\n  ('regression: mark validation',\n   (['7.3', '1.0', '4.0', '0.0', '7.0', '0.5', '2.0'], '2.0'),\n   'invalid mark 7.3'),\n  ('variant scenario 1',\n   (['11.0', '2.5', '10.0', '6.0', '2.0', '6.0'], '2.0'),\n   'invalid mark 11.0'),\n  ('variant scenario 2', (['6.0', '1.5', '6.0', '5.5', '10.0'], '1.6'), '28.00')],\n [('control five-judge panel', (['7.0', '7.5', '6.5', '8.0', '7.0'], '2.0'), '43.00'),\n  ('control seven-judge panel',\n   (['8.0', '8.5', '9.0', '7.5', '8.0', '8.5', '6.0'], '3.1'),\n   '75.95'),\n  ('boundary perfect tens', (['10.0', '10.0', '10.0', '10.0', '10.0'], '3.4'), '102.00'),\n  ('boundary non-half mark', (['7.3', '7.0', '7.0', '7.0', '7.0'], '2.0'), 'invalid mark 7.3'),\n  ('boundary six judges', (['7.0', '7.0', '7.0', '7.0', '7.0', '7.0'], '2.0'), 'invalid panel'),\n  ('regression: mark validation',\n   (['10.0', '8.0', '8.0', '9.5', '0.5', '6.0'], '3.1'),\n   'invalid panel'),\n  ('regression: mark validation',\n   (['7.3', '0.5', '6.5', '0.5', '3.0', '0.0', '9.5'], '2.0'),\n   'invalid mark 7.3'),\n  ('variant scenario 1', (['8.0', '0.0', '4.5', '8.0', '5.5'], '2.0'), '36.00'),\n  ('variant scenario 2', (['9.5', '8.0', '3.5', '7.5', '4.0', '9.0'], '3.1'), 'invalid panel')]]\nfor label, args, expected in cases[N - 1]:\n    check(label, run(args), expected)\nprint(json.dumps({\"observations\": observations, \"passed\": all(x[\"passed\"] for x in observations)}, ensure_ascii=False))\nraise SystemExit(0 if all(x[\"passed\"] for x in observations) else 1)\n"},"broken":{"sha256":"393c8e00958cfdcefa084b23cd55219bc31a099c08721b8c03fb0addbf9ec6c2","source":"\"\"\"Failure Map reference implementation. Python standard library only.\"\"\"\nimport json\nfrom fractions import Fraction\nN = 1\nobservations = []\ndef solve(scores, dd):\n    marks = []\n    for s in scores:\n        v = Fraction(s)\n        if v < 0 or v > 10:\n            return 'invalid mark ' + s\n        marks.append(v)\n    if len(marks) == 5:\n        drop = 1\n    elif len(marks) == 7:\n        drop = 2\n    else:\n        return 'invalid panel'\n    kept = sorted(marks)[drop:len(marks) - drop]\n    total = sum(kept) * Fraction(dd)\n    return '%.2f' % float(total)\ndef check(label, actual, expected):\n    observations.append({\"check\": label, \"actual\": actual, \"expected\": expected, \"passed\": actual == expected})\ndef run(args):\n    try:\n        return solve(*args)\n    except Exception as exc:\n        return 'raised ' + type(exc).__name__\ncases = [[('control five-judge panel', (['7.0', '7.5', '6.5', '8.0', '7.0'], '2.0'), '43.00'),\n  ('control seven-judge panel',\n   (['8.0', '8.5', '9.0', '7.5', '8.0', '8.5', '6.0'], '3.1'),\n   '75.95'),\n  ('boundary perfect tens', (['10.0', '10.0', '10.0', '10.0', '10.0'], '3.4'), '102.00'),\n  ('boundary non-half mark', (['7.3', '7.0', '7.0', '7.0', '7.0'], '2.0'), 'invalid mark 7.3'),\n  ('boundary six judges', (['7.0', '7.0', '7.0', '7.0', '7.0', '7.0'], '2.0'), 'invalid panel'),\n  ('regression: mark validation', (['3.5', '10.0', '4.0', '3.5', '7.5'], '2.0'), '30.00'),\n  ('regression: mark validation', (['7.3', '8.0', '3.5', '1.5', '0.5'], '2.0'), 'invalid mark 7.3'),\n  ('variant scenario 1', (['1.5', '8.0', '3.0', '7.0', '7.5'], '3.1'), '54.25'),\n  ('variant scenario 2',\n   (['-1.0', '2.5', '2.0', '3.5', '0.0', '7.5', '6.5'], '2.0'),\n   'invalid mark -1.0')],\n [('control five-judge panel', (['7.0', '7.5', '6.5', '8.0', '7.0'], '2.0'), '43.00'),\n  ('control seven-judge panel',\n   (['8.0', '8.5', '9.0', '7.5', '8.0', '8.5', '6.0'], '3.1'),\n   '75.95'),\n  ('boundary perfect tens', (['10.0', '10.0', '10.0', '10.0', '10.0'], '3.4'), '102.00'),\n  ('boundary non-half mark', (['7.3', '7.0', '7.0', '7.0', '7.0'], '2.0'), 'invalid mark 7.3'),\n  ('boundary six judges', (['7.0', '7.0', '7.0', '7.0', '7.0', '7.0'], '2.0'), 'invalid panel'),\n  ('regression: mark validation',\n   (['1.5', '8.0', '9.0', '1.5', '7.0', '10.0', '0.0'], '1.6'),\n   '26.40'),\n  ('regression: mark validation',\n   (['7.3', '1.5', '7.0', '1.5', '1.5', '0.0'], '2.8'),\n   'invalid mark 7.3'),\n  ('variant scenario 1', (['8.5', '3.0', '8.5', '2.0', '4.0', '4.0'], '1.6'), 'invalid panel'),\n  ('variant scenario 2', (['8.0', '3.0', '7.5', '3.0', '0.0', '7.5'], '2.0'), 'invalid panel')],\n [('control five-judge panel', (['7.0', '7.5', '6.5', '8.0', '7.0'], '2.0'), '43.00'),\n  ('control seven-judge panel',\n   (['8.0', '8.5', '9.0', '7.5', '8.0', '8.5', '6.0'], '3.1'),\n   '75.95'),\n  ('boundary perfect tens', (['10.0', '10.0', '10.0', '10.0', '10.0'], '3.4'), '102.00'),\n  ('boundary non-half mark', (['7.3', '7.0', '7.0', '7.0', '7.0'], '2.0'), 'invalid mark 7.3'),\n  ('boundary six judges', (['7.0', '7.0', '7.0', '7.0', '7.0', '7.0'], '2.0'), 'invalid panel'),\n  ('regression: mark validation', (['9.0', '4.5', '10.0', '9.5', '4.5'], '3.4'), '78.20'),\n  ('regression: mark validation',\n   (['7.3', '5.0', '4.5', '5.0', '3.5', '7.5', '0.5'], '1.6'),\n   'invalid mark 7.3'),\n  ('variant scenario 1', (['0.0', '6.5', '8.0', '7.5', '5.0'], '2.8'), '53.20'),\n  ('variant scenario 2', (['8.5', '4.5', '1.0', '6.5', '6.5'], '3.4'), '59.50')],\n [('control five-judge panel', (['7.0', '7.5', '6.5', '8.0', '7.0'], '2.0'), '43.00'),\n  ('control seven-judge panel',\n   (['8.0', '8.5', '9.0', '7.5', '8.0', '8.5', '6.0'], '3.1'),\n   '75.95'),\n  ('boundary perfect tens', (['10.0', '10.0', '10.0', '10.0', '10.0'], '3.4'), '102.00'),\n  ('boundary non-half mark', (['7.3', '7.0', '7.0', '7.0', '7.0'], '2.0'), 'invalid mark 7.3'),\n  ('boundary six judges', (['7.0', '7.0', '7.0', '7.0', '7.0', '7.0'], '2.0'), 'invalid panel'),\n  ('regression: mark validation', (['0.0', '1.0', '2.0', '10.0', '8.0'], '1.6'), '17.60'),\n  ('regression: mark validation',\n   (['7.3', '1.0', '4.0', '0.0', '7.0', '0.5', '2.0'], '2.0'),\n   'invalid mark 7.3'),\n  ('variant scenario 1',\n   (['11.0', '2.5', '10.0', '6.0', '2.0', '6.0'], '2.0'),\n   'invalid mark 11.0'),\n  ('variant scenario 2', (['6.0', '1.5', '6.0', '5.5', '10.0'], '1.6'), '28.00')],\n [('control five-judge panel', (['7.0', '7.5', '6.5', '8.0', '7.0'], '2.0'), '43.00'),\n  ('control seven-judge panel',\n   (['8.0', '8.5', '9.0', '7.5', '8.0', '8.5', '6.0'], '3.1'),\n   '75.95'),\n  ('boundary perfect tens', (['10.0', '10.0', '10.0', '10.0', '10.0'], '3.4'), '102.00'),\n  ('boundary non-half mark', (['7.3', '7.0', '7.0', '7.0', '7.0'], '2.0'), 'invalid mark 7.3'),\n  ('boundary six judges', (['7.0', '7.0', '7.0', '7.0', '7.0', '7.0'], '2.0'), 'invalid panel'),\n  ('regression: mark validation',\n   (['10.0', '8.0', '8.0', '9.5', '0.5', '6.0'], '3.1'),\n   'invalid panel'),\n  ('regression: mark validation',\n   (['7.3', '0.5', '6.5', '0.5', '3.0', '0.0', '9.5'], '2.0'),\n   'invalid mark 7.3'),\n  ('variant scenario 1', (['8.0', '0.0', '4.5', '8.0', '5.5'], '2.0'), '36.00'),\n  ('variant scenario 2', (['9.5', '8.0', '3.5', '7.5', '4.0', '9.0'], '3.1'), 'invalid panel')]]\nfor label, args, expected in cases[N - 1]:\n    check(label, run(args), expected)\nprint(json.dumps({\"observations\": observations, \"passed\": all(x[\"passed\"] for x in observations)}, ensure_ascii=False))\nraise SystemExit(0 if all(x[\"passed\"] for x in observations) else 1)\n"}},"limitations":"Stipulated, bounded toy contract stated in the contract field; not a claim of conformance with any governing body rulebook or operator house rules. This reproducer isolates one failure mechanism. Results cover the supplied fixtures. Variants within a family share a test contract and should remain grouped when constructing evaluation splits. Related mechanisms with a shared evaluation_group must also remain together; these controlled models are not independent production incidents.","method":"Deterministic executable model with adversarial boundary fixtures.","provenance":{"created_by":"Failure Map","dependencies":"Python standard library","family":"w2-sports-scoring-diving-judges-trim-mark-validation","generated_at":"2026-09-29T14:50:29.181793+00:00","license":"CC0-1.0","python":"3.12.14","seed":1,"split":"open-access"},"relevance":"Meet management systems compute dive scores from judge panels of different sizes.","root_cause":"Validation only checks the 0-10 range, not the half-point grid.","sha256":"3e40edb2d4e44e4cbc62d2b61bb9a348f460bcd4cdb136b36a59b07e4300638b","title":"Judge mark validation accepts off-grid or rejects a perfect ten · case 01","variant":1,"variant_policy":"Five numbered records share a model and may reuse boundary fixtures.","verified":true,"visibility":"public","verification":{"attempt":{"elapsed_ms":45.218,"exit_code":1,"observations":[{"actual":"43.00","check":"control five-judge panel","expected":"43.00","passed":true},{"actual":"75.95","check":"control seven-judge panel","expected":"75.95","passed":true},{"actual":"invalid mark 10.0","check":"boundary perfect tens","expected":"102.00","passed":false},{"actual":"invalid mark 7.3","check":"boundary non-half mark","expected":"invalid mark 7.3","passed":true},{"actual":"invalid panel","check":"boundary six judges","expected":"invalid panel","passed":true},{"actual":"invalid mark 10.0","check":"regression: mark validation","expected":"30.00","passed":false},{"actual":"invalid mark 7.3","check":"regression: mark validation","expected":"invalid mark 7.3","passed":true},{"actual":"54.25","check":"variant scenario 1","expected":"54.25","passed":true},{"actual":"invalid mark -1.0","check":"variant scenario 2","expected":"invalid mark -1.0","passed":true}],"passed":false,"stderr":"","stdout":"{\"observations\": [{\"check\": \"control five-judge panel\", \"actual\": \"43.00\", \"expected\": \"43.00\", \"passed\": true}, {\"check\": \"control seven-judge panel\", \"actual\": \"75.95\", \"expected\": \"75.95\", \"passed\": true}, {\"check\": \"boundary perfect tens\", \"actual\": \"invalid mark 10.0\", \"expected\": \"102.00\", \"passed\": false}, {\"check\": \"boundary non-half mark\", \"actual\": \"invalid mark 7.3\", \"expected\": \"invalid mark 7.3\", \"passed\": true}, {\"check\": \"boundary six judges\", \"actual\": \"invalid panel\", \"expected\": \"invalid panel\", \"passed\": true}, {\"check\": \"regression: mark validation\", \"actual\": \"invalid mark 10.0\", \"expected\": \"30.00\", \"passed\": false}, {\"check\": \"regression: mark validation\", \"actual\": \"invalid mark 7.3\", \"expected\": \"invalid mark 7.3\", \"passed\": true}, {\"check\": \"variant scenario 1\", \"actual\": \"54.25\", \"expected\": \"54.25\", \"passed\": true}, {\"check\": \"variant scenario 2\", \"actual\": \"invalid mark -1.0\", \"expected\": \"invalid mark -1.0\", \"passed\": true}], \"passed\": false}\n"},"broken":{"elapsed_ms":43.995,"exit_code":1,"observations":[{"actual":"43.00","check":"control five-judge panel","expected":"43.00","passed":true},{"actual":"75.95","check":"control seven-judge panel","expected":"75.95","passed":true},{"actual":"102.00","check":"boundary perfect tens","expected":"102.00","passed":true},{"actual":"42.00","check":"boundary non-half mark","expected":"invalid mark 7.3","passed":false},{"actual":"invalid panel","check":"boundary six judges","expected":"invalid panel","passed":true},{"actual":"30.00","check":"regression: mark validation","expected":"30.00","passed":true},{"actual":"24.60","check":"regression: mark validation","expected":"invalid mark 7.3","passed":false},{"actual":"54.25","check":"variant scenario 1","expected":"54.25","passed":true},{"actual":"invalid mark -1.0","check":"variant scenario 2","expected":"invalid mark -1.0","passed":true}],"passed":false,"stderr":"","stdout":"{\"observations\": [{\"check\": \"control five-judge panel\", \"actual\": \"43.00\", \"expected\": \"43.00\", \"passed\": true}, {\"check\": \"control seven-judge panel\", \"actual\": \"75.95\", \"expected\": \"75.95\", \"passed\": true}, {\"check\": \"boundary perfect tens\", \"actual\": \"102.00\", \"expected\": \"102.00\", \"passed\": true}, {\"check\": \"boundary non-half mark\", \"actual\": \"42.00\", \"expected\": \"invalid mark 7.3\", \"passed\": false}, {\"check\": \"boundary six judges\", \"actual\": \"invalid panel\", \"expected\": \"invalid panel\", \"passed\": true}, {\"check\": \"regression: mark validation\", \"actual\": \"30.00\", \"expected\": \"30.00\", \"passed\": true}, {\"check\": \"regression: mark validation\", \"actual\": \"24.60\", \"expected\": \"invalid mark 7.3\", \"passed\": false}, {\"check\": \"variant scenario 1\", \"actual\": \"54.25\", \"expected\": \"54.25\", \"passed\": true}, {\"check\": \"variant scenario 2\", \"actual\": \"invalid mark -1.0\", \"expected\": \"invalid mark -1.0\", \"passed\": true}], \"passed\": false}\n"}},"member_only":{"stages":["fixed"],"fields":["implementations.fixed","verification.fixed","harness","repair"],"note":"The verified repair, its recorded checks, the repair description, and the scoring harness are available to members."}}