"""Agent Olympics — a hackathon-competition benchmark inside hack-house. This package layers a fourth benchmark axis on top of ``bench/``: teams of LLM agents deliberate in a real hack-house room, implement code in an isolated VM, and are scored on correctness/speed/quality/collaboration. See SPEC.md for the full design. M1 is the arena spine (one same-model team, one MBPP problem, real-room/local-bus deliberation -> driver implement -> PodmanRuntime -> hidden tests -> replayable transcript -> deterministic score). """