# Run 20260924T185619-dlants-magen-polyglot_python_bowling - harness: [dlants-magenta.nvim](https://harnessreport.com/harnesses/dlants-magenta.nvim.md) @ 5af1a2139a26 - task: [aider_polyglot/polyglot_python_bowling](https://harnessreport.com/tasks/aider_polyglot/polyglot_python_bowling.md) - model: bedrock/us.anthropic.claude-haiku-4-5-20251001-v1:0 - verifier says: reward 0 · 3 failed, 28 passed - model calls: 12 tokens in/out: 2656/5856 - agent seconds: 49 - started: 2026-09-24T18:58:49 - full bundle (calls, logs, files): https://harnessreport.com/api/run/20260924T185619-dlants-magen-polyglot_python_bowling ## Failed tests - test_cannot_roll_if_game_already_has_ten_frames - test_the_second_bonus_rolls_after_a_strike_in_the_last_frame_cannot_be_a_strike_if_the_first_one_is_not_a_strike - test_two_bonus_rolls_after_a_strike_in_the_last_frame_cannot_score_more_than_10_points Last agent action: `bash_command: cd /app && python3 -c " from bowling import BowlingGame # Test three consecutive strikes game = BowlingGame() game.roll(…` --- Harness Report runs agent harnesses from their GitHub repos on Harbor tasks and records every model call. Every page is also `.md` and `.json`; index: https://harnessreport.com/llms.txt · MCP: https://harnessreport.com/mcp