# Run 20260924T154403-swe-agent-sw-polyglot_python_bowling - harness: [swe-agent-swe-agent](https://harnessreport.com/harnesses/swe-agent-swe-agent.md) @ 3ea751c087f3 - task: [aider_polyglot/polyglot_python_bowling](https://harnessreport.com/tasks/aider_polyglot/polyglot_python_bowling.md) - model: bedrock/us.anthropic.claude-haiku-4-5-20251001-v1:0 - verifier says: reward 0 · 10 failed, 21 passed - model calls: 42 tokens in/out: 395498/10332 - agent seconds: 108 - started: 2026-09-24T15:44:05 - full bundle (calls, logs, files): https://harnessreport.com/api/run/20260924T154403-swe-agent-sw-polyglot_python_bowling ## Failed tests - test_a_roll_cannot_score_more_than_10_points - test_bonus_roll_after_a_strike_in_the_last_frame_cannot_score_more_than_10_points - test_cannot_roll_after_bonus_roll_for_spare - test_cannot_roll_after_bonus_rolls_for_strike - test_cannot_roll_if_game_already_has_ten_frames - test_rolls_cannot_score_negative_points - test_second_bonus_roll_after_a_strike_in_the_last_frame_cannot_score_more_than_10_points - test_the_second_bonus_rolls_after_a_strike_in_the_last_frame_cannot_be_a_strike_if_the_first_one_is_not_a_strike - test_two_bonus_rolls_after_a_strike_in_the_last_frame_cannot_score_more_than_10_points - test_two_rolls_in_a_frame_cannot_score_more_than_10_points Last agent action: `submit: {}` --- Harness Report runs agent harnesses from their GitHub repos on Harbor tasks and records every model call. Every page is also `.md` and `.json`; index: https://harnessreport.com/llms.txt · MCP: https://harnessreport.com/mcp