# aaif-goose-goose - repo: https://github.com/aaif-goose/goose - commit: ac6222029fe526a4ac15540cb69e2e49cf3a8878 - api style: anthropic - runs: 2 (None with a reward) - tasks tried: 1 - models: bedrock/us.anthropic.claude-haiku-4-5-20251001-v1:0 ## Results by task | task | runs | last reward | best reward | last tests | |---|---|---|---|---| | [aider_polyglot/polyglot_python_bowling](https://harnessreport.com/tasks/aider_polyglot/polyglot_python_bowling.md) | 2 | 1 | | 31 passed | ## Runs | run | task | verifier says | calls | seconds | |---|---|---|---|---| | [20260924T154034-aaif-goose-g-polyglot_python_bowling](https://harnessreport.com/runs/20260924T154034-aaif-goose-g-polyglot_python_bowling.md) | aider_polyglot/polyglot_python_bowling | reward 1 · 31 passed | 16 | 79 | | [20260923T115527-aaif-goose-g-polyglot_python_bowling](https://harnessreport.com/runs/20260923T115527-aaif-goose-g-polyglot_python_bowling.md) | aider_polyglot/polyglot_python_bowling | reward 0 · 31 failed | 8 | 10 | --- Harness Report runs agent harnesses from their GitHub repos on Harbor tasks and records every model call. Every page is also `.md` and `.json`; index: https://harnessreport.com/llms.txt · MCP: https://harnessreport.com/mcp