# google-gemini-gemini-cli - repo: https://github.com/google-gemini/gemini-cli - commit: bedef96ef42905bd84a86dbec021c706168e7e2f - api style: openai - runs: 6 (None with a reward) - tasks tried: 5 - models: bedrock/us.anthropic.claude-haiku-4-5-20251001-v1:0 ## Results by task | task | runs | last reward | best reward | last tests | |---|---|---|---|---| | [aider_polyglot/polyglot_go_palindrome-products](https://harnessreport.com/tasks/aider_polyglot/polyglot_go_palindrome-products.md) | 1 | 0 | | | | [aider_polyglot/polyglot_python_bowling](https://harnessreport.com/tasks/aider_polyglot/polyglot_python_bowling.md) | 2 | 1 | | 31 passed | | [bird-bench/card_games__372](https://harnessreport.com/tasks/bird-bench/card_games__372.md) | 1 | 1 | | | | [humanevalfix/python-12](https://harnessreport.com/tasks/humanevalfix/python-12.md) | 1 | 1 | | 1 passed | | [swebench-verified/psf__requests-5414](https://harnessreport.com/tasks/swebench-verified/psf__requests-5414.md) | 1 | 0 | | 1 failed, 130 passed, 1 xfailed, 158 errors | ## Runs | run | task | verifier says | calls | seconds | |---|---|---|---|---| | [20260924T232100-sw-gemini-cli-card_games__372](https://harnessreport.com/runs/20260924T232100-sw-gemini-cli-card_games__372.md) | bird-bench/card_games__372 | reward 1 | 4 | 15 | | [20260924T231732-sw-gemini-cli-python-12](https://harnessreport.com/runs/20260924T231732-sw-gemini-cli-python-12.md) | humanevalfix/python-12 | reward 1 · 1 passed | 5 | 19 | | [20260924T230944-sw-gemini-cli-psf__requests-5414](https://harnessreport.com/runs/20260924T230944-sw-gemini-cli-psf__requests-5414.md) | swebench-verified/psf__requests-5414 | reward 0 · 1 failed, 130 passed, 1 xfailed, 158 errors · verifier exited 1 | 32 | 290 | | [20260924T225606-sw-gemini-cli-polyglot_go_palindrome-products](https://harnessreport.com/runs/20260924T225606-sw-gemini-cli-polyglot_go_palindrome-products.md) | aider_polyglot/polyglot_go_palindrome-products | reward 0 | 31 | 264 | | [20260924T154427-google-gemin-polyglot_python_bowling](https://harnessreport.com/runs/20260924T154427-google-gemin-polyglot_python_bowling.md) | aider_polyglot/polyglot_python_bowling | reward 1 · 31 passed | 23 | 121 | | [20260923T153226-google-gemin-polyglot_python_bowling](https://harnessreport.com/runs/20260923T153226-google-gemin-polyglot_python_bowling.md) | aider_polyglot/polyglot_python_bowling | reward 0 · 31 failed | 30 | 286 | --- Harness Report runs agent harnesses from their GitHub repos on Harbor tasks and records every model call. Every page is also `.md` and `.json`; index: https://harnessreport.com/llms.txt · MCP: https://harnessreport.com/mcp