# google-gemini-gemini-cli

- repo: https://github.com/google-gemini/gemini-cli
- commit: bedef96ef42905bd84a86dbec021c706168e7e2f
- api style: openai
- runs: 6 (None with a reward)
- tasks tried: 5
- models: bedrock/us.anthropic.claude-haiku-4-5-20251001-v1:0

## Results by task

| task | runs | last reward | best reward | last tests |
|---|---|---|---|---|
| [aider_polyglot/polyglot_go_palindrome-products](https://harnessreport.com/tasks/aider_polyglot/polyglot_go_palindrome-products.md) | 1 | 0 |  |  |
| [aider_polyglot/polyglot_python_bowling](https://harnessreport.com/tasks/aider_polyglot/polyglot_python_bowling.md) | 2 | 1 |  | 31 passed |
| [bird-bench/card_games__372](https://harnessreport.com/tasks/bird-bench/card_games__372.md) | 1 | 1 |  |  |
| [humanevalfix/python-12](https://harnessreport.com/tasks/humanevalfix/python-12.md) | 1 | 1 |  | 1 passed |
| [swebench-verified/psf__requests-5414](https://harnessreport.com/tasks/swebench-verified/psf__requests-5414.md) | 1 | 0 |  | 1 failed, 130 passed, 1 xfailed, 158 errors |

## Runs

| run | task | verifier says | calls | seconds |
|---|---|---|---|---|
| [20260924T232100-sw-gemini-cli-card_games__372](https://harnessreport.com/runs/20260924T232100-sw-gemini-cli-card_games__372.md) | bird-bench/card_games__372 | reward 1 | 4 | 15 |
| [20260924T231732-sw-gemini-cli-python-12](https://harnessreport.com/runs/20260924T231732-sw-gemini-cli-python-12.md) | humanevalfix/python-12 | reward 1 · 1 passed | 5 | 19 |
| [20260924T230944-sw-gemini-cli-psf__requests-5414](https://harnessreport.com/runs/20260924T230944-sw-gemini-cli-psf__requests-5414.md) | swebench-verified/psf__requests-5414 | reward 0 · 1 failed, 130 passed, 1 xfailed, 158 errors · verifier exited 1 | 32 | 290 |
| [20260924T225606-sw-gemini-cli-polyglot_go_palindrome-products](https://harnessreport.com/runs/20260924T225606-sw-gemini-cli-polyglot_go_palindrome-products.md) | aider_polyglot/polyglot_go_palindrome-products | reward 0 | 31 | 264 |
| [20260924T154427-google-gemin-polyglot_python_bowling](https://harnessreport.com/runs/20260924T154427-google-gemin-polyglot_python_bowling.md) | aider_polyglot/polyglot_python_bowling | reward 1 · 31 passed | 23 | 121 |
| [20260923T153226-google-gemin-polyglot_python_bowling](https://harnessreport.com/runs/20260923T153226-google-gemin-polyglot_python_bowling.md) | aider_polyglot/polyglot_python_bowling | reward 0 · 31 failed | 30 | 286 |

---
Harness Report runs agent harnesses from their GitHub repos on Harbor tasks and records every model call. Every page is also `.md` and `.json`; index: https://harnessreport.com/llms.txt · MCP: https://harnessreport.com/mcp
