# yc-software-qm

- repo: https://github.com/yc-software/qm
- commit: a61db71b91158d7239047738ae0a07c8c5a829bc
- api style: anthropic
- runs: 3 (None with a reward)
- tasks tried: 1
- models: bedrock/us.anthropic.claude-haiku-4-5-20251001-v1:0

## Results by task

| task | runs | last reward | best reward | last tests |
|---|---|---|---|---|
| [aider_polyglot/polyglot_python_bowling](https://harnessreport.com/tasks/aider_polyglot/polyglot_python_bowling.md) | 3 | 1 |  | 31 passed |

## Runs

| run | task | verifier says | calls | seconds |
|---|---|---|---|---|
| [20260924T180430-yc-software--polyglot_python_bowling](https://harnessreport.com/runs/20260924T180430-yc-software--polyglot_python_bowling.md) | aider_polyglot/polyglot_python_bowling | reward 1 · 31 passed | 15 | 89 |
| [20260924T154338-yc-software--polyglot_python_bowling](https://harnessreport.com/runs/20260924T154338-yc-software--polyglot_python_bowling.md) | aider_polyglot/polyglot_python_bowling | reward 1 · 31 passed | 4 | 20 |
| [20260923T145938-yc-software--polyglot_python_bowling](https://harnessreport.com/runs/20260923T145938-yc-software--polyglot_python_bowling.md) | aider_polyglot/polyglot_python_bowling | reward 0 · 31 failed | 1 | 3 |

---
Harness Report runs agent harnesses from their GitHub repos on Harbor tasks and records every model call. Every page is also `.md` and `.json`; index: https://harnessreport.com/llms.txt · MCP: https://harnessreport.com/mcp
