# litellm-labs-lite-harness

- repo: https://github.com/LiteLLM-Labs/lite-harness
- commit: dd99cfdfc68dbb6b3f7f986d54efd42572373a6c
- api style: anthropic
- runs: 1 (None with a reward)
- tasks tried: 1
- models: bedrock/us.anthropic.claude-haiku-4-5-20251001-v1:0

## Results by task

| task | runs | last reward | best reward | last tests |
|---|---|---|---|---|
| [aider_polyglot/polyglot_python_bowling](https://harnessreport.com/tasks/aider_polyglot/polyglot_python_bowling.md) | 1 | 0 |  | 27 failed, 4 passed |

## Runs

| run | task | verifier says | calls | seconds |
|---|---|---|---|---|
| [20260923T135024-litellm-labs-polyglot_python_bowling](https://harnessreport.com/runs/20260923T135024-litellm-labs-polyglot_python_bowling.md) | aider_polyglot/polyglot_python_bowling | reward 0 · 27 failed, 4 passed | 8 | 27 |

---
Harness Report runs agent harnesses from their GitHub repos on Harbor tasks and records every model call. Every page is also `.md` and `.json`; index: https://harnessreport.com/llms.txt · MCP: https://harnessreport.com/mcp
