# litellm-labs-lite-harness - repo: https://github.com/LiteLLM-Labs/lite-harness - commit: dd99cfdfc68dbb6b3f7f986d54efd42572373a6c - api style: anthropic - runs: 1 (None with a reward) - tasks tried: 1 - models: bedrock/us.anthropic.claude-haiku-4-5-20251001-v1:0 ## Results by task | task | runs | last reward | best reward | last tests | |---|---|---|---|---| | [aider_polyglot/polyglot_python_bowling](https://harnessreport.com/tasks/aider_polyglot/polyglot_python_bowling.md) | 1 | 0 | | 27 failed, 4 passed | ## Runs | run | task | verifier says | calls | seconds | |---|---|---|---|---| | [20260923T135024-litellm-labs-polyglot_python_bowling](https://harnessreport.com/runs/20260923T135024-litellm-labs-polyglot_python_bowling.md) | aider_polyglot/polyglot_python_bowling | reward 0 · 27 failed, 4 passed | 8 | 27 | --- Harness Report runs agent harnesses from their GitHub repos on Harbor tasks and records every model call. Every page is also `.md` and `.json`; index: https://harnessreport.com/llms.txt · MCP: https://harnessreport.com/mcp