# charmbracelet-crush > Writes, edits, tests, and debugs code in a terminal by executing shell commands, reading and modifying files, and using LSP for code intelligence. - repo: https://github.com/charmbracelet/crush - commit: c9b45348a4f1b7695e70830a4408aa4c31688851 - api style: openai - runs: 1 (0 with a reward) - tasks tried: 1 - models: bedrock/us.anthropic.claude-haiku-4-5-20251001-v1:0 - domains: swe - languages: bash, cpp, go, java, javascript, nix, python, rust - capabilities: edits-files, runs-shell, runs-tests, uses-git, reads-docs, long-horizon, calls-apis ## How it runs here Crush (github.com/charmbracelet/crush) is a terminal-based AI coding assistant built in Go. Its CLI entrypoint is `crush run [prompt]` for non-interactive use, which sends the prompt to an LLM and streams the response while the agent uses tools (view, edit, bash, glob, grep, etc.) to complete coding tasks. The harness is built from source using `golang:latest` with `CGO_ENABLED=0 GOEXPERIMENT=greenteagc`. A crushrc config file at `$CRUSH_GLOBAL_CONFIG/crushrc` configures a custom OpenAI-compatible provider pointing to `$PROXY_URL/v1` with model `claude-3-5-sonnet-20241022`, disables the default provider catalog and auto-update to avoid network calls, and pre-approves all tool permissions so the agent can run without user interaction. The run command is `crush run --quiet "$TASK"` from the task working directory. ## Results by task | task | runs | last reward | best reward | last tests | |---|---|---|---|---| | [aider_polyglot/polyglot_cpp_allergies](https://harnessreport.com/tasks/aider_polyglot/polyglot_cpp_allergies.md) | 1 | | | | ## Tests to run next _ranked by llm_ | task | why | |---|---| | [evoeval/14](https://harnessreport.com/tasks/evoeval/14.md) | Start with a straightforward Python implementation task to establish a working baseline after the initial C++ error. | | [aider_polyglot/polyglot_javascript_triangle](https://harnessreport.com/tasks/aider_polyglot/polyglot_javascript_triangle.md) | Test JavaScript, a core language, with a geometry classification problem that's fundamentally simpler than the allergies domain. | | [aider_polyglot/polyglot_rust_bowling](https://harnessreport.com/tasks/aider_polyglot/polyglot_rust_bowling.md) | Assess Rust capability on a complex stateful scoring problem that requires careful algorithmic thinking and comprehensive test validation. | | [quixbugs/quixbugs-java-flatten](https://harnessreport.com/tasks/quixbugs/quixbugs-java-flatten.md) | Evaluate Java support on a focused bug-fixing task with a single-line constraint, testing code analysis and precise modification ability. | | [swebench-verified/sphinx-doc__sphinx-8595](https://harnessreport.com/tasks/swebench-verified/sphinx-doc__sphinx-8595.md) | Test real-world debugging on an established open-source project with git history, test suites, and documentation—exercising the harness's full toolset. | ## Runs | run | task | verifier says | calls | seconds | |---|---|---|---|---| | [20260926T073808-charmbracele-308ef1539f9a-polyglot_cpp_allergies](https://harnessreport.com/runs/20260926T073808-charmbracele-308ef1539f9a-polyglot_cpp_allergies.md) | aider_polyglot/polyglot_cpp_allergies | no reward written | 32 | 68 | --- Harness Report runs agent harnesses from their GitHub repos on Harbor tasks and records every model call. Every page is also `.md` and `.json`; index: https://harnessreport.com/llms.txt · MCP: https://harnessreport.com/mcp