# Model Agent Env Runtime Score Cost Trials
synth-scaffold harbor-sandbox synth-router 88.0 $0.17 4.0k 2 harbor-react docker/py3.12 vLLM 0.9 82.0 $0.26 3.5k 3 OpenHands docker/py3.12 SGLang 76.0 $0.35 3.0k 4 SWE-agent e2b TGI 3.0 70.0 $0.44 2.4k 5 aider modal vLLM 0.9 64.0 $0.53 1.9k 6 harbor-react harbor-sandbox SGLang 58.0 $0.62 1.3k Ranked by resolve rate from verified trials across every submitted model, agent, environment, and runtime combination. Cost and tokens are per-attempt averages.
Model Synth-R1 32B
Agent synth-scaffold
Environment harbor-sandbox
Runtime synth-router 88.0 Score
$0.17 Cost
14.0k Tokens
How ranking works Each model, agent, environment, and runtime combination needs at least 50 verified trials to appear. Ties break on cost, then tokens. Scores refresh as new trials land.