Institution

TokenRhythm Technologies

AI systems company publishing work on coding-agent benchmarks, harnesses, and evaluation infrastructure.

Claw-SWE-Bench: Why Coding Agent Harnesses Matter

Claw-SWE-Bench evaluates OpenClaw-style coding-agent harnesses on 350 GitHub issue tasks. OpenClaw jumps from 19.1% to 73.4% Pass@1 with a full adapter.