ScaffBench
Measuring coding agents on real fullstack scaffolding tasks — time, tokens, cost, and whether the result actually builds.
TokensOutput
Avg output tokens generated per scaffold run.
CostUSD
Avg API cost per scaffold in USD.
StepsActions
Avg tool steps and commands executed per run.
CodeLoC
Avg lines of code written per scaffold, lockfiles excluded.
ScaffBench 3
Full pass rate
Leaderboard
Scored across 13 specs on a clean machine. Higher pass rate is better; lower cost, time, and tokens are better.
PassPrimary
Installs, builds, type-checks, and clears all quality gates cold.
WiredLibs
Required spec libraries present and imported.
StatsMean / run
Mean time, cost, tokens, steps, and LoC.
Give your agent the fast path.
One MCP server, every spec-to-scaffold tool the benchmark used. Pick your agent, paste, done.
all supported clients$ claude mcp add --transport stdio better-fullstack -- npx -y create-better-fullstack@latest mcprun in your terminal