NebulaiSEALED BENCHMARK · HONOJS/HONO · 2026-09-04

21–28% lower agent cost on a real codebase.

Across six questions about Hono, a 451-file open-source framework, Nebulai lowered measured session cost and token use while giving the agent relevant Project Context.

MEASURED-SESSION COST21–28%

lower, across two repeated runs of the same question set

TOKENS13–24%

fewer, same runs

QUESTIONS CHEAPER6 / 6

First run: 13 of 17 successful pairs

Savings compare attempts where both agents passed the recorded grader: 17 pairs covering six questions in the first run, and 15 pairs covering five in the second. Recorded success was 100% versus 94.4%, then 94.4% versus 83.3% (baseline versus Nebulai). All failed attempts missed one literal naming check; factual correctness needs separate review. These figures exclude memory construction and hosting costs and do not measure cross-session discovery reuse.

First run · means over successful pairs · up to three pairs per question
QuestionRounds unaidedSaved
context-response8.0+34.3%
error-propagation9.0+29.5%
middleware-composition10.0+18.6%
request-params15.3+17.4%
smart-router-selection6.0+5.9%
regexp-route-compilation7.0+3.6%

More project context. Less agent overhead.

Nebulai helped the agent reach the relevant parts of the codebase with fewer tokens and lower measured-session cost across both benchmark runs.

That is the outcome Nebulai is built for: engineering teams and coding agents start with the project’s existing understanding instead of paying to reconstruct it.

Nebulai is currently in private beta. We’re opening early access to teams that want their software projects to remember.

Request early access