1/ Last week, Kensa read through a production agent's traces and caught a silent bug regressing the system.
1 in 4 of its generated summaries were truncating mid-sentence (max output tokens set to 150!)
That bug is one of many reasons we're releasing a new version today.
1/ Last week, Kensa read through a production agent's traces and caught a silent bug regressing the system.
1 in 4 of its generated summaries were truncating mid-sentence (max output tokens set to 150!)
That bug is one of many reasons we're releasing a new version today.
6/ The evals are files you own and commit, nothing leaves your repository. And Kensa remains open source.
Paste this in Claude Code, Codex, or Cursor and let it wire itself:
"Fetch https://t.co/pKNx3M8Oc4 and follow it"
People of the world, friends and relatives - prepare yourself for the next version of Kensa, built from the ground up to make your agents even more reliable.
Drops on 22nd.
Introducing kensa.
Tell your coding agent to eval any agent: It reads your code, sets up telemetry, finds failure modes, writes scenarios and judges.
kensa CLI + skills orchestrates it.
You review the diffs, not the scaffold.
Open source & MIT license, link below.