سيناريوهات إصلاح الأخطاء
كل تشغيل يبدأ من اختبار فاشل قابل للتكرار وينتهي بالتحقق.
Measured across 87 bug-fix tasks on identical checkouts: Codna verified 87/87 (100%) at ~$0.02 per fix — 5× fewer tokens and 1.7× faster than Cursor, and 6.6× cheaper than Cline (which verified only 73.6%).
| Metric | Codna | Cursor | Cline |
|---|---|---|---|
| Fixes verified | 87/87 | 87/87 | 64/87 |
| Accuracy | 100% | 100% | 73.6% |
| Avg wall time | 13.4s | 22.6s | 72.9s |
| Avg tokens / fix | 16.2K | 81.0K | 64.8K |
| Avg cost / fix | $0.021 | n/a | $0.139 |
| Repo understanding | 0 tokens | LLM context | LLM context |
المنهجية
تشمل مجموعة المعايير أخطاء منفردة وأخطاء متعددة في مستودع واحد ومستودعات لم يرها الوكيل. كل سيناريو يُحسب فقط عند نجاح الاختبار.
كل تشغيل يبدأ من اختبار فاشل قابل للتكرار وينتهي بالتحقق.
Codna فهم 130 مستودعًا عبر 110 لغة في 9.2 ثانية.
Codna يُرسل للوكيل حزمة أدلة بدلًا من المستودع بالكامل.
Measured spend to ship one verified fix — about 28× cheaper than a typical agentic edit.
Measured 2026-06-15 on identical isolated checkouts. Every scenario counts only when your own test passes. Download the raw dataset (JSON) →
Codna ran head-to-head against OpenAI Codex CLI and Google Gemini CLI across 8 real bug-fix scenarios. Every fix was verified by the project's own tests. No synthetic tasks, no self-reported numbers.
A fix counts only when the test suite passes. Codna went 8 for 8. A fix that compiles but breaks tests is not counted.
Codna sends the AI agent an evidence bundle measured at roughly 600 tokens — 162x less context than reading the full repository. That translates directly to lower API costs, typically pennies per fix.
The deterministic engine maps the dependency and blast-radius graph in about 60 ms using zero LLM tokens. The agent receives only the relevant evidence, so there is far less to process before a fix is produced.
Yes. In measured testing, Codna mapped 130 repositories in 9.2 seconds, using zero tokens for the mapping step.
The benchmark compared Codna against the default configurations of OpenAI Codex CLI and Google Gemini CLI as available at time of testing. Codna is bring-your-own-key, so you choose the underlying model.
codna fix . --issue "the failing test"