Measured, not marketed
Every number below comes from controlled same-model A/B runs — the same brain, with and without Libra OS — certified across three repetitions. The gain is the system layer, not a bigger model.
| Result | What it means | |
|---|---|---|
| GAIA — public agentic benchmark | +18.9 pts | A full model generation of improvement, from the orchestration layer alone |
| Grounded scale | 256 employees per 8×H100 node, 98.3% grounded success | We measure grounded answers, not HTTP 200s |
| Token economy | 23× fewer tokens per task (~27K vs ~615K) | Under a tenth of the cost per answer |
| AgentDojo — prompt-injection defense | attack success 26.8% → 17.3%, benign utility held at 90.7% | Hardened without making the system useless |
Full methodology and per-run data available on request — contact@meganova.ai.