37.1% lower API proxy cost across larger tasks.
ContextGuard keeps Codex on its normal shell workflow while injecting evidence once, compacting eligible noisy output locally, and preventing successful results from creating another retrieval roundtrip.
Four layers between your shell and Codex, local-first with evidence preserved.
Three larger maintenance tasks, six counterbalanced pairs, GPT-5.6 Luna on low reasoning. The pooled sample cut total tokens by 21.6% and the API cost proxy by 37.1%; individual task medians still vary.
API costs are an OpenAI Standard API proxy calculated from measured per-turn usage. Codex subscription billing is not exposed by the CLI. The rate card includes GPT-6 Astra for comparison.
Six paired GPT-5.6 Luna runs at low reasoning · GPT-6 Astra Standard: $10/M input, $1/M cached, $50/M output for short context · long context: $20/M, $2/M, $75/M.
Install from the GimingerConsulting/ContextGuard marketplace source, supported on macOS, Linux, and Windows wherever Codex runs. Start a new thread after setup.