Modules§
- anthropic
- cache_
aligner - Cache-aligner (#940 detect, #974 relocate) — Headroom “cache aligner” port.
- cache_
attribution - Prompt-cache miss attribution (#986, cache-economics telemetry).
- cache_
breakpoint - Active prompt-cache breakpoint injection (#939, Headroom “cache aligner” adjacent).
- cache_
policy - Net-cost policy for cache-busting rewrites (#986, cache-economics).
- cache_
safety - Cache-preservation telemetry for the proxy’s frozen-region prose rewrites (#710).
- ccr
- Content-addressed recovery (CCR) for the proxy’s lossy rewrites (#482).
- chatgpt
- chatgpt_
cookies - chatgpt_
ws - WebSocket passthrough for ChatGPT’s
/backend-apirail (#597). - cold_
prefix - Big-gap cold-prefix repack prediction (#480).
- compress
- compress_
api POST /v1/compress— deterministic messages-in / messages-out compression.- cost
- Per-model proxy savings accounting.
- effort
- Cache-safe, cross-provider reasoning-effort control (#834).
- forward
- history_
prune - holdout
- Deterministic output-savings holdout (#895 Track B).
- introspect
- metrics
- openai
- openai_
responses - openai_
responses_ ws - WebSocket bridge for the OpenAI/Codex Responses transport (#440).
- output_
savings - Output-token savings reporting (#895 Track B).
- prose
- Frozen-region prose compression for the proxy (#710).
- prose_
ranker - Cache-safe wire prose squeeze (#895).
- tool_
kind - Classifies what produced a
tool_resultso the proxy never lossy-compresses a file/source-code read the model still needs (e.g. mid-refactor). - tool_
output - usage
- Real provider-reported token usage extraction.
- usage_
meter - Measured per-model spend meter.
- verbosity
- Cache-safe wire verbosity steer (#895 Track B).