Technical deep dives on the self-hosted AI gateway — pre-model optimization, prompt cache, and spend attribution.
September 27, 2026
Anyray sits in front of the model: a self-hosted gateway that compresses unused LLM context, keeps prompt cache stable, attributes spend, and fails open.
Dean Rubin · CTO