Execution Layer
Optimisation layer that sits between the models and your orchestrator, cutting latency and cost across the STT → LLM → TTS pipeline. Applied on top of your Gateway calls:- Listen (STT) — higher transcription quality
- Think (LLM) — skips redundant calls to save tokens and latency
- Speak (TTS) — caching for lower latency and cost