Inference
Available
An OpenAI-compatible endpoint served from GPUs co-located with the network points of presence, so a model call from a voice agent is a local hop.
What you get
- OpenAI-compatible API — existing clients work unchanged
- GPUs co-located with the voice edge, in-region
- Per-region pinning for data residency commitments