INFERENCE
Every model. One API.
100+ models from every major provider, behind one OpenAI-compatible API. Point your existing code at our base URL and it works.
Run it direct. Or run it guarded.
Every request goes one of two ways. You choose, per request.
CAPABILITIES
Inside a Guarded request.
Guarded is the trust layer: three checkpoints on every request, before, at, and after the model. Direct skips all three.
- 1
Before the model
every request passes a three-layer safety check: fast pattern matching, a Gloo-trained multilingual classifier, and an AI arbiter for the hard calls.
- 2
At the model
your values shape the response: alignment instructions, faith-tradition awareness, and grounding in your content when you ask for it.
- 3
After the model
the response is moderated as it streams. A bad answer gets stopped before your user sees it.
MODEL CATALOG
New models land behind the API you already integrated. No rewrites, no migrations, no keeping up.
Some of the labs behind the catalog.
- OpenAI
- Anthropic
- Google
- Meta
- Mistral AI
- Qwen
- DeepSeek
- Moonshot AI
FAQ
THE REST OF THE PLATFORM