Skip to content

INFERENCE

Every model. One API.

100+ models from every major provider, behind one OpenAI-compatible API. Point your existing code at our base URL and it works.

Run it direct. Or run it guarded.

Every request goes one of two ways. You choose, per request.

Direct

provider prices

  • The same models, at the price the provider charges. We add nothing.
  • Built for pipelines, batch jobs, agents and internal tools: anywhere you need fast, reliable model access with failover already underneath.

Guarded

Provider price + platform fee.

  • The same request, passed through the trust layer before and after the model: safety checks tuned for real human moments, alignment to your values, moderation of what comes back, citations when grounded.
  • Built for anything your users will see: chat experiences, member care, youth-facing tools, and especially moments of care, crisis, or conviction.

CAPABILITIES

  • Failover built in

    direct providers first, health-checked fallbacks behind them, automatic recovery.

  • Prompt caching

    cache-aware billing with TTL you control.

  • Grounded completions

    answers cited to your sources, structured for your UI.

  • Web search, cited

    current answers with URL citations, server-side.

  • Multimodal

    vision in, image generation out, tools throughout.

  • Streaming everywhere

    SSE with keepalives and truncated-stream detection.

Inside a Guarded request.

Guarded is the trust layer: three checkpoints on every request, before, at, and after the model. Direct skips all three.

  1. 1

    Before the model

    every request passes a three-layer safety check: fast pattern matching, a Gloo-trained multilingual classifier, and an AI arbiter for the hard calls.

  2. 2

    At the model

    your values shape the response: alignment instructions, faith-tradition awareness, and grounding in your content when you ask for it.

  3. 3

    After the model

    the response is moderated as it streams. A bad answer gets stopped before your user sees it.

MODEL CATALOG

New models land behind the API you already integrated. No rewrites, no migrations, no keeping up.

Some of the labs behind the catalog.

  • OpenAI
  • Anthropic
  • Google
  • Meta
  • Mistral AI
  • Qwen
  • DeepSeek
  • Moonshot AI

FAQ

THE REST OF THE PLATFORM

Build with Trusted AI.