Enterprise On-Premise

Run the whole loop inside your own network.

The same stack, deployed to your own infrastructure. Speech, inference and rendering stay inside your perimeter, and no conversation data leaves the network.

Deploy

Helm chart, Docker Compose

Inference

Any OpenAI-compatible endpoint

Egress

None required at runtime

Audit

File, syslog or OTLP

01

Air-Gapped Operation

No outbound calls are required at runtime, and model weights are mounted from your own registry.

02

Your Own Models

Any OpenAI-compatible endpoint, plus a local speech-to-text and text-to-speech provider. Swap providers per organization without redeploying.

03

Deploys As Containers

A Helm chart and a plain Compose file are both supported. A GPU node is only required for rendering and local inference.

04

Audit Trail

Every utterance, tool call and credential grant is recorded to your own sink, whether that is a file, syslog or OTLP.

Have a question about this surface?

We read everything you send us.

Talk to Us