Platform · Runtime and models
The models are ours.
The compute is yours.
Speech, language and voice models ship with the runtime and run on CPU inside your boundary. No GPU, and no third-party endpoint in the call path.
One runtime, every surface.
Voice AI Agents, Real-Time Translation, Knowledge Mesh and Agent Assist ship together on one deployment, and Voice Identity Transformation and AI Observability run on the same runtime.
Voice AI Agents
Inbound and outbound calls, long context, and a handoff to a person mid-sentence.
Ships togetherReal-Time Translation
Simultaneous translation on a live call, with accent normalisation on both sides.
Ships togetherKnowledge Mesh
Retrieval and governance: comprehensive in, salient out, permission-filtered and cited.
Ships togetherAgent Assist
The same cited answer on the human agent's desktop while the call is still live.
Ships together
The models ship with the runtime.
Speech, language and voice models are built and versioned here, deployed with the runtime, and moved on your change schedule rather than on a vendor's.
Speech
Recognition and synthesis on the call itself, with no audio sent anywhere else for either.
In the deploymentLanguage
The reasoning that reads the caller's intent, chooses the tool and writes the reply.
In the deploymentVoice
The spoken output, and the identity work that holds every agent to one brand standard.
In the deployment
Models, at work
One turn of a call, and the three models it passes through.
- The caller speaks, and Speech hears it
- Language reads intent and picks the tool
- Voice speaks the reply in one brand voice
- The audio leaves on the trunk it came in on
What the drawing marks
- Your trunk: The carrier your calls use today
- On the call itself: No audio goes anywhere else
- Versioned with the runtime: Moves on your change schedule
- Your deployment: Every model loads from inside it
Voicing consensus router 1.7% LS-clean, 3.1% LS-other, 5.8% Open ASR avg; 150 ms to a final transcript; real-time factor 0.05X on CPUEnglish word error rate, quantised on-premise, no GPU
The caller speaks, and Speech hears it
Voicing consensus router 1.7% LS-clean, 3.1% LS-other, 5.8% Open ASR avg; 150 ms to a final transcript; real-time factor 0.05X on CPUEnglish word error rate, quantised on-premise, no GPU
Language reads intent and picks the tool
97% function-calling accuracy on transcribed speech, graded schema-strict across an eight-tool suite2,800 voice-transcribed calls
Voice speaks the reply in one brand voice
Voicing TTS router MOS 4.46, RTF 0.04x, TTFA 38 msUTMOS benchmark, internal locale eval, Aug 2026
The audio leaves on the trunk it came in on
700 ms end-to-end response latency at the 80th percentileBFSI, on-premises, real telephony, 80th percentile, 2026
Inference runs on hardware you own.
The runtime is optimised for CPU, so a deployment is sized on servers an enterprise already runs instead of on a GPU estate it has to buy first.
CPU-optimised
Sized on ordinary server cores, so a deployment does not wait on specialised hardware.
No GPUInside the boundary
Every model in the call path is loaded from your own deployment, never from an endpoint.
No egressOn your schedule
Model versions move when your change board says they move, and roll back the same way.
Versioned
Inference, under the hood
Five techniques that take the GPU out of the call path.
Real-time factor
0.05X
on CPU, quantised, no GPU required
It joins the systems the call already touches.
Telephony, CRM and the systems of record stay where they are. The runtime reaches them from inside your network, with credentials your own team issues and revokes.
Telephony
The numbers, carriers and contact-centre platform your calls already route through.
Your carrierSystems of record
CRM, core banking, policy and scheduling systems, read and written under your permissions.
Your credentialsPeople
A handoff to a person mid-call, with the transcript and the state of the call attached.
Your team
Integration, by name
The systems a call already touches, reached from inside your network.
Telephony
Genesys
NICE CXone
Avaya
Your carriers and the contact-centre platform you run.
Systems of record
Salesforce
Epic EHR
Oracle BRM
Read and written under permissions your team controls.
People
Five9
WhatsApp Business
ServiceNow ITSM
A mid-call handoff, with transcript and state attached.
And it does not end here.
System marks belong to their owners.
A documented API
Every action the agent takes, as a call.
An MCP server
Your tools, over the Model Context Protocol.
Webhooks
Events pushed to whatever listens.
Bring one call type. Leave with an architecture.

Six surfaces on one runtime
A working session with an engineer who has deployed inside a bank’s perimeter. We map your telephony, data boundary and handoff rules, and tell you what we would not automate.