Soniqo · Private speech intelligence

Private speech intelligence.
In your cloud, on-premise, or on-device.

Speaker-separated transcription, automated conversation QA, real-time guidance for operators and sales representatives, and speech models adapted to your organization. Runs in an approved cloud, a private VPC, or on your premises — recordings and transcripts stay inside the environment you choose.

Deploy anywhere

Cloud, on-prem, or on-device — same models.

Start on the managed API. Move to your own infrastructure when compliance demands it. Or run on the edge with the open-source SDKs. The models travel with you.

Cloud API

Managed and metered per audio-second. Sign in, get a key, and call the endpoint — nothing to host. This is what you are looking at.

Sign in
On-premise

Run the whole stack inside your own network — audio and transcripts never leave it, for regulated, air-gapped, and data-residency requirements. Speaker-separated transcription and automated conversation QA, real-time operator and sales guidance, and speech models adapted to your language, geography, vocabulary, and workflows.

On-device SDK

Ship speech on the edge with our open-source SDKs — speech-swift for Apple platforms, speech-core for Linux / Windows / macOS / Android — with model adaptations tuned to your hardware. No network round-trip. Wake words and voice commands run fully offline. A CPU-only Android build runs the full pipeline in 1.2 GB of RAM — part of our private speech runtime, available on request. On Linux, our proprietary runtime runs a complete command loop in under 400 MB of RAM, for commercial deployments.

Watch the demo
Model adaptation

Adapt our own models to your organization, on our own model stack: Steno transcription tuned to your language, vocabulary and workflows, and Soniqo Full Duplex — full-duplex speech-to-speech pipelines running on our own runtime. Delivered with on-premise and private deployments.

What's included

Everything you need to ship voice features.

Diarization, dialect coverage, and drop-in compatibility with the OpenAI audio API.

Speakers, attributed

Every utterance comes back labelled with a speaker. Register speaker profiles once and get stable named identities across meetings, interviews, and call recordings.

Languages and dialects

25+ languages with regional-dialect coverage on the core model, plus a long-tail model spanning 1,600+ languages (Hindi, Arabic, Indonesian, Vietnamese, and more). You set the language on each request.

OpenAI-compatible

Code written against the OpenAI audio API works against Soniqo by changing one line — the base URL. Send your key as a Bearer token or X-Api-Key. No client rewrite, no SDK migration.