Useful from the first word
Follow the conversation live, then improve the transcript automatically as more context becomes available.
Guarding your sensitive conversations.
nanosamur.ai is a complete open-source speech AI platform for organisations that cannot send sensitive conversations to a third party. Capture, transcribe, refine, and process speech entirely inside infrastructure you control.
Watch nanosamur.ai turn a live conversation into speaker-aware transcripts, workflow results, a searchable final record, and a fully traced session.
Follow the conversation live, then improve the transcript automatically as more context becomes available.
Speaker-aware transcripts, word timings, and synchronized playback make conversations easier to review and verify.
Send transcripts to your own AI workflows, applications, and webhooks during or after a session.
Store recordings, refined transcripts, and final session records for search, replay, and downstream processing.
Use the browser interface or Windows desktop application to capture, review, and replay conversations.
Connect through REST, realtime WebSockets, the Python SDK, Kafka events, workflows, and webhooks.
Distributed processing stages, multitenancy, and persistent records provide a foundation for shared services.
Follow a session across the stack with correlated metrics, logs, and distributed traces.
Run the complete platform on-premises, in a private cloud, or on Kubernetes using infrastructure operated by your organisation.
Process speech without cloud APIs, including in networks with restricted or no external connectivity.
Transcribe interviews, briefings, meetings, and operational debriefs inside controlled infrastructure, including isolated networks.
Capture consultations and clinical discussions while keeping audio and transcripts inside infrastructure governed by your organisation.
Clone the Apache-2.0 repository and start the complete platform locally with Docker Compose.
cp .env.example .env docker compose pull docker compose up -d
Not just a model or SDK. The browser UI, desktop application, APIs, orchestration, speech services, and deployment configuration are available as open source.
Capture, transcription, storage, and downstream processing can all run inside infrastructure operated by your organisation.
Keep recordings, transcripts, application data, and operational telemetry under your organisation’s policies and controls.
Clone the repository and start the complete speech AI stack on your own hardware.