From conversation to insight
Speaker-aware transcripts, workflows, playback, and a searchable final record. Follow a session from start to finish.
Download product tourGuarding your sensitive conversations.
nanosamur.ai is a complete open-source speech AI platform for organisations that cannot send sensitive conversations to a third party. Capture, transcribe, refine, and process speech entirely inside infrastructure you control.
Speaker-aware transcripts, workflows, playback, and a searchable final record. Follow a session from start to finish.
Download product tourQwen3-ASR and faster-whisper, running side by side. Compare their live transcripts as the same conversation unfolds.
Read the model comparisonFollow the conversation live, then improve the transcript automatically as more context becomes available.
Speaker-aware transcripts, word timings, and synchronized playback make conversations easier to review and verify.
Send transcripts to your own AI workflows, applications, and webhooks during or after a session.
Store recordings, refined transcripts, and final session records for search, replay, and downstream processing.
Use the browser interface or Windows desktop application to capture, review, and replay conversations.
Connect through REST, realtime WebSockets, the Python SDK, Kafka events, workflows, and webhooks.
Distributed processing stages, multitenancy, and persistent records provide a foundation for shared services.
Follow a session across the stack with correlated metrics, logs, and distributed traces.
Run the complete platform on-premises, in a private cloud, or on Kubernetes using infrastructure operated by your organisation.
Process speech without cloud APIs, including in networks with restricted or no external connectivity.
Transcribe interviews, briefings, meetings, and operational debriefs inside controlled infrastructure, including isolated networks.
Capture consultations and clinical discussions while keeping audio and transcripts inside infrastructure governed by your organisation.
Clone the Apache-2.0 repository and start the complete platform locally with Docker Compose.
cp .env.example .env docker compose pull docker compose up -d
Not just a model or SDK. The browser UI, desktop application, APIs, orchestration, speech services, and deployment configuration are available as open source.
Capture, transcription, storage, and downstream processing can all run inside infrastructure operated by your organisation.
Keep recordings, transcripts, application data, and operational telemetry under your organisation’s policies and controls.
Clone the repository and start the complete speech AI stack on your own hardware.