Speech-to-text, text-to-speech, speaker diarization, speech enhancement, source separation, and VAD using next-gen Kaldi with onnxruntime without Internet connection. Support embedded systems, Android, iOS, HarmonyOS, Raspberry Pi, RISC-V, RK NPU, Axera NPU, Ascend NPU, x86_64 servers, websocket server/client, support 12 programming languages
aarch64
android
arm32
asr
cpp
csharp
dotnet
ios
lazarus
linux
macos
mfc
object-pascal
onnx
raspberry-pi
risc-v
speech-to-text
text-to-speech
vits
windows
Updated 2026-10-09 07:49:02 +00:00