StarWorkshop/wavrail
Streaming speech recognition toolkit: log-mel features, CTC greedy and prefix-beam decoding, transducer-style frame-synchronous emission, forced alignment, Chinese/English CER-WER metrics, and a chunked streaming pipeline with latency accounting. NumPy core with a CPU-only torch extra; trainable demos on synthetic audio, fully offline.
GitHub repository with 31 stars and 235 forks.
Language: Python
Topics: asr, cer, ctc, forced-alignment, log-mel, numpy, python, speech-recognition, streaming, transducer