--- license: cc-by-4.0 language: - en tags: - asr - speech - coreml - nemo - parakeet - nvidia - int8-per-channel-symmetric library_name: coremltools pipeline_tag: automatic-speech-recognition base_model: nvidia/parakeet-ctc-0.6b-Vietnamese --- # parakeet-ctc-0.6b-vi-coreml-int8 CoreML conversion of [nvidia/parakeet-ctc-0.6b-Vietnamese](https://huggingface.co/nvidia/parakeet-ctc-0.6b-Vietnamese) — INT8 PER CHANNEL SYMMETRIC quantized. | | | |---|---| | **Architecture** | CTC | | **Language** | English | | **Sample rate** | 16000 Hz | | **Max audio** | 15.0s | | **Vocab size** | 1024 | | **Framework** | NVIDIA NeMo → CoreML (coremltools) | ## Components | File | Component | Best compute | |------|-----------|--------------| | `parakeet_mel_encoder.mlpackage` | mel_encoder | ANE / GPU | | `parakeet_ctc_decoder.mlpackage` | ctc_decoder | ANE / GPU | ## Usage ```bash pip install ovos-stt-plugin-coreml ``` ```python from ovos_stt_plugin_coreml import CoremlSTT from ovos_plugin_manager.utils.audio import AudioFile stt = CoremlSTT(config={"metadata": "metadata.json"}) with AudioFile("speech.wav") as f: audio = f.read() print(stt.execute(audio)) ``` ## Source model [nvidia/parakeet-ctc-0.6b-Vietnamese](https://huggingface.co/nvidia/parakeet-ctc-0.6b-Vietnamese)