Lorqa transcription โ€” Fast (320 ms)

Byte-identical mirror of aufklarer/Nemotron-3.5-ASR-Streaming-0.6B-CoreML-INT8 at commit 447095fe87b480b5e6a15367f135303d479de8ac. Based on NVIDIA Nemotron 3.5 ASR streaming 0.6B. Core ML INT8 encoder, FP16 decoder and joint; 16 kHz mono PCM. All weights and language/tokenizer/configuration files are unchanged.

Download: 642,196,943 bytes of model/runtime data (642 MB decimal / 612 MiB). Runtime memory is higher than download size. 320 ms is chunk duration, not guaranteed end-to-end latency. Chinese accuracy needs evaluation on the intended recordings.

Languages and prompt slots: see languages.json. Auto detection and explicit BCP-47 language hints are supported by the companion runtime. Do not use this bundle with FluidAudio's 2240 ms manager; the graph and cache layouts differ.

Runtime reference: soniqo/speech-swift (Apache-2.0). Model weights retain OpenMDW 1.1, with upstream attribution retained. This repository is a distribution mirror, not a retraining or accuracy claim.

The Balanced (2240 ms) model remains in Lorqa/nemotron-3.5-asr-streaming-multilingual-0.6b-coreml, under multilingual/2240ms.

Downloads last month
15
Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support