ONNX
GGUF
imatrix
conversational

Dictto offline models

Model files used by the Dictto offline dictation app. They are downloaded by the app at first run (not bundled in the APK). This repo aggregates models under two licenses — see below.

Contents & licenses

Path Model Upstream source License
speech/core/ Parakeet TDT 0.6B v3 (PT + EN), ONNX int8 nvidia/parakeet-tdt-0.6b-v3, ONNX/int8 export by k2-fsa/sherpa-onnx CC-BY-4.0
speech/en-stream/ NeMo streaming fast-conformer transducer EN (1040 ms) nvidia/stt_en_fastconformer_hybrid_large_streaming_1040ms CC-BY-4.0
ai/ Qwen3-1.7B (abliterated) GGUF Q4_K_L Qwen/Qwen3-1.7Bmlabonne/Qwen3-1.7B-abliterated → GGUF by bartowski Apache-2.0

Attribution

  • Speech models © NVIDIA, licensed under CC-BY-4.0. Changes: exported to ONNX and quantized to int8 by the k2-fsa/sherpa-onnx project.
  • AI model: Qwen3-1.7B by the Qwen team, licensed under Apache-2.0; "abliterated" fine-tune by mlabonne; GGUF quantization by bartowski.

All three models permit commercial and non-commercial use.

Downloads last month
1,068
GGUF
Model size
5B params
Architecture
gemma4
Hardware compatibility
Log In to add your hardware

2-bit

8-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support