OmniVoice β€” corrected bidirectional ONNX export (int4 variant)

Self-contained int4 variant (smallest download: 4-bit llm_decoder weights, dequantized to fp32 at compute β†’ clean on CUDA; audio_embeddings is bf16, lossless). Backbone at repo root + audio_tokenizer/ (Higgs fp32) + voices/. Siblings: -bf16 (default), -fp32. Bidirectional re-export of k2-fsa/OmniVoice for Sokuji (#351); plain onnxruntime.

License β€” NON-COMMERCIAL

CC-BY-NC-4.0, inherited from k2-fsa/OmniVoice (Emilia training set). Unofficial re-export, not affiliated with/endorsed by the authors. Some re-uploads mislabel it apache-2.0 β€” incorrect for the weights; do not use commercially.

Downloads last month
36
Inference Providers NEW
This model isn't deployed by any Inference Provider. πŸ™‹ Ask for provider support

Model tree for jiangzhuo9357/omnivoice-onnx-bidi-int4

Finetuned
Qwen/Qwen3-0.6B
Finetuned
k2-fsa/OmniVoice
Quantized
(25)
this model