doggy8088/Qwen3-ASR-1.7B-MLX-fp16

This repo contains an MLX conversion of Qwen/Qwen3-ASR-1.7B.

Variant

  • Format: MLX safetensors
  • Weights dtype: float16
  • Custom architecture helper: qwen3_asr_mlx.py

Files

  • model*.safetensors: converted MLX weights
  • config.json: converted model config
  • preprocessor_config.json, tokenizer_config.json, vocab.json, merges.txt: tokenizer / processor assets copied from the source model
  • qwen3_asr_mlx.py: custom MLX architecture + processor helper used for this conversion

Notes

  • Source model: Qwen/Qwen3-ASR-1.7B
  • Local conversion directory: Qwen3-ASR-1.7B-MLX-fp16
  • This is an Apple Silicon / MLX-oriented conversion.
  • For quantized variants, quantizable linear layers are quantized while non-quantizable layers remain in floating point.
Downloads last month
16
Safetensors
Model size
2B params
Tensor type
F16
·
MLX
Hardware compatibility
Log In to add your hardware

Quantized

Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for doggy8088/Qwen3-ASR-1.7B-MLX-fp16

Finetuned
(116)
this model