Qwen2.5-0.5B-Instruct β€” CoreML (ANE+GPU Optimized)

Converted from Qwen/Qwen2.5-0.5B-Instruct for on-device inference on Apple devices.

File Size Description
model.mlpackage 302 MB Monolithic decoder with stateful KV cache (int4)
  • HF-exact match: "The capital of France is Paris." βœ…
  • iOS 18+ required (MLState API)

See CoreML-LLM for full details.


More models in this format: Core ML Model Zoo β€” 46 models, each with the recipe that produced it.

Want a different model on-device? Open a request β€” free, open weights only; the export and its measured numbers get published publicly.

Downloads last month
18
Inference Providers NEW
This model isn't deployed by any Inference Provider. πŸ™‹ Ask for provider support

Model tree for mlboydaisuke/qwen2.5-0.5b-coreml

Quantized
(267)
this model

Collection including mlboydaisuke/qwen2.5-0.5b-coreml