head to head · same machine, same method
Qwen2.5-0.5B-Instruct-4bit vs Qwen3-0.6B
Two tiny text-generation models compared on the facts this hub states: what it is, how big it is, and how long one fixed input took on one machine. No benchmark scores from publishers — those live, with their conditions, on each model page.
| mlx-community/Qwen2.5-0.5B-Instruct-4bit | Qwen/Qwen3-0.6B | |
|---|---|---|
| task | text-generation | text-generation |
| format | mlx | safetensors |
| parameters | 494M | 752M |
| file size | 278 MB | 1.50 GB |
| licence | apache-2.0 | apache-2.0 |
| publisher | mlx-community | Qwen |
| one fixed input (ms) | 241.4 | 469.6 |
| downloads | 20 | 20 |
The millisecond row is one fixed input per task type, measured on this hub's machine with one warm-up run then one timed run. How these numbers were taken. Saved filenames: Qwen2.5-0.5B-Instruct-4bit.safetensors · Qwen3-0.6B.safetensors.