head to head · same machine, same method
Qwen2.5-0.5B-Instruct vs Qwen2.5-0.5B-Instruct-4bit
Two tiny text-generation models compared on the facts this hub states: what it is, how big it is, and how long one fixed input took on one machine. No benchmark scores from publishers — those live, with their conditions, on each model page.
| Qwen/Qwen2.5-0.5B-Instruct | mlx-community/Qwen2.5-0.5B-Instruct-4bit | |
|---|---|---|
| task | text-generation | text-generation |
| format | safetensors | mlx |
| parameters | 494M | 494M |
| file size | 988 MB | 278 MB |
| licence | apache-2.0 | apache-2.0 |
| publisher | Qwen | mlx-community |
| one fixed input (ms) | 266.8 | 241.4 |
| downloads | 16 | 20 |
The millisecond row is one fixed input per task type, measured on this hub's machine with one warm-up run then one timed run. How these numbers were taken. Saved filenames: Qwen2.5-0.5B-Instruct.safetensors · Qwen2.5-0.5B-Instruct-4bit.safetensors.