tinymodels
head to head · same machine, same method

Qwen2.5-0.5B-Instruct-4bit vs Qwen2.5-0.5B-Instruct-GGUF

Two tiny text-generation models compared on the facts this hub states: what it is, how big it is, and how long one fixed input took on one machine. No benchmark scores from publishers — those live, with their conditions, on each model page.

mlx-community/Qwen2.5-0.5B-Instruct-4bit Qwen/Qwen2.5-0.5B-Instruct-GGUF
tasktext-generationtext-generation
formatmlxgguf
parameters494M494M
file size278 MB491 MB
licenceapache-2.0apache-2.0
publishermlx-communityQwen
one fixed input (ms)241.4255.4
downloads2021

The millisecond row is one fixed input per task type, measured on this hub's machine with one warm-up run then one timed run. How these numbers were taken. Saved filenames: Qwen2.5-0.5B-Instruct-4bit.safetensors · qwen2.5-0.5b-instruct-q4_k_m.gguf.