tinymodels
head to head · same machine, same method

Qwen2.5-0.5B-Instruct-GGUF vs SmolLM2-135M-Instruct

Two tiny text-generation models compared on the facts this hub states: what it is, how big it is, and how long one fixed input took on one machine. No benchmark scores from publishers — those live, with their conditions, on each model page.

Qwen/Qwen2.5-0.5B-Instruct-GGUF HuggingFaceTB/SmolLM2-135M-Instruct
tasktext-generationtext-generation
formatggufsafetensors
parameters494M135M
file size491 MB269 MB
licenceapache-2.0apache-2.0
publisherQwenHuggingFaceTB
one fixed input (ms)255.4110.2
downloads2121

The millisecond row is one fixed input per task type, measured on this hub's machine with one warm-up run then one timed run. How these numbers were taken. Saved filenames: qwen2.5-0.5b-instruct-q4_k_m.gguf · SmolLM2-135M-Instruct.safetensors.