tasks · tinymodels.co
Tiny image-text-to-text models
Every image-text-to-text model in the catalogue, with the size and licence stated and one measured millisecond figure per model — one fixed input, one machine, one method.
Browse by format instead: gguf, mlx, safetensors.
| model | params | size | licence | one fixed input (ms) | |
|---|---|---|---|---|---|
| Florence-2-base-ft | 232M | 463 MB | mit | 2,124.4 | Get |
The millisecond figure is one fixed input per task type, measured on this hub's machine with one warm-up run then one timed run. How these numbers were taken; the whole run, every model, is the measured bench dataset.