Model Catalog
4 items
Applied Filters
ggml-org /
InternVL3-14B-Instruct-GGUF
Image-Text-to-Text
Llama.cpp
Accelerated llama.cpp
Q8_0
GPU 1x Nvidia L4
$ 0.8
ggml-org /
Qwen2.5-VL-3B-Instruct-GGUF
Image-Text-to-Text
Llama.cpp
Accelerated llama.cpp
Q8_0
GPU 1x Nvidia T4
$ 0.5
ggml-org /
Qwen2.5-VL-7B-Instruct-GGUF
Image-Text-to-Text
Llama.cpp
Accelerated llama.cpp
Q8_0
GPU 1x Nvidia T4
$ 0.5
ggml-org /
SmolVLM2-2.2B-Instruct-GGUF
Image-Text-to-Text
Llama.cpp
Accelerated llama.cpp
F16
GPU 1x Nvidia T4
$ 0.5