SakThai Vision 7B ๐Ÿ–ผ๏ธ

Multimodal vision model (LLaVA 1.5 7B) quantized to GGUF Q4_K_M. Understands images and answers questions about them.

Usage

llama-cli -m llava-1.5-7b-hf-q4_k_m.gguf --mmproj mmproj-model-f16.gguf \
  -p "Describe this image" --image photo.jpg

Specs

Property Value
Base model LLaVA 1.5 7B
Format GGUF Q4_K_M
Size 3.9 GB
RAM needed 8 GB+
Use case Image understanding, visual QA

SakThai Model Family

Model Type Downloads
sakthai-embedding sentence-similarity 28
sakthai-context-0.5b-tools text-generation 7
sakthai-context-0.5b-merged text-generation 785
sakthai-context-1.5b-tools text-generation 115
sakthai-context-1.5b-merged text-generation 942
sakthai-context-7b-tools text-generation 147
sakthai-context-7b-merged text-generation 534
sakthai-context-7b-128k text-generation 324
sakthai-coder-1.5b text-generation 15

๐Ÿ‘‰ Full Collection

Downloads last month
-
GGUF
Model size
7B params
Architecture
llama
Hardware compatibility
Log In to add your hardware

4-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Space using Nanthasit/sakthai-vision-7b 1

Collection including Nanthasit/sakthai-vision-7b