Comfortable on the Q4 GGUF + projector profile.
Edition: Uncensored edition by HauHau
- Model size
- 9B dense · 262K max
- Model format
- Q4 GGUF + projector
- Runs with
- llama.cpp
The Q4 weights are 5.3 GB; vision adds a separate projector and long context adds KV cache.