How to use from
Unsloth Studio
Install Unsloth Studio (macOS, Linux, WSL)
curl -fsSL https://unsloth.ai/install.sh | sh
# Run unsloth studio
unsloth studio -H 0.0.0.0 -p 8888
# Then open http://localhost:8888 in your browser
# Search for llmware/qwen-3-vl-8b-gguf to start chatting
Install Unsloth Studio (Windows)
irm https://unsloth.ai/install.ps1 | iex
# Run unsloth studio
unsloth studio -H 0.0.0.0 -p 8888
# Then open http://localhost:8888 in your browser
# Search for llmware/qwen-3-vl-8b-gguf to start chatting
Using HuggingFace Spaces for Unsloth
# No setup required
# Open https://huggingface.co/spaces/unsloth/studio in your browser
# Search for llmware/qwen-3-vl-8b-gguf to start chatting
Quick Links

qwen3-vl-8b-gguf

qwen3-vl-8b-gguf is a GGUF Q4_K_M quantized version of Qwen3-VL-8B-Instruct providing a fast, small inference implementation, optimized for AI PCs.

Model Description

  • Developed by: Qwen
  • Quantized by: unsloth
  • Model type: qwen3vl
  • Parameters: 8 billion
  • Model Parent: Qwen/Qwen3-VL-8B-Instruct
  • Language(s) (NLP): English
  • License: Apache 2.0
  • Uses: Chat, general-purpose LLM
  • Quantization: int4

Model Card Contact

llmware on hf

llmware website

Downloads last month
20
GGUF
Model size
8B params
Architecture
qwen3vl
Hardware compatibility
Log In to add your hardware

4-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for llmware/qwen-3-vl-8b-gguf

Quantized
(82)
this model