llava-hf / image-text-to-text updated 1 year ago

llava-onevision-qwen2-0.5b-ov-hf

Check out also the Google Colab demo to run Llava on a free-tier Google Colab instance:

Params
890 M
Context
Downloads 30d
1.1 M
Likes
57
Commercial use: allowed apache-2.0 Not gated SAFETENSORS 2 languages View on Hugging Face ↗

Download history

daily snapshots · 10 days
1.1 M1.0 M
Jul 28Jul 31Aug 3Aug 6

Can you run it?

Estimated VRAM at 8K context unless noted. Pick your hardware to see the verdict per quantization.

FileQuantSizeEst. VRAMVerdict on RTX 4090 · 24 GB
model.safetensors f16 1.8 GB 2.6 GB ✅ Runs comfortably
Estimate: file size × 1.1 + KV cache at 8K + 0.5 GB overhead. Not a benchmark — how we calculate this.

Specifications

Architecture
LlavaOnevisionForConditionalGeneration
Parameters
890 M
Tensor type
F16
Licence
apache-2.0
First seen on the Hub
2024-08-13
Training datasets
lmms-lab/LLaVA-OneVision-Data
Added to our catalog
2026-07-28