granite3.2-vision

A compact and efficient vision-language model, specifically designed for visual document understanding, enabling automated content extraction from tables, charts, infographics, plots, diagrams, and more.

Görüntü Araç Kullanımı 2b
Hızlı Kurulum (Ollama kuruluysa)
ollama run granite3.2-vision

Ollama kurulu değil mi? ollama.com/download — Windows, macOS ve Linux için ücretsiz. İlk çalıştırmada model indirilir, sonrası tamamen çevrimdışıdır.

Varyantlar

Boyut büyüdükçe kalite artar, donanım ihtiyacı yükselir. Başlangıç için küçük varyantı deneyin.

EtiketBoyutBağlamGirdiKomut
latest 2.4GB 16K Text, Image ollama run granite3.2-vision:latest
2b 2.4GB 16K Text, Image ollama run granite3.2-vision:2b
2b-q4_K_M 2.4GB 16K Text, Image ollama run granite3.2-vision:2b-q4_K_M
2b-q8_0 3.6GB 16K Text, Image ollama run granite3.2-vision:2b-q8_0
2b-fp16 6.0GB 16K Text, Image ollama run granite3.2-vision:2b-fp16

Model Detayları ve Benchmarklar (kaynak: ollama.com)

Note: this model requires Ollama 0.5.13.

A compact and efficient vision-language model, specifically designed for visual document understanding, enabling automated content extraction from tables, charts, infographics, plots, diagrams, and more. The model was trained on a meticulously curated instruction-following dataset, comprising diverse public datasets and synthetic datasets tailored to support a wide range of document understanding and general image tasks. It was trained by fine-tuning a Granite large language model with both image and text modalities.

References

Hugging Face