20 3 days ago

Ornith-1.5-9B quantized to IQ2_M (3.77 GB) — the most aggressive compression available for this model. Ideal for low-VRAM GPUs. Runs with a 16k context window while leaving 2.5 GB of VRAM free. Slightly faster at the cost of some precision

vision
ollama run madkoding/ornith-1.5-9b-iq2m

Models

View all →

Readme

No readme