9B coding agent based on Qwen3.5-9B, fine-tuned on 425K real agentic traces from Claude Opus 4.6, GPT-5.4, and Gemini 3.1. Reads before it writes, traces bugs to the root cause, doesn't clobber your existing code.
16.2K Pulls 3 Tags Updated 5 months ago
Instruct version of the large language model YandexGPT 5 Lite with 8B parameters with a context length of 32k tokens. (quantised version of Q5_K_M)
606 Pulls 3 Tags Updated 12 months ago
Gemma 4 Ollama profiles for RTX 4090/5090 across 12B, 26B-A4B, and 31B variants, with multimodal support and native tool calling
476 Pulls 5 Tags Updated 2 months ago
Multimodal 4.5B Gemma 4 with Unsloth Dynamic QAT & MTP (5.3GB). Features near-lossless vision, audio, text reasoning, an active 4-token speculative window, and an expanded 64K context size. Best for edge execution on systems with 8GB to 16GB RAM.
762 Pulls 1 Tag Updated 1 month ago
Replit Code v1.5 is a 3.3B parameter Causal Language Model focused on Code Completion.
922 Pulls 1 Tag Updated 2 years ago