Happy, fun, helper robot. LLMs work for hermes, openclaw, and picoclaw. Make new frens and be nice to everyone.
-
qwen3.5-hermes
A compact, agent-tuned Qwen 3.5 model built for local agent harnesses like Hermes Agent. Based on `qwen3.5:latest` with Q4_K_M quantization for a lightweight 6.6 GB footprint, while retaining strong tool-calling, vision, and thinking capabilities.
vision tools thinking1,495 Pulls 1 Tag Updated 1 month ago
-
qwen3.6-hermes
Qwen 3.6 27B (Q4_K_M) optimized for Hermes Agent — 64K context, 8192 max tokens, MTP for speed, flash attention + Q8 KV cache.
vision tools thinking718 Pulls 1 Tag Updated 1 month ago
-
gemma4
Gemma 4 12B agent-tuned for Hermes Agent and other local agent harnesses. One built for 20GB GPUs the other for 8GB GPUs.
vision tools thinking716 Pulls 2 Tags Updated 1 month ago
-
gemma4-26b-a4b-it-qat-agent
Gemma 4 26B tuned for agentic use. 64K context window, flash attention + Q8 KV cache quantization for reduced VRAM. Temperature 0.7, output capped at 8192 tokens.
vision tools thinking221 Pulls 1 Tag Updated 1 month ago
-
gemma4-31b-hermes
Gemma 4 31B (Q4_0) optimized for Hermes Agent — 64K context, 8192 max tokens, flash attention + Q8 KV cache. Fits 24 GB VRAM with optimizations.
vision tools thinking161 Pulls 1 Tag Updated 1 month ago
-
lightning-agent
Lightning Agent — a 32.9B Nemotron MoE model tuned for Hermes agentic workflows.
15 Pulls 1 Tag Updated 1 week ago
-
muse-glimmer-agent
a 27.9B vision-language model tuned for Hermes and other agentic workflows.
vision11 Pulls 1 Tag Updated 1 week ago
-
gemma4-256kvision tools thinking
9 Pulls 1 Tag Updated 3 weeks ago