949 Pulls 1 Tag Updated 2 weeks ago
178 Pulls 1 Tag Updated 2 weeks ago
Nanbeige4.1-3B illustrates that compact models can simultaneously achieve robust reasoning, preference alignment, and effective agentic behaviors
7,035 Pulls 6 Tags Updated 5 months ago
3B model that shouldn't be this good - crushes benchmarks through deep chain-of-thought reasoning
1,801 Pulls 1 Tag Updated 5 months ago
26.2.27. Nanbeige4.1-3B-Q4 illustrates that compact models can simultaneously achieve robust reasoning, preference alignment, and effective agentic behaviors.
807 Pulls 1 Tag Updated 5 months ago
Fine-tuned version of Nanbeige 4.1 3B specialized for Python code generation with direct, focused output.
750 Pulls 3 Tags Updated 5 months ago
Nanbeige4.1-3B-q4_K_M no think tools fit 4G-6G GPU OpenClaw local free tokens LobsterAI
444 Pulls 1 Tag Updated 4 months ago
Ollama version of de-censored Nanbeige4.1-3B-heretic
400 Pulls 1 Tag Updated 5 months ago
Optimized its system prompt and parameters for not to overthink too much.
337 Pulls 1 Tag Updated 5 months ago
nanbeige4.1-3b-tools no think fit 4G-6G GPU OpenClaw local free tokens
212 Pulls 1 Tag Updated 4 months ago
107 Pulls 1 Tag Updated 4 months ago
59 Pulls 1 Tag Updated 5 months ago
36 Pulls 1 Tag Updated 4 months ago
32 Pulls 1 Tag Updated 4 months ago
The Nanbeige2-16B-Chat is the latest 16B model developed by the Nanbeige Lab, which utilized 4.5T tokens of high-quality training data during the training phase.
147 Pulls 1 Tag Updated 2 years ago
Model şu an beta aşamasındadır. 1M parametrenin getirdiği sınırlardan dolayı karmaşık cümle yapılarında bozulmalar yaşanabilir. Geliştirme süreci devam etmektedir.
22 Pulls 3 Tags Updated 6 months ago
Instinct is Continue's state-of-the-art open Next Edit model. Robustly fine-tuned from Qwen2.5-Coder-7B, Instinct intelligently predicts your next move to keep you in flow.
6,737 Pulls 1 Tag Updated 11 months ago
A 62M-parameter GPT trained from scratch on a single 8GB-RAM NVIDIA Jetson, offering a compact open-weights base model for raw text completion. Not instruction-tuned.
1 Pull 1 Tag Updated 1 week ago
Llama-3.1-Nemotron-70B-Instruct is a large language model customized by NVIDIA to improve the helpfulness of LLM generated responses to user queries.
3,052 Pulls 6 Tags Updated 1 year ago