A strong Mixture-of-Experts (MoE) language model with 671B total parameters with 37B activated for each token.
3.8M Pulls 5 Tags Updated 1 year ago
A general-purpose model ranging from 3 billion parameters to 70 billion, suitable for entry-level hardware.
3M Pulls 119 Tags Updated 2 years ago
Atom-Olmo3-7B is a specialized language model fine-tuned from Olmo-3 7B Instruct for collaborative problem-solving and creative exploration.
222 Pulls 1 Tag Updated 9 months ago
Single file version with (Dynamic Quants) A strong Mixture-of-Experts (MoE) language model with 671B total parameters with 37B activated for each token.
117 Pulls 4 Tags Updated 1 year ago
A 7B math reasoning model from Allen AI, trained with RL-Zero to solve problems step-by-step like a skilled tutor. Supports 65K context for complex multi-step problems - runs on any laptop.
260 Pulls 7 Tags Updated 9 months ago
olmo 3.7b instruct x starcoder, small but packs a punch
35 Pulls 1 Tag Updated 2 months ago
StarCoder2 is the next generation of transparently trained open code LLMs that comes in three sizes: 3B, 7B and 15B parameters.
3M Pulls 67 Tags Updated 1 year ago
7,196 Pulls 2 Tags Updated 1 year ago
4,094 Pulls 5 Tags Updated 1 year ago
(Unsloth Dynamic Quants) A strong Mixture-of-Experts (MoE) language model with 671B total parameters with 37B activated for each token.
2,064 Pulls 3 Tags Updated 1 year ago
25 Pulls 1 Tag Updated 1 year ago