ornith-1.5:397b

745 2 hours ago

Chirp Chirp! 馃惁 We are introducing Ornith-1.5, a major step toward building foundation models through end-to-end self-improvement.

vision 9b 35b 397b
ollama run ornith-1.5:397b

Details

2 hours ago

337c6db6be12 路 242GB 路

qwen35moe
396B
Q4_K_M
clip
456M
BF16

Readme

ornith_logo.png

Chirp Chirp! 馃惁 We are introducing Ornith-1.5, a major step toward building foundation models through end-to-end self-improvement.

Ornith-1.5 extends Ornith-1.0 by expanding the self-improvement loop from scaffold and rollout optimization to jointly optimizing task generation, scaffold construction, and solution rollouts. Rather than relying on a fixed set of human-curated tasks and manually designed harnesses, Ornith-1.5 continuously generates new training tasks, discovers effective strategies for solving them, and improves the policy through reinforcement learning. For more details on the task, harness, and rollout reward design, please refer to our blog.

Ornith -1.5-397B

ornith_397b_eval.png

Ornith -1.5-35B-A3B

ornith_35b_eval.png

Ornith-1.5-9B

ornith_9b_eval.png