732 2 hours ago

Chirp Chirp! 馃惁 We are introducing Ornith-1.5, a major step toward building foundation models through end-to-end self-improvement.

vision 9b 35b 397b
ollama run ornith-1.5:9b

Details

3 hours ago

e5df7dcdd8a2 路 6.6GB 路

qwen35
8.95B
Q4_K_M
clip
456M
BF16

Readme

ornith_logo.png

Chirp Chirp! 馃惁 We are introducing Ornith-1.5, a major step toward building foundation models through end-to-end self-improvement.

Ornith-1.5 extends Ornith-1.0 by expanding the self-improvement loop from scaffold and rollout optimization to jointly optimizing task generation, scaffold construction, and solution rollouts. Rather than relying on a fixed set of human-curated tasks and manually designed harnesses, Ornith-1.5 continuously generates new training tasks, discovers effective strategies for solving them, and improves the policy through reinforcement learning. For more details on the task, harness, and rollout reward design, please refer to our blog.

Ornith -1.5-397B

ornith_397b_eval.png

Ornith -1.5-35B-A3B

ornith_35b_eval.png

Ornith-1.5-9B

ornith_9b_eval.png