Back to feed
Prompt Engineering
Prompt Engineering
6/27/2026
The original headline is: "Ornith 1.0: This is new class of self-improving model"

The original headline is: "Ornith 1.0: This is new class of self-improving model"

Original: Ornith 1.0: This is new class of self-improving model

Short summary

Ornith 1 is a new family of open-weight models trained to generate both code solutions and task-specific harnesses using reinforcement learning, with the 397B variant approaching Opus 4.8 performance. The 9B model achieves 3-20× cost reduction vs. closed-source at comparable accuracy; scale (35B+) required for multi-step reasoning. Creator validates benchmarks on M2 Max and discusses reward-hacking defenses.

  • Ornith generates code solutions + task-specific harnesses in one loop using GRPO reinforcement learning
  • 9B model achieves 3-20× cost reduction vs. closed-source with matching accuracy on some tasks
  • Creator validates on M2 Max hardware; 397B variant approaches Opus 4.8 performance

Generated with AI, which can make mistakes.

Is this a good recommendation for you?

Comments

Failed to load comments. Please try again.

Explore more