ELSEIF
Your brief EB
352 stories from 111 feeds 413 clusters Refreshed 10 minutes ago next pull 17:22

TECH Signal 491

Ornith-1.5 closes the self-improvement loop by generating its own tasks, scaffolds, and solution rollouts

Ornith-1.5 trains itself by generating tasks, scaffolds, and solution rollouts, with models from 9B to 397B parameters.

WHY IT MATTERS

For engineers building on open models, Ornith-1.5 suggests a path where the training data is no longer fixed but continuously generated by the model itself. This could reduce the need for human-curated task sets and hand-designed agent harnesses, though it also means the training distribution is shaped by the model's own capabilities and biases. The reported performance matches or exceeds larger closed models on agentic coding benchmarks, so the approach is worth watching.

Written by elseif from the cluster below · every claim links back to a source

The three things worth knowing

01

Ornith-1.5 spans three scales: 397B MoE, 35B MoE, and 9B dense, with a quantized mobile version for edge deployment.

02

The self-improvement loop jointly optimizes task generation, scaffold construction, and solution rollouts, with reward propagated across all three stages.

03

Ornith-1.5-397B scores 86.1 on Terminal-Bench 2.1 and 56.0 on DeepSWE, matching Claude Opus 4.8 and outperforming GLM-5.2 and DeepSeek-V4-Flash-0731.

THE CLUSTER

Same story, 1 feed.

ORDERED BY FIRST SEEN
ornith.ai via Hacker News Ornith-1.5: From Self-Scaffolding to Self-Improvement Open ↗