Product
Oxagen, the agent control plane
More
Research Field manual Docs Get a demo

Research · Pillar

Self-improving models

Models that train on their own output.

Bootstrapped reasoning, self-rewarding training, and the hard limits the research keeps finding: model collapse and the myth of unaided self-correction.

2 posts in Self-improving models