Hanami

Loading...

I increasingly think recursive self-improvement could challenge the current RL-centric paradigm.RL improves a model by u | Hanami