RT @TrapitBansal: Congratulations Levent, Tristan, and OpenAI! What a miraculous time to be alive!
I view this as the first successful achievement of Recursive Self Improvement or RSI. Models get increasingly better at Math and get increasingly used by mathematicians to solve all kinds of difficult problems in their domain, thereby contributing very rare and very hard tokens that are then used in the model training to further improve the model's math capabilities: generalizing and drawing connections across sessions and everything else the models learn from; which is then served back, and so on and so forth.
This happened first for math but will eventually happen for every domain.
引用推文
We congratulate Levent Alpöge and Tristan Buckmaster on their remarkable mathematical work.
We (the researchers and the agents) did not see any of their work through any means until they released it publicly — in particular, no specific user data was accessed in order to solve this problem.
While unlikely, we cannot rule out that de-identified data derived from their usage of our products helped improve our models.
However, our proofs differ significantly and even the precise results proved are different in the Euler case (forced vs. unforced).