r/LocalLLaMA · · 1 min read

If the weights never change, is it really recursive self-improvement?

Mirrored from r/LocalLLaMA for archival readability. Support the source by reading on the original site.

If the weights never change, is it really recursive self-improvement?

https://preview.redd.it/e0ydm43a55kh1.png?width=2902&format=png&auto=webp&s=8c9b88ff5e4157811e8996ba5a1e96cc55c8ae6a

This paper is using a much narrower definition of recursive self-improvement than the phrase usually suggests.

AQuA stores validated evidence in a persistent research state that shapes later hypotheses. The underlying language model and evaluator remain fixed. I still find the narrower claim interesting, even if it sits closer to memory-augmented research automation than to a model rewriting itself.

The paper does not establish any weight-level capability gain. Is persistent memory that improves later research decisions enough to call a system RSI, or should the term require changes to the system’s underlying capabilities?

submitted by /u/derspenti
[link] [comments]

Discussion (0)

Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.

Sign in →

No comments yet. Sign in and be the first to say something.

More from r/LocalLLaMA