Press Space to continue
Finding signal on Twitter is more difficult than it used to be. We curate the best tweets on topics like AI, startups, and product development every weekday so you can focus on what matters.
Press Space to continue
Press Space to continue
It works by "replaying its past discovery attempts" while "testing thousands of alternative strategies cheaply", then "deploying the better strategy in the next round" and "it improves the exploration policy, not the underlying model weights." ...so RSI is prompt engineering lol
big AI news Google just demonstrated a recursive self improvement loop for AI discovery Google/DeepMind researchers introduced Dream-RSI, a system where an AI agent improves how it explores problems by replaying its past discovery attempts, testing thousands of alternative strategies cheaply, then deploying the better strategy in the next round. Across algorithm design, mathematical optimization, and GPU kernel engineering, it matched or improved discovery quality while cutting search costs dramatically, in one setting reducing agent calls by up to 162x. 👀 Importantly, it improves the exploration policy, not the underlying model weights.