3 likes · 315 views · 1 bookmark
In reply to @lilianweng · Jul 7, 2026new post on harness engineering for AI self-improvement: t.co/ZYvGfVs61k
It is hard to forecast how much the future of RSI will rely on harnesses. Likely harness engineering will evolve in the direction of self-improvement and enable auto-research, and, in turn, smarter models keeps harnesses simple.
Even when many harness improvement get eventually internalized into core model, the need to specify goals and context will not disappear.
@lilianweng Yup! I believe auto-research loops will be part of it. And kind of are already. I experimented with a Karpathy-style DIY research loop on a personal quest and, yeah, this stuff works like today.
PS: The thing I made:
alexisrondeau.me"Which city in Paris are you staying in?" - A post-training study on small-model specialization with an autoresearch loopAcross 77 experiments — every lever from prompting and retrieval to fine-tuning and 2025's rubric-grounded RL — an on-device 3B model landed within about one rubric point of a frontier cloud model on every quality dimension: close, but consistently a notch below. None of the post-training playbook closed the last notch, and the controlled tests show why — the gap is a broad capability limit, not a…↗ alexisrondeau.me