Alexis Rondeau

Newsfeed

Reply · Tuesday, July 7, 2026 · 16:54

3 likes · 315 views · 1 bookmark
In reply to @lilianweng · Jul 7, 2026

new post on harness engineering for AI self-improvement: t.co/ZYvGfVs61k

It is hard to forecast how much the future of RSI will rely on harnesses. Likely harness engineering will evolve in the direction of self-improvement and enable auto-research, and, in turn, smarter models keeps harnesses simple.

Even when many harness improvement get eventually internalized into core model, the need to specify goals and context will not disappear.

@lilianweng Yup! I believe auto-research loops will be part of it. And kind of are already. I experimented with a Karpathy-style DIY research loop on a personal quest and, yeah, this stuff works like today.

PS: The thing I made:

alexisrondeau.me"Which city in Paris are you staying in?" - A post-training study on small-model specialization with an autoresearch loopAcross 77 experiments — every lever from prompting and retrieval to fine-tuning and 2025's rubric-grounded RL — an on-device 3B model landed within about one rubric point of a frontier cloud model on every quality dimension: close, but consistently a notch below. None of the post-training playbook closed the last notch, and the controlled tests show why — the gap is a broad capability limit, not a…↗ alexisrondeau.me