
🤗 Upvotes: 33 | cs.AI Authors: Haozhe Liu, Tian Ye, Sensen Gao, Qihang Cao, Yitong Li, Mingchen Zhuge, Duomin Wang, Ruihua Zhang, Ping Luo, Jiawang Bian, Lei Zhu, Ligeng Zhu, Enze Xie, Song Han Title: SoL-Pi: Recursively Scaling Auto-Research Loops for Efficient Agent Harness Arxiv: http://arxiv.org/abs/2609.20519v1 Abstract: As coding agents move from supervised code completion to unattended, around-the-clock exploration, their work expands from isolated predictions into long trajectories of reasoning, tool use, and feedback. Token efficiency therefore becomes important for scaling recursive self-improvement. We take an RSI-inspired approach at the harness layer, scaling auto-research loops across increasingly numerous and diverse environments for harness rollouts. At this scale, the process yields reusable improvements that transfer beyond their development setting, moving automated harness discovery toward production-level outcomes. Four mechanisms survive selection and form SoL-Pi, spanning action execution, context compaction, observation handling, and delegated reading. On the 51-task EdgeBench evaluation, SoL-Pi achieves performance comparable to Pi across GPT-5.6 Sol and Opus 5 while reducing recorded token traffic by 44.7-49.0% and API cost by about one third. In other words, estimated hourly savings are \$8.75-\$13.50 relative to native Codex and Claude Code harnesses, and \$4.36-\$5.71 relative to Pi.
Podzilla Summary coming soon
Sign up to get notified when the full AI-powered summary is ready.
Free forever for up to 3 podcasts. No credit card required.

An Empirical Study of Harness Design for Coding Agents

DeepSeek-V4.1-Flash: Pushing the Limits of KV Cache Compression

JEPA-Anything: Learning Predictive Models across Different Worlds

ScienceIDE: Turning World's Scientific Codebase into Agent Learnable Environments
Free AI-powered recaps of Daily Paper Cast and your other favorite podcasts, delivered to your inbox.
Free forever for up to 3 podcasts. No credit card required.