
Argues that most of the adaptable surface of an lives in its skills and rather than its weights, and shows a train-free loop that improves that surface using the same frozen model.
articleBetter, Faster, Stronger: Programmatic Skill Learning Best Reduces Agent CostZixi Huang, Xiheng Wang, Andrew Wang, William Jurayj, Bernal Jim\'enez Guti\'errez, Daniel Khashabi, Nicholas Andrews
articleRethinking Self-Evolving Agent Skills: Feedback Dynamics over Multiple RoundsYuxuan Liu, Zhaochen Su, Yuhao Zhang, Jiahe Guo, Zhongwei Xie, Huihao Jing, Lingyun Xie, Qing Zong, Yauwai Yim, Zhixiong Zhang, Haoran Li, Yangqiu Song
articleOne Recipe, Many Harnesses: What Self-Evolution Encodes Across Languages and ModelsSiqi Yang, Qianlan Yang, Yu-Xiong Wang, Saurabh Pujar, Martin HirzelSign in to comment.
Loading comments…