
Shows a local model reading docs well enough to drive fuzzing at scale, and quantifies how much of an API surface has cross-parameter constraints that naive generators break on.
articleLost in Compaction: Evaluating Side-Constraint Loss under Context CompactionZhiqi Wang, Yichi Zhang, Dongwon Lee, Yuchen Yang
articleThe Devil Is in the Interface: Evaluating How Tool Architecture Shapes Coding Agent BehaviorXiangzhe Xu, Hamidreza Saghir, Qianhui Wu, Marc-Alexandre C\^ot\'e, Tong Wang, Kiran Lakkaraju, Kexin Pei, Xiangyu Zhang
articleHow well LLM-based test generation techniques perform with newer LLM versions?Michael Konstantinou, Renzo Degiovanni, Mike PapadakisSign in to comment.
Loading comments…