
If you use contamination probes to judge whether a code model memorized a , those probes likely fail on exactly the large models you ship — the methodology, not the model, is what breaks.
articleMeasure, Don't Optimize: Forecasting Recovery in LLM UnlearningZirui Song, Huaxing Liu, Xiang Wang, Shuai Li, Xinye Li, Lang Gao, Jinghui Zhang, Zheng Lu, Fengxian Ji, Xiaojun Chang, Xiuying Chen
articleBeyond FLOPs: Energy-Aware Knowledge Distillation for Sustainable LLMs on Code-Related TaskEnrique Barba Roque, Lu\'is Cruz, Annibale Panichella
articleBeyond the Traceback: Using LLMs for Adaptive Explanations of Programming ErrorsAlexandru-Radu Moraru, Shreyan Biswas, Ujwal GadirajuChecking sign-in…
Loading comments…