← All IntelIntel / article 
Beyond FLOPs: Energy-Aware Knowledge Distillation for Sustainable LLMs on Code-Related Task
- Source
- arxiv.org
- Author
- Enrique Barba Roque, Lu\'is Cruz, Annibale Panichella
- Date

Why it matters
Model-size decisions get justified with FLOPs; if that number does not track energy on real hardware, the sizing argument needs a different basis.
Terms in this piece · Glossary
- distillation — Training a small, cheap model to imitate a big one's outputs, keeping much of the capability at a fraction of the cost.
Read the source arxiv.org
Recommended reads
articleMemorization Diagnostics for Code LLMs Should be Scale-AwarePrateek Kumar Rajput, Abdoul Aziz Bonkoungou, Alberick Euraste Djir\'e, Xunzhu Tang, Yewei Song, Iyiola Emmanuel Olatunji, El Hacen Diallo, Jacques Klein, Tegawend\'e F. Bissyand\'e
articleAn Empirical Evaluation of Cost-Efficient Large Language Models on Algorithmic Programming TasksChandimal Adikari, Nandika Herath
articleTwo Truths and A Lie? Benchmarking Off-the-Shelf LLMs for Requirements Quality Assessment: Performance, False Alarms, and MissesJannatul Shefa, Alejandro Salado, Paul Wach, Taylan G. Topcu
Comments
Checking sign-in…
Loading comments…