# Preprint finds energy-aware distillation can cut code-model inference energy

_Wednesday, August 19, 2026 at 12:00 AM EDT · AI, Science · Latest · Tier 2 — Notable_

A preprint on software-engineering language models reports that energy-surrogate-guided knowledge distillation reduced inference energy use by as much as 90% and memory use by 86% in its experiments, with modest accuracy trade-offs. The researchers tested clone detection, vulnerability prediction and code summarization, and found FLOPs did not consistently indicate actual energy consumption. The work uses direct CPU and GPU energy estimates during distillation rather than optimizing only for operation counts.

## Sources

- [cs.AI updates on arXiv.org](https://arxiv.org/abs/2608.17515)

---
Canonical: https://techandbusiness.org/newswire/5omkjabOm8v5CxndSEiLWZ
Retrieved: 2026-08-19T11:35:56.691Z
Publisher: Tech & Business (techandbusiness.org)
