Brevity is the soul of sustainability: Characterizing LLM response lengths

Soham Poddar; Paramita Koley; Janardan Misra; Niloy Ganguly; Saptarshi Ghosh

Brevity is the soul of sustainability: Characterizing LLM response lengths

Soham Poddar, Paramita Koley, Janardan Misra, Niloy Ganguly, Saptarshi Ghosh

Abstract

A significant portion of the energy consumed by Large Language Models (LLMs) arises from their inference processes; hence developing energy-efficient methods for inference is crucial. While several techniques exist for inference optimization, output compression remains relatively unexplored, with only a few preliminary efforts addressing this aspect. In this work, we first benchmark 12 decoder-only LLMs across 5 datasets, revealing that these models often produce responses that are substantially longer than necessary. We then conduct a comprehensive quality assessment of LLM responses, formally defining six information categories present in LLM responses. We show that LLMs often tend to include redundant or additional information besides the minimal answer. To address this issue of long responses by LLMs, we explore several simple and intuitive prompt-engineering strategies.Empirical evaluation shows that appropriate prompts targeting length reduction and controlling information content can achieve significant energy optimization between 25-60% by reducing the response length while preserving the quality of LLM responses.

Anthology ID:: 2025.findings-acl.1125
Volume:: Findings of the Association for Computational Linguistics: ACL 2025
Month:: July
Year:: 2025
Address:: Vienna, Austria
Editors:: Wanxiang Che, Joyce Nabende, Ekaterina Shutova, Mohammad Taher Pilehvar
Venues:: Findings | WS
SIG:
Publisher:: Association for Computational Linguistics
Note:
Pages:: 21848–21864
Language:
URL:: https://preview.aclanthology.org/acl25-workshop-ingestion/2025.findings-acl.1125/
DOI:
Bibkey:
Cite (ACL):: Soham Poddar, Paramita Koley, Janardan Misra, Niloy Ganguly, and Saptarshi Ghosh. 2025. Brevity is the soul of sustainability: Characterizing LLM response lengths. In Findings of the Association for Computational Linguistics: ACL 2025, pages 21848–21864, Vienna, Austria. Association for Computational Linguistics.
Cite (Informal):: Brevity is the soul of sustainability: Characterizing LLM response lengths (Poddar et al., Findings 2025)
Copy Citation:
PDF:: https://preview.aclanthology.org/acl25-workshop-ingestion/2025.findings-acl.1125.pdf

PDF Cite Search Fix data