cs.LGOct 1, 2026

Exposing the Cost of Deep Learning Audio Development

Authors: Constance Douwes, Paul Magron, Romain Serizel

Organizations: Universit´e de Lorraine, CNRS, Inria, LORIA, F-54000 Nancy, France

Abstract

The environmental impact of deep learning has attracted increasing attention over the past decade. Existing studies mainly focus on the energy and carbon emissions of model training and inference, while the whole development phase is often overlooked. Yet, architecture prototyping and intensive experiments are conducted during this stage, which is highly energy-demanding. In this article, we propose a methodology to estimate these costs, based on activity logs from the Grid5000 shared computing platform used by the LORIA laboratory. As a case-study, we focus on audio projects developed in the Multispeech research team. We evaluate the overall energy cost of four projects, and we compare them to those of training the reported models. Our results show that the energy required for the development phase is 3 to 256 times greater than that required to train the best-performing model alone. These results advocate for a more systematic reporting of energy consumption across the entire life cycle of deep learning-based audio projects.

Figures & tables

Explore similar work

CardsList
  1. Assessing the Energy and Carbon Emissions of Neural Speaker Verification Model in Training and Inference

    Jun 6, 2026Hugo Leguillier, Driss Matrouf, Guillaume Lechien +1Automatic Speaker VerificationSpeaker

  2. Towards Resource-Efficient LLMs: End-to-End Energy Accounting of Distillation Pipelines

    May 13, 2026Katherine Lambert, Sasha LuccioniEnergy ConsumptionResidual Distillation

  3. LongAudioSpan: Spanning the Duration and Depth of Audio Comprehension

    Aug 26, 2026Wen Huang, Yunfei Chu, Meng Gao +2Audio UnderstandingLarge Audio Language Models