paper-0116 · paper · 2019
Emma Strubell, Ananya Ganesh, Andrew McCallum
Put training cost and carbon on the research agenda.
Abstract
Recent progress in hardware and methodology for training neural networks has ushered in a new generation of large networks trained on abundant data. These models have obtained notable gains in accuracy across many NLP tasks. However, these accuracy improvements depend on the availability of exceptionally large computational resources that necessitate similarly substantial energy consumption. As a result these models are costly to train and develop, both financially, due to the cost of hardware and electricity or cloud compute time, and environmentally, due to the carbon footprint required to fuel modern tensor processing hardware. In this paper we bring this issue to the attention of NLP researchers by quantifying the approximate financial and environmental costs of training a variety of recently successful neural network models for NLP. Based on these findings, we propose actionable recommendations to reduce costs and improve equity in NLP research and practice. [OpenAlex]
Academic, score -0.1991
| Metric | Status | Value | Norm. | Weight | Contribution | Source | Confidence | License | Provenance |
|---|---|---|---|---|---|---|---|---|---|
| citation_count | present | 422.0 | 0.001895 | 0.5 | 0.000948 | OpenAlex | high | OpenAlex, CC0 metadata | link |
| library_holdings | missing | recorded as missing, penalized by rule, never imputed | −0.1 | recorded as missing; penalized by rule, never imputed | |||||
| readership_persistence | present | 8.0 | 0.5 | 0.05 | 0.025 | OpenAlex | medium | OpenAlex, CC0 metadata | link |
| syllabus_adoptions | missing | recorded as missing, penalized by rule, never imputed | −0.125 | recorded as missing; penalized by rule, never imputed | |||||
Broad Influence, score 0.0004
| Metric | Status | Value | Norm. | Weight | Contribution | Source | Confidence | License | Provenance |
|---|---|---|---|---|---|---|---|---|---|
| citation_count | present | 422.0 | 0.001895 | 0.2 | 0.000379 | OpenAlex | high | OpenAlex, CC0 metadata | link |
| library_holdings | missing | recorded as missing, penalized by rule, never imputed | −0.125 | recorded as missing; penalized by rule, never imputed | |||||
| readership_persistence | present | 8.0 | 0.5 | 0.4 | 0.2 | OpenAlex | medium | OpenAlex, CC0 metadata | link |
| syllabus_adoptions | missing | recorded as missing, penalized by rule, never imputed | −0.075 | recorded as missing; penalized by rule, never imputed | |||||
Governance Practitioner, score -0.2745
| Metric | Status | Value | Norm. | Weight | Contribution | Source | Confidence | License | Provenance |
|---|---|---|---|---|---|---|---|---|---|
| citation_count | present | 422.0 | 0.001895 | 0.25 | 0.000474 | OpenAlex | high | OpenAlex, CC0 metadata | link |
| library_holdings | missing | recorded as missing, penalized by rule, never imputed | −0.15 | recorded as missing; penalized by rule, never imputed | |||||
| readership_persistence | present | 8.0 | 0.5 | 0.1 | 0.05 | OpenAlex | medium | OpenAlex, CC0 metadata | link |
| syllabus_adoptions | missing | recorded as missing, penalized by rule, never imputed | −0.175 | recorded as missing; penalized by rule, never imputed | |||||
A rank is not a verdict on intrinsic worth. It is a transparent output of declared evidence, weights, and missing-data rules at a specific release date.
Disagree with this rank or a number? Challenge it with your evidence. Every challenge gets a public identifier and a published resolution.