Pith. sign in

REVIEW 7 cited by

LLMCarbon: Modeling the end-to-end Carbon Footprint of Large Language Models

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2309.14393 v2 pith:OWKQAFO4 submitted 2023-09-25 cs.CL cs.AIcs.CYcs.LG

classification cs.CLcs.AIcs.CYcs.LG
keywords carbonfootprintllmstrainingmlco2cannotcarbdense
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

The carbon footprint associated with large language models (LLMs) is a significant concern, encompassing emissions from their training, inference, experimentation, and storage processes, including operational and embodied carbon emissions. An essential aspect is accurately estimating the carbon impact of emerging LLMs even before their training, which heavily relies on GPU usage. Existing studies have reported the carbon footprint of LLM training, but only one tool, mlco2, can predict the carbon footprint of new neural networks prior to physical training. However, mlco2 has several serious limitations. It cannot extend its estimation to dense or mixture-of-experts (MoE) LLMs, disregards critical architectural parameters, focuses solely on GPUs, and cannot model embodied carbon footprints. Addressing these gaps, we introduce \textit{\carb}, an end-to-end carbon footprint projection model designed for both dense and MoE LLMs. Compared to mlco2, \carb~significantly enhances the accuracy of carbon footprint estimations for various LLMs. The source code is released at \url{https://github.com/SotaroKaneda/MLCarbon}.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 7 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Throttling Web Agents Using Reasoning Gates

    cs.AI 2025-09 conditional novelty 6.0 of 10

    Rebus-based reasoning gates, puzzles built from random word/domain clue sets, impose token costs on LM web agents that are up to 9.2x the generator's cost.

  2. CEO-DC: Driving Decarbonization in HPC Data Centers with Actionable Insights

    cs.AR 2025-07 conditional novelty 6.0 of 10

    A decision framework using new carbon and price efficiency metrics shows most AI platform improvements cannot keep pace with demand growth, and short upgrade cycles need carbon prices far above current levels.

  3. Analysis of Propaganda in Tweets From Politically Biased Sources

    cs.SI 2025-07 conditional novelty 6.0 of 10

    Journalists at politically extreme news outlets tweet propaganda-like language more often than those at mild outlets, and large language models outperform a fine-tuned BERT in detecting it.

  4. Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time

    cs.CL 2025-05 conditional novelty 6.0 of 10

    SITAlign is an inference-time constrained decoder that maximizes a primary reward while enforcing thresholds on secondary rewards, and it reports better primary-reward win-tie rates than weighted-objective decoding.

  5. Calculating Software's Energy Use and Carbon Emissions: A Survey of the State of Art, Challenges, and the Way Ahead

    cs.SE 2025-06 conditional novelty 5.0 of 10

    A structured survey of 21 software energy and carbon calculation tools, organized as Monitoring, Estimation, or Black-Box approaches, with a component-wise comparison and a list of open challenges.

  6. Scaling Fine-Grained MoE Beyond 50B Parameters: Empirical Evaluation and Practical Insights

    cs.LG 2025-06 conditional novelty 5.0 of 10

    At 56B total parameters, fine-grained MoE with smaller, more numerous experts beats standard Switch and Mixtral-style MoE on validation loss and average downstream accuracy at matched FLOPs.

  7. Performance is not All You Need: Sustainability Considerations for Algorithms

    cs.CV 2025-08 reject novelty 4.0 of 10

    The paper introduces FMS and ASC, composite sustainability scores that fuse accuracy and energy consumption, and evaluates them on multiple vision tasks.

Pith tools