Pith. sign in

REVIEW 2 cited by

How Green are Neural Language Models? Analyzing Energy Consumption in Text Summarization Fine-tuning

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2501.15398 v3 pith:5SUR6ZHR submitted 2025-01-26 cs.CL

classification cs.CL
keywords modelslanguageneuralcarbonconsumptionenergyenvironmentalfine-tuning
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Artificial intelligence systems significantly impact the environment, particularly in natural language processing (NLP) tasks. These tasks often require extensive computational resources to train deep neural networks, including large-scale language models containing billions of parameters. This study analyzes the trade-offs between energy consumption and performance across three neural language models: two pre-trained models (T5-base and BART-base), and one large language model (LLaMA-3-8B). These models were fine-tuned for the text summarization task, focusing on generating research paper highlights that encapsulate the core themes of each paper. The carbon footprint associated with fine-tuning each model was measured, offering a comprehensive assessment of their environmental impact. It is observed that LLaMA-3-8B produces the largest carbon footprint among the three models. A wide range of evaluation metrics, including ROUGE, METEOR, MoverScore, BERTScore, and SciBERTScore, were employed to assess the performance of the models on the given task. This research underscores the importance of incorporating environmental considerations into the design and implementation of neural language models and calls for the advancement of energy-efficient AI methodologies.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. F$^2$Agent: Financial Fusion of Agentic Intelligence for Multimodal Trading

    cs.MA 2026-08 reject novelty 5.0 of 10

    F2Agent, a hierarchy of specialized LLM and Transformer agents with adaptive cross-modal attention and consistency regularization, is reported to beat 16 trading baselines on six assets, though appendix results from a...

  2. Electricity Demand and Grid Impacts of AI Data Centers: Challenges and Prospects

    eess.SY 2025-09 conditional novelty 2.0 of 10

    A review paper synthesizes evidence that AI data center electricity demand is large, bursty, and power-electronics-dominated, creating multi-timescale grid challenges.

Pith tools