← back to paper
arxiv: 2509.08867 · 2 revisions
Benchmarking Energy Efficiency of Large Language Models Using vLLM