Repository logo

MARLIN: multi-agent game-theoretic reinforcement learning for sustainable LLM inference in cloud datacenters

dc.contributor.authorMoore, Hayden, author
dc.contributor.authorQi, Sirui, author
dc.contributor.authorMilojicic, Dejan, author
dc.contributor.authorBash, Cullen, author
dc.contributor.authorPasricha, Sudeep, author
dc.contributor.authorACM, publisher
dc.date.accessioned2026-09-17T18:27:01Z
dc.date.issued2026-06-22
dc.description.abstractLarge Language Models (LLMs) have become increasingly prevalent in cloud-based platforms, propelled by the introduction of AI-based consumer and enterprise services. LLM inference requests in particular account for up to 90% of total LLM lifecycle energy use, dwarfing training energy costs. The rising volume of LLM inference requests is increasing environmental footprints, particularly carbon emissions and water consumption. To improve sustainability for LLM inference serving in cloud datacenter environments, we propose a novel multi-agent game-theoretic reinforcement learning framework called MARLIN to co-optimize time-to-first token (TTFT), carbon emissions, water usage, and energy costs associated with LLM inference. MARLIN demonstrates a reduction of at least 18% in TTFT, 33% in carbon emissions, 43% in water usage, and 11% in energy costs compared to state-of-the-art LLM inference management frameworks.
dc.format.mediumborn digital
dc.format.mediumarticles
dc.identifierFACF_ACMOA_3797248.3815404.pdf
dc.identifier.bibliographicCitationHayden Moore, Sirui Qi, Dejan Milojicic, Cullen Bash, and Sudeep Pasricha. 2026. MARLIN: Multi-Agent Game-Theoretic Reinforcement Learning for Sustainable LLM Inference in Cloud Datacenters. In International Green and Sustainable Computing Conference (IGSC 2026), June 22-24, 2026, Canandaigua, NY, USA. ACM, New York, NY, USA, 8 pages. https://doi.org/10.1145/3797248.3815404
dc.identifier.doihttps://doi.org/10.1145/3797248.3815404
dc.identifier.urihttps://hdl.handle.net/10217/245565
dc.languageEnglish
dc.language.isoeng
dc.publisherColorado State University. Libraries
dc.relation.ispartofPublications
dc.relation.ispartofACM DL Digital Library
dc.rights.licenseThis work is licensed under a Creative Commons Attribution 4.0 International License.
dc.rights.urihttps://creativecommons.org/licenses/by/4.0
dc.subjectlarge language models
dc.subjectsustainability
dc.subjectcarbon emissions
dc.subjectwater usage
dc.subjectenergy costs
dc.subjectcloud datacenters
dc.subjectreinforcement learning
dc.titleMARLIN: multi-agent game-theoretic reinforcement learning for sustainable LLM inference in cloud datacenters
dc.typeText
dc.typeImage

Files

Original bundle

Now showing 1 - 1 of 1
Loading...
Thumbnail Image
Name:
FACF_ACMOA_3797248.3815404.pdf
Size:
1.15 MB
Format:
Adobe Portable Document Format

Collections