arXiv (OAI Expanded)open access
Characterization of Request and Token Energy Costs for LLM Inference Workloads on GPU Platforms
performance, distributed, parallel, and cluster computing, machine learning, c.4
This document is indexed with metadata only — full text is not available in the archive for this record.
Open the official source ↗
Related documents
Record · ID 783222
Retrieved via
Conceptio — every document is proof-bundled with source, license, and retrieval metadata.