“Token-Burst-Aware Capacity Planning for LLM Inference Services: Request Arrival, Token Demand, and Failure Risk Modeling from BurstGPT Traces”. Journal of Technology Informatics and Engineering 5, no. 2 (August 3, 2026): 142–164. Accessed September 3, 2026. https://jtie.stekom.ac.id/index.php/jtie/article/view/565.