“Token-Burst-Aware Capacity Planning for LLM Inference Services: Request Arrival, Token Demand, and Failure Risk Modeling from BurstGPT Traces” (2026) Journal of Technology Informatics and Engineering, 5(2), pp. 142–164. doi:10.51903/jtie.v5i2.565.