[1]
“Token-Burst-Aware Capacity Planning for LLM Inference Services: Request Arrival, Token Demand, and Failure Risk Modeling from BurstGPT Traces”, JTIE, vol. 5, no. 2, pp. 142–164, Aug. 2026, doi: 10.51903/jtie.v5i2.565.