Uppsats
EC2 Autoscaling Under Bursty Traffic: An Evaluation of Tail Latency and Scaling Responsiveness : EC2 Autoscaling Under Bursty Traffic: A Tail Latency focused Comparison with AWS Lambda
Yrkesexamen på avancerad nivå
Uppsala universitet/Datorteknik
Publicerad: 2026
Språk: Engelska
Nyckelord
klicka för att sökaSammanfattning
Cloud-based services must remain responsive under bursty and unpredictable traffic patterns, yet the relationship between autoscaling configuration choices and tail latency behaviour under such conditions remains insufficiently understood. This thesis evaluates how different Amazon EC2 autoscaling configurations influence system performance under bursty high-load traffic, with a focus on tail latency at the 95th and 99th percentiles (p95 and p99), time-to-stability (TTS), and service level objective (SLO) compliance. The experiments were conducted on real AWS infrastructure using a controlled experimental design. Four configuration parameters were varied relative to a fixed baseline: scaling policy type, scaling policy aggressiveness, minimum instance capacity, and instance size. Three traffic patterns representing steep bursts, gradual ramps, and repeated bursts were evaluated across all configurations. AWS Lambda was included as a serverless reference architecture under identical workload conditions. The central finding is that reactive autoscaling on EC2 is primarily constrained by instance provisioning delay rather than detection latency. This delay, averaging 5 to 6.5 minutes across configurations, accounts for 60 to 65 percent of total scaling delay and cannot be eliminated through policy tuning alone. Minimum capacity proved to be the most effective configuration parameter in the evaluated setup: pre-provisioning five instances eliminated SLO violations entirely across all evaluated traffic patterns. AWS Lambda achieved the same result, with cold starts affecting only around 0.013% of requests and remaining within SLO headroom. The results provide empirical guidance for practitioners designing latency-sensitive systems under bursty workloads, highlighting that burst resilience depends more on pre-provisioned capacity than on scaling policy aggressiveness, and that the choice between EC2 and Lambda reflects a trade-off between configuration control and operational simplicity.
Information
- Författare
- Ehrengren, Arvid
- Lärosäte / institution
- Uppsala universitet/Datorteknik
- Publiceringsdatum
- 2026
- Uppsatstyp
- Yrkesexamen på avancerad nivå
- Språk
- Engelska
Utforska vidare
Liknande uppsatser
Uppsatser med liknande ämnen och nyckelord.
Master-uppsats, Linköpings universitet/Institutionen för datavetenskap
Svensson, Petter, Norberg, Philip
Publicerad: 2025
Yrkesexamen på avancerad nivå, Blekinge Tekniska Högskola/Institutionen för industriell ekonomi
Shirzad, Shams, Musliu, Clirim
Publicerad: 2024
Kandidat-uppsats, Linnéuniversitetet/Institutionen för datavetenskap och medieteknik (DM)
Ksouri, Zied, Sai, Adam
Publicerad: 2026
Master-uppsats, Uppsala universitet/Institutionen för informatik och media
Krishna Murthy, Priyanka
Publicerad: 2026
Master-uppsats, Luleå tekniska universitet/Institutionen för system- och rymdteknik
Sheta, Ahmed
Publicerad: 2026
Master-uppsats, Malmö universitet/Fakulteten för teknik och samhälle (TS)
Gerlach, Jakob
Publicerad: 2026