×

Predictive Scaling In Kubernetes Using Machine Learning

Author : Sneha Iyer Journa Name: International Journal of Science, Engineering and Technology Volume: 7 issue: 1 Year: Volume-7-issue-1 Views : 167
Abstract:
Predictive scaling represents a transformative shift in Kubernetes resource management, moving away from reactive thresholds toward proactive, data-driven orchestration. Traditional mechanisms, such as the Horizontal Pod Autoscaler (HPA), rely on observed metrics like CPU and memory utilization, which often results in a \"lag\" where resources are provisioned only after performance degradation has begun. By integrating machine learning (ML) models—including Time Series Analysis, Recurrent Neural Networks (RNNs), and Long Short-Term Memory (LSTM) networks—Kubernetes clusters can now anticipate traffic surges and workload spikes before they occur. This review explores the architectural integration of ML providers with the Kubernetes Metrics API, the efficacy of various algorithmic approaches in reducing latency, and the cost-optimization benefits of predictive modeling. As cloud-native environments grow in complexity, predictive scaling emerges as a critical component for maintaining high availability while minimizing resource wastage in dynamic, large-scale microservices architectures.

Related Indexing Platform

Indexed

Zenodo Logo
Zenodo
Research Data Repository
https://zenodo.org/records/19481685
DOI
DOI Resolver
Global Persistent Identifier
https://doi.org/10.5281/zenodo.19481685
GS
Google Scholar
Search this title on Scholar
Search on Google Scholar
SS
Semantic Scholar
Search this title
Search on Semantic Scholar
Lens
Lens.org
Check citations via DOI
Search on Lens.org
Leave Your Comment

Related Reviewers

Chat with Expert