Autoscaler
Also called: Auto Scaler
A system that automatically adjusts computing resources based on workload demand to maintain performance, availability, and efficient resource utilization.
Explore more about Infrastructure
Related terms
The automatic adjustment of computing resources in response to changing workloads to optimize performance, availability, and cost efficiency.
Load BalancingThe practice of distributing workloads or requests across multiple computing resources to improve scalability, availability, and resource utilization.
Kubernetes ClusterA group of interconnected machines running Kubernetes that work together to deploy, schedule, scale, and manage containerized applications across multiple worker nodes under the control of a centralized control plane.
Compute InstanceA virtual or physical compute resource that provides CPU, memory, storage, and networking for running applications and workloads.
GPUGPUA Graphics Processing Unit (GPU) is a specialized parallel processor designed to perform large-scale mathematical computations efficiently, making it the primary hardware for training and running AI models.