Auto Scaling

197 단어·1 분·원문(.md)

Auto Scaling #

Monitors applications and automatically adjusts capacity to maintain stable, predictable performance at the lowest possible cost.

With AWS Auto Scaling, you can easily set up application scaling across multiple services and resources in minutes.

Uses of Auto Scaling #

Use a minimum number of instances

Maintain the desired number of instances as a target

Keep instances below the maximum instance count

Distribute instances evenly across Availability Zones

Ensure instances are always available to maintain service

EC2 Auto Scaling Configuration #

  • Launch Configuration: What and how to launch?
    • EC2 type, size
    • AMI
    • Security Group, Key, IAM
    • User Data
  • Monitoring: When to launch? + Status check
    • Example: Launch additional instances when CPU utilization exceeds a certain percentage, or when one EC2 instance dies in a stack requiring two or more.
    • Integrate with CloudWatch (AND/OR) ELB
  • Desired Capacity: How many to launch?
    • Example: Minimum 1 ~ Maximum 3
  • Lifecycle Hook: Callback on instance launch/termination
    • Can perform pre/post-processing in conjunction with other services -> CloudWatch Event / SNS / SQS
    • Transition to Terminating:wait/Terminating:Proceed state
    • Waits for 3600 seconds by default (allowing tasks like image backup or log backup during this time)

EC2 Auto Scaling Flowchart #

DevOps/aws/as.md