Overview

Automatically keeps the number of servers matched to actual demand.

HiTechCloud Auto Scaling provides effective policy-based management of server resources. A business can schedule a policy to run on a regular basis, or create a real-time monitoring policy, to manage the number of Elastic Compute instances in an AS group.

When demand rises, Auto Scaling adds elastic compute instances automatically to maintain server performance. When demand falls, it removes them according to the conditions you have configured, saving cost and reducing idle resources.

Overview of the HiTechCloud Auto Scaling model

Balance performance and cost

Add EC automatically when demand rises and release resources that are no longer needed when load falls, keeping the system stable without wasting budget.

Scaling policies by schedule or by monitoring

Set fixed schedules, recurring cycles or real-time monitoring thresholds to stay ahead of shifts in application load.

Automate server lifecycle

AS supports instance creation, group membership, load distribution, alerting and EC reclamation under standardized operating rules.

Auto Scale resources up and down according to configured policies
24/7 Continuously tracks load and scaling events
N+1 Maintain processing capacity during traffic spikes
Cost Reduces surplus servers during off-peak periods
Features

Scale infrastructure proactively, with no manual steps.

From automatic EC creation and load balancing to event notifications, HiTechCloud Auto Scaling helps operations teams respond quickly to changes in load.

Auto scaling

Automatically creates or removes EC instances based on actual workload, reducing reliance on manual deployment.

Automatically adjust instances

Apply scheduled, recurring, or monitoring-based policies to scale at the right moment — for example, adding capacity ahead of peak hours.

Load balancing

Distributes traffic across the EC instances in an auto scaling group, improving application availability, scalability and performance.

Automated notifications

Sends an alert when an EC instance is launched, added to a group, removed, or when a scaling policy is triggered.

Set up schedules and recurring tasks

Adjust resources automatically by hour, by day or on a defined business cycle, freeing IT teams for higher-value work.

AS policy

Scale flexibly by schedule, cycle or monitoring data.

Businesses can standardize multiple policy types to handle peak hours, scheduled tasks, or real-time load fluctuations.

01

Scheduled Scaling

Scale resources up or down on a fixed schedule for time windows with predictable demand.

02

Periodic Scaling

Repeat scaling actions automatically on a daily, weekly, or monthly cycle to match how you operate.

03

Metric-based Scaling

Trigger scaling based on CPU, RAM, connection or request metrics, or on operational alert thresholds.

Use case

Serves a range of business-critical applications and compute tasks.

Auto Scaling suits systems with variable load that need to add processing capacity quickly and release resources as soon as they are no longer needed.

Web applications with fluctuating traffic

When a marketing campaign, flash sale or online event drives load up quickly, AS automatically adds EC instances to maintain the user experience.

Scheduled batch jobs

Create an EC group at a scheduled time to run compute tasks, then automatically scale it down or release it when the job finishes.

Flexible Dev/Test environment

Add resources during working hours for development teams and reduce EC instances outside them to control infrastructure cost.

Services that require high availability

Combine Auto Scaling with load balancing to maintain a minimum number of servers, reducing the risk of disruption during incidents or unusual load.

Benefits

Optimize cost, improve availability and simplify infrastructure operations.

Cost savings

Keeps the number of EC instances matched to actual demand, adding capacity when needed and removing surplus resources as load falls.

Convenient performance monitoring

Observe scaling state, performance and compute resources to track system health over time.

Reliable operations

Standardized scaling policies make applications respond more consistently to load changes and reduce the risk of manual error.

Complete cloud ecosystem

Works easily with Elastic Compute, Load Balancer, monitoring, networking and other HiTechCloud services.

Pricing

Choose the right Auto Scaling size and billing cycle.

Prices exclude VAT. Filter by configuration size and select a billing period to see the corresponding cost.

3536
Small

C6 Small 1

A compact configuration for a starter AS group, suited to web applications or background services with steady load.

1 vCPU 1 GB RAM
225,000 VND / 1 month
  • 500 Mbps internal VPC bandwidth
  • Scale on a schedule or on monitoring metrics
  • Suited to light workloads
Order now
3537
Medium

C6 Medium 2

More memory for applications that need extra RAM, while keeping running costs optimized.

1 vCPU 2 GB RAM
300,000 VND / 1 month
  • 500 Mbps internal VPC bandwidth
  • Automatically add or remove EC by policy
  • Suitable for small web apps and APIs
Order now
3538
Large

C6 Large 1

Balances CPU and cost for AS groups that need to handle more requests at peak hours.

2 vCPU 2 GB RAM
450,000 VND / 1 month
  • 800 Mbps internal VPC bandwidth
  • Load balancer integration
  • Suited to services with fluctuating traffic
Order now
3540
Memory

C6 Large 6

Optimized for memory-hungry workloads such as workers, cache layers or scheduled background tasks.

2 vCPU 12 GB RAM
1,200,000 VND / 1 month
  • 800 Mbps internal VPC bandwidth
  • Suited for batch/worker jobs needing high RAM
  • Monitor scaling status 24/7
Order now
3541
XLarge

C6 XLarge 1

A high-CPU configuration for applications handling concurrency, heavy API traffic or burst compute tasks.

4 vCPU 4 GB RAM
900,000 VND / 1 month
  • Internal VPC bandwidth of 1,000 Mbps
  • Scale quickly when load spikes
  • Suitable for CPU-intensive workloads
Order now
Deploy

A clear, controllable process for configuring auto scaling.

HiTechCloud supports you from application load assessment and Auto Scaling group design through monitoring, threshold tuning, and post-deployment cost optimization.

01

Assess load and SLA targets

Identify peak traffic, performance thresholds, EC configuration, and the minimum and maximum server counts.

02

Design Auto Scaling groups

Standardize the launch template, network, security groups and health checks, and integrate load balancing where needed.

03

Configure scaling policies

Set policies by schedule, cycle or monitoring metric to scale resources up and down automatically.

04

Continuous monitoring and optimization

Review alerts, scaling logs and application performance, and tune thresholds to balance SLA against cost.

FAQs

Frequently asked questions about Auto Scaling.

Key information to help businesses configure Auto Scaling safely and effectively.

Why use Auto Scaling?

Auto Scaling adds resources automatically as load rises and releases them as load falls, keeping performance stable, reducing interruptions and optimizing operating costs.

How do I configure Auto Scaling?

You need to define the EC group, the template server configuration, minimum/maximum counts, the scaling policy, monitoring thresholds, notifications and the load balancing mechanism if your application requires one.

How does Auto Scaling work?

The service watches the configured conditions. When a scale-out condition is met, the system provisions additional ECs; when demand falls, surplus ECs are removed under the scale-in policy.

What types of Auto Scaling are available?

Common patterns include scheduled scaling, recurring scaling and scaling driven by monitoring metrics such as CPU, RAM, requests, connections or performance alerts.

How do I monitor and control Auto Scaling?

You can track the number of EC instances, scaling events, health checks, operational logs and performance alerts, and adjust policy thresholds as the load profile changes.

Which applications can Auto Scaling work with?

The service suits web and app servers, APIs, batch workers, background processing systems, Dev/Test environments and any application that can run across multiple independent EC instances behind a load balancer.

HiTechCloud Auto Scaling

Ready to automate resource scaling for your application?

Contact HiTechCloud for advice on Auto Scaling configuration, load balancing, monitoring and cost optimization policies suited to your workload.

Deployment consultation