Google Cloud | Compute | Circuit

Let Google Cloud’s Predictive Services Autoscale Your Infrastructure

At Google Cloud, we believe you get most benefits from the cloud when you scale infrastructure based on changing demand. Compute Engine allows you to configure autoscaling to save costs during periods of low demand, and add capacity to support peak loads. 

When you use a managed instance group (MIG), you can have an autoscaler automatically create or delete virtual machine (VM) instances based on increases or decreases in load. However, if your application takes several minutes to initialize, creating VMs in response to growing load might not increase your application’s capacity quickly enough. For example, if there’s a large increase in load (like when users first wake up in the morning), some users might experience delays while your application is initializing on new instances.

A good way to solve this problem would be to create VMs ahead of demand so that your application has enough time to initialize beforehand. This requires knowing upcoming demand. If only we could predict the future… Well, now we can!

Introducing predictive autoscaling

Predictive autoscaling uses Google Cloud’s machine learning capabilities to forecast capacity needs. It creates VMs ahead of growing demand allowing enough time for your application to initialize.

Figure 1. Autoscaling creates VMs as demand grows leaving no buffer for application to initialize. Predictive autoscaling creates VMs ahead of demand allowing enough time for your application to initialize and start serving new load.

How does it work?

Predictive autoscaling uses your instance group’s CPU history to forecast future load and calculate how many VMs are needed to meet your target CPU utilization. Our machine learning adjusts the forecast based on recurring load patterns for each MIG. 

You can specify how far in advance you want autoscaler to create new VMs by configuring the application initialization period. For example, if your app takes 5 minutes to initialize, autoscaler will create new instances 5 minutes ahead of the anticipated load increase. This allows you to keep your CPU utilization within the target and keep your application responsive even when there’s high growth in demand. 

Many of our customers have different capacity needs during different times of the day or different days of the week. Our forecasting model understands weekly and daily patterns to cover for these differences. For example, if your app usually needs less capacity on the weekend our forecast will capture that. Or, if you have higher capacity needs during working hours, we also have you covered.

Why should you try it?

Predictive autoscaling continuously adapts forecasted capacity to best match upcoming demand. Autoscaler checks the forecast several times per minute and creates or deletes VMs to match its prediction. The forecast itself is updated every few minutes to match recent load trends so if your growth rate is higher or lower than usual we will adjust the forecast accordingly. This gives you capacity needed to cover peak load while saving on cost when demand goes down. 

You can start using predictive autoscaling without worry as it’s fully compatible with the current autoscaler. Autoscaler will calculate enough VMs to cover both forecasted as well as real-time CPU load—whichever is higher. This works with other autoscaling features as well: you can scale based on schedule, your Load Balancer request target or Cloud Monitoring metrics. Autoscaler provides enough capacity to all of your configurations by taking the highest number of VMs needed to meet all your targets.

Getting started

You can enable predictive autoscaling in the Google Cloud Console. Select an autoscaled MIG from the instance groups page and click Edit group. Change predictive autoscaling configuration from Off to Optimize for availability.

To better understand whether predictive autoscaling is good for your application, click the link See if predictive autoscaling can optimize your availability. This will show you a comparison of the last seven days with your current autoscaling configuration vs. with predictive autoscaling enabled.

In the above chart, 

  • Average VM minutes overloaded per day shows how often your VMs exceed your CPU utilization target. This happens when demand is higher than available capacity. Predictive autoscaling can reduce this by starting VMs ahead of anticipated load. 
  • Average VMs per day is a proxy for cost. This shows how much additional VM capacity you need to keep your CPU utilization within the target you have set. You can optimize your cost by adjusting Minimum instances andCPU utilization as explained below. 

Optimizing your configuration

Make sure your Cool down period reflects how long it takes for your application to initialize from VM boot time until it’s ready to serve the load. Predictive autoscaling will use this value to start VMs ahead of forecasted load. If you set it to 10 minutes (600 seconds) your VMs will start 10 minutes before the load is expected to increase.

Review your autoscaling CPU utilization target and Minimum number of instances. With predictive autoscaling you no longer need a buffer to compensate for the time it takes for a VM to start. If your application works best at 70% CPU utilization you don’t need to set target to a much lower value as predictive autoscaling will start VMs ahead of usual load. A higher CPU utilization and lower Minimum number of instances allows you to reduce the cost as you don’t need to pay for additional capacity to prepare for growing demand.

Try predictive autoscaling today

Predictive autoscaling is generally available across all Google Cloud regions. For more information on how to configure, simulate and monitor predictive autoscaling, consult the documentation.

By: Pawel Wenda (Product Manager)
Source: Google Cloud Blog



For enquiries, product placements, sponsorships, and collaborations, connect with us at hello@globalcloudplatforms.com. We'd love to hear from you!


Our humans need coffee too! Your support is highly appreciated, thank you!

Total
0
Shares
Previous Article
Google Cloud | Security

New Research: Enterprises More Confident Than Ever In Cloud Security

Next Article

Rubin Observatory Offers First Astronomy Research Platform In The Cloud

Related Posts

IBM Unveils New Capabilities for Preserving Aging Infrastructure Using AI, 3D Modeling and Data Capture

IBM Maximo for Civil Infrastructure is designed to assist organizations in better managing, monitoring and maintaining their infrastructure assets. ARMONK, NY Oct. 14, 2020 -- IBM (NYSE: IBM) today announced new capabilities in IBM Maximo for Civil Infrastructure to help prolong the lifespan of aging bridges, tunnels, highways, and railways. New enhancements include the ability to deploy on Red Hat OpenShift for hybrid cloud environments, as well as new AI and 3D model annotation tools that can provide deep industry and task-specific insights to support engineers. By helping facilitate off-site inspections, advanced analytics and predictive maintenance, IBM Maximo for Civil Infrastructure is designed to assist organizations in better managing, monitoring and maintaining their infrastructure assets. In the United States roughly $2 trillion in infrastructure repairs were unfunded in 2015, according to the 2017 American Society of Civil Engineers Infrastructure Report Card. And around the world, the prevalence of aging infrastructure threatens the continuity of day-to-day life for citizens worldwide. Owners, operators and engineers need to be able to improve their ability to decide where, when and how to address infrastructure issues with critical assets that must endure for generations. [embedded content] IBM Maximo for Civil Infrastructure can consolidate numerous sources of data including maintenance and design details; near real-time IoT data generated from sensors; wearables from workers; stationary cameras and drones and weather data from The Weather Company, to identify and measure the impact of damage such as cracks, rust and corrosion, as well as displacement vibrations and stress. These insights can help organizations more proactively manage and prioritize infrastructure repair and reduce the need for time-intensive manual inspections and unnecessary costs. “Tools like AI, predictive maintenance, drones and hybrid cloud will play an important role in meeting the challenge of rising infrastructure costs, and helping these vital structures endure for future generations,” said Bjarne Jørgensen, Executive Director, Asset Management at Sund and Baelt. “These solutions can help determine the exact need for maintenance in near real-time to assist organizations in extending the lifetime of structures.” IBM Maximo for Civil Infrastructure incorporates AI visual recognition tools developed by IBM Research to allow civil engineers to make structures come alive using 3D modeling. Capabilities like Maximo Visual Inspection can make it easier to identify defects, their root-cause, and place them in the context of the greater structure to perform rapid assessments to better prioritize maintenance decisions that target critical repairs. These tools can be increasingly important for future engineers as skills availability may be a challenge. "Infrastructure maintenance is a problem that’s being compounded from all sides: Bridges are getting older, payloads are getting larger, and the necessary preventive actions and maintenance are often postponed due to lack of funding,” Jørgensen added. "With Maximo for Civil Infrastructure, IBM is introducing a solution that addresses the problem from all sides, using IoT and AI technology to administer more proactive repairs, maintain invaluable institutional and engineering knowledge, and better prioritize resources.” “Maximo for Civil Infrastructure was developed with input from some of the largest operators of infrastructure in the world so that IBM’s powerful technology across AI and IoT is carefully adapted to the unique needs of civil engineers,” said Joe Berti, VP of AI Applications at IBM. “With these tools we believe civil engineers will be able to innovate and improve the process of monitoring, maintaining and preserving important structures around the world.” IBM Maximo for Civil Infrastructure provides the following new capabilities, in addition to the core offerings available as part of the IBM Maximo Application Suite. Maximo Application Suite licensing and Open Shift deployment: With a single license, customers can now deploy asset management, sensor integration, advanced data analytics including AI functionality and visual analytic capability. Capabilities including Monitor and Health can be deployed now on RedHat Open Shift, which allows customers to more quickly deploy, manage and scale their hybrid cloud deployments with ease. Defect Management: A new user interface that records defect information, adds multi-variable defect rankings, attaches pictures, and stores history for defects. Structural defects do not exist in isolation, they can affect everything they touch and the integrity of the overall structure. By comparing detected defects against work history, sensor data, weather and traffic data and more, AI can help engineers better identify root causes and patterns that indicate a defect may reoccur.  Improved 3D Visualization: Most serious defects are located within the structure and are not necessarily visible from the outside. However, new tools within the Maximo BIM viewer plugin allow users to add annotations to their standard 3D models, giving users access to a 3D representation of an asset, for example a pillar or beam, where all the defects have been annotated on that asset.  Asset Loader Improvements: While every piece of civil infrastructure is unique, many share common hierarchies of assets, and some organizations have hundreds or even thousands of similar bridges that need to be defined in the asset management system. To streamline the process, a new tool to better manage import and export of an asset hierarchy, including a new UI to manage imports and exports and the process of selecting files, is now available.  IBM Maximo for Civil Infrastructure integrates 30 years of recognized industry-leading infrastructure asset management with best-in-class models from the world's premier infrastructure firms. It helps operators and engineers address one of the world's largest and most complex challenges — extending the lifespan of critical structures under frequently changing conditions. You can read more about how some of its capabilities were developed here. About IBM MaximoPowered by IBM's investments in artificial intelligence, fueled by IoT data, and built for hybrid cloud, The IBM Maximo Application Suite is extending its leadership as one of the most trusted enterprise asset management systems on the planet. And with new investments in remote monitoring, computer vision and AI-powered anomaly detection, it is poised to remain a leading solution for tomorrow's asset management challenges, empowering Operational Technology (OT) and Information Technology (IT) leaders with a comprehensive view into asset performance. For more information please visit: www.ibm.com/products/maximo. Media Contact:Holli HaswellIBM Director, External Relationshhaswell@us.ibm.com