Skip to main content

Auto Scaling Group

This post is part of a bigger topic Autoscaling Publishers in AWS.

Now that we have a Launch Configuration, based on it, we can create an Auto Scaling Group. This will be in charge of creating/terminating publisher instances based on some rules (Scaling Policies).

We base our Auto Scaling Group on the Launch Configuration created earlier (sdl_publisher_lc). The very important properties are the Desired, Min and Max number of instances in the group. The Desired is set usually by the Scaling Policies, but more about those later. The Min and Max are the limits of this group. In my case I use Min=1 and Max=3, meaning that I want to have 1 publisher running at all times and when needed, based on load, 2 additional publishers can be added to the group by an 'Increasing size' Scaling Policy.

Once load passes, a 'Decreasing size' Scaling Policy reduces the number of instances in the group.


Scaling Policies

These policies represent the rules for adding/removing instances to the group. They can monitor metrics on the instance itself (e.g. CPU), CloudWatch metrics, or even CloudWatch alarms (e.g. Publish Alarm defined earlier) in order to increase and decrease the number of instances.

We define an 'increase_group_size' as a Scaling Policy with Steps in order to add more publisher instances as the size of the Publish Queue increases.

We also define a 'decrease_group_size' as Simple Scaling Policy that reduces the size of the group. But more details about these policies in a followup post.




Lifecycle Hooks

We are going to use a lifecycle hook when scaling-in (decreasing) the size of out group, when publishing load has passed.

More details about the termination hook in a later post. For now, we create one hook that is going to be raised when the group attempts to terminate an instance. This hook can be intercepted in a CloudWatch event that can then trigger a Lambda Function that will instruct the publisher to shutdown gracefully. Once that happens, the termination hook is released and termination occurs normally.

The Lifecycle Hook Name is important, because in our Lambda Function we will instruct this particular hook to continue termination.

Heartbeat Timeout specifies the time needed to this hook to expire. This means that in the case the termination Lambda did not release the hook in the meantime, the hook will be automatically released once this timeout expires.



Comments

Popular posts from this blog

Running sp_updatestats on AWS RDS database

Part of the maintenance tasks that I perform on a MSSQL Content Manager database is to run stored procedure sp_updatestats . exec sp_updatestats However, that is not supported on an AWS RDS instance. The error message below indicates that only the sa  account can perform this: Msg 15247 , Level 16 , State 1 , Procedure sp_updatestats, Line 15 [Batch Start Line 0 ] User does not have permission to perform this action. Instead there are several posts that suggest using UPDATE STATISTICS instead: https://dba.stackexchange.com/questions/145982/sp-updatestats-vs-update-statistics I stumbled upon the following post from 2008 (!!!), https://social.msdn.microsoft.com/Forums/sqlserver/en-US/186e3db0-fe37-4c31-b017-8e7c24d19697/spupdatestats-fails-to-run-with-permission-error-under-dbopriveleged-user , which describes a way to wrap the call to sp_updatestats and execute it under a different user: create procedure dbo.sp_updstats with execute as 'dbo' as...

Toolkit - Performance

This post if part of a series about the File System Toolkit  - a custom content delivery API for SDL Tridion. This post presents performance data that was captured for each major functionality with and without caching Linking, CP Assembler, Component Presentation Factory, Dynamic Content Queries, Model Factory. The data was captured on a 2014 Macbook Pro 15", 16 GB RAM, 2.6 GHz Intel Core i7 running OS X El Capitan. Test methodology: each test was run for 3 minutes and the total number of successful Toolkit API calls was measured. Then the number of calls per second was computed 'with cache' and 'without cache' test runs. Then a cache boost factor was calculated by diving (the number of API calls with cache) / (number of API calls without cache). Each cache test was executed 3 times, with different cache time-to-live values of 1 second, 5 seconds and 0 seconds (eternal cache, no expiration). The rationale is to see what impact different cache expiration/evi...

Debugging a Tridion 2011 Event System

OK, so you wrote your Tridion Event System. Now it's time to debug it. I know this is a hypothetical situtation -- your code never needs any kind of debugging ;) but indulge me... Recently, Alvin Reyes ( @nivlong ) blogged about being difficult to know how exactly to debug a Tridion Event System. More exactly, the question was " What process do I attach to for debugging even system code? ". Unfortunately, there is no simple or generic answer for it. Different events are fired by different Tridion CM modules. These modules run as different programs (or services) or run inside other programs (e.g. IIS). This means that you will need to monitor (or debug) different processes, based on which events your code handles. So the usual suspects are: dllhost.exe (or dllhost3g.exe ) - running as the MTSUser is the SDL Tridion Content Manager COM+ application and it fires events on generic TOM objects (e.g. events based on Tridion.ContentManager.Extensibility.Events.CrudEven...