Amazon ECS now automatically detects and repairs container instances with impaired agent connectivity
Amazon ECS automates detection and recovery of agent connectivity issues, improving application availability
Amazon ECS now continuously monitors agent connections across container instances. It detects disconnections caused by infrastructure problems such as EBS volume degradation, host thermal events, or network failures. For AWS Fargate and ECS managed instances, ECS automatically drains tasks and launches replacements while deregistering affected instances. Customers using ECS on EC2 can utilize these health change events to trigger instance replacement workflows. This capability is available at no extra cost in all AWS commercial and GovCloud (US) regions
Why it matters
Users of container management services can benefit from reduced workload failures and improved application availability without manual intervention