Thank you for sending your enquiry! One of our team members will contact you shortly.
Thank you for sending your booking! One of our team members will contact you shortly.
Course Outline
Core Foundations of AWS Cloud Operations
- Defining operational roles and responsibilities in cloud environments.
- AWS account structures, organizational setup, and multi-account strategies.
- Key operational services: CloudWatch, CloudTrail, and AWS Config.
Infrastructure as Code (IaC) and Provisioning
- Principles of IaC and the concept of immutable infrastructure.
- Provisioning resources using Terraform and AWS CloudFormation.
- Managing state, modules, and promoting configurations across environments.
CI/CD and Deployment Methodologies
- Designing CI/CD pipelines for cloud-native applications.
- Implementing blue/green, canary, and rolling deployment techniques.
- Automating rollbacks, health checks, and release validation.
Monitoring, Observability, and Alerting
- Handling metrics, logs, and traces: ingestion, storage, and analysis.
- Utilizing CloudWatch, X-Ray, and third-party observability solutions.
- Establishing SLOs/SLIs, alerting policies, and on-call workflows.
Security Operations and Identity Governance
- IAM best practices, least privilege principles, and cross-account access.
- Managing secrets, KMS, and secure parameter stores.
- Operational security measures: patching, vulnerability scanning, and audit trails.
Resilience, Backup, and Disaster Recovery
- Designing for fault tolerance and high availability.
- Backup strategies, automated snapshots, and restoration procedures.
- Disaster recovery planning and the creation of operational runbooks.
Cost Optimization and Governance
- Cost visibility through billing, tagging, and allocation strategies.
- Rightsizing resources, reserved instances/savings plans, and budget controls.
- Governance through policies, guardrails, and compliance automation.
Containers, Serverless, and Runtime Management
- Operational considerations for ECS, EKS, and Lambda.
- Service discovery, autoscaling, and resource constraints.
- Logging, tracing, and debugging containerized workloads.
Incident Response, Playbooks, and Chaos Engineering
- Runbook-based incident response and post-mortem analysis.
- Automating remediation and self-healing mechanisms.
- Introduction to chaos experiments for resilience validation.
Practical Workshop: Managing a Sample Workload
- Deploying a sample application using IaC and a CI/CD pipeline.
- Setting up monitoring, alerts, and automated remediation scripts.
- Simulating incidents and practicing runbook-driven response.
Wrap-up and Future Pathways
Requirements
- Foundational knowledge of cloud concepts and networking.
- Proficiency with the Linux command line and scripting.
- Experience with source control (Git) and fundamental CI/CD principles.
Target Audience
- Cloud operations engineers.
- SREs and platform engineers.
- DevOps engineers and technical leaders.
21 Hours
Testimonials (1)
I've find out new interesting things about Lambda and Serverless