FREE ACCESS
5,000–10,000 jobs/day
See all jobs on Scoutfield
Search thousands of fresh jobs every day.
Discover
- Fresh listings
- Fast filters
- No subscription required
Create a free account and start exploring right away.
Core Competencies
Role fitCore Competencies
Use this summary to align your resume positioning with the role.
Demonstrates expertise in advanced Terraform and AWS environments, with a strong focus on automation, reliability, and incident response. Capable of collaborating with engineering teams to enhance service availability and streamline operations while integrating AI tools.
Highest-signal resume keywords
Advanced Terraform ExpertiseAWS Multi-Account ManagementCI/CD Workflow DevelopmentIncident Response and ObservabilityAI Services Integration
ATS Keywords
Tailor your resumeApplicant Tracking System Keywords
Tip: use these terms in your resume and cover letter to boost ATS matches.
Hard Skills
TerraformAWS ECSAWS RDSAWS ALBAWS VPCAWS IAMAWS Route53AWS LambdaAWS DynamoDBBash/Python Scripting
Soft Skills
CollaborationProblem-SolvingDocumentation
Tools & Technologies
GitHub ActionsDockerCloudWatchSecrets ManagerKMS
Industry Keywords
Service Level BaselinesDisaster RecoveryRoot Cause AnalysisOperational Toil ReductionInfrastructure Topology
Tech Stack
Tools & technologiesAWSCloudDockerDynamoDBLinuxPythonTerraform
About the role
Key responsibilities & impact- Create monitoring queries and establish service level baselines
- Support senior engineers during incidents
- Contribute to post-mortems and root cause analyses
- Participate in disaster recovery tests
- Implement automation and execute code in production environments
- Contribute to SRE knowledge documentation
- Support deployment, monitoring, and reliability of services integrating AI tools
- Support architecture and senior engineers in creating infrastructure topology drawings and deployment workflows
- Test availability, reliability, and recoverability in non-production environments
- Lead complex reliability initiatives and drive automation to reduce operational toil
- Design and implement solutions that improve service availability, streamline operations, and enhance system recovery capabilities
- Collaborate with engineering teams and host-function stakeholders
- Support handover and capability-building so solutions remain owned and operable after the squad moves on
Requirements
What you’ll need- Expertise in advanced Terraform, including modules, providers, state management, lifecycle controls, drift detection, safe refactoring, remote state, locking, and cross-stack dependencies
- Hands-on experience managing production, multi-account, multi-region AWS environments across ECS, RDS, ALB, VPC, IAM, Route53, ECR, S3, Lambda, DynamoDB, SQS, Secrets Manager, KMS, and CloudWatch
- Experience building and troubleshooting reusable GitHub Actions CI/CD workflows, OIDC authentication, approval gates, runners, Terraform deployments, application deployments, and migration pipelines
- Knowledge of Docker, ECR, ECS task definitions and services, IAM roles, health checks, autoscaling, ALB integration, and deployment rollbacks
- Proficiency in AWS networking and security, including VPCs, networking, ALBs, Route53, ACM/TLS, IAM, OIDC, Secrets Manager, KMS, and cloud security best practices
- Skilled in incident response and observability using logs, metrics, alarms, deployment history, root cause analysis, rollback decisions, and operational runbooks
- Strong Linux and Git fundamentals with Bash/Python scripting for AWS CLI automation, CI/CD, and operational tooling
- Hands-on experience integrating and operating AI services and APIs in production, including monitoring, reliability, and security practices for AI-powered features
- Ability to support multiple engineering teams, troubleshoot across infrastructure and application layers, document solutions, and enable secure self-service practices
Benefits
Comp & perks- Annual incentive bonus
- Country-specific benefits
- Accommodation or adjustment support during the hiring process
