FREE ACCESS
5,000–10,000 jobs/day
See all jobs on Scoutfield
Search thousands of fresh jobs every day.
Discover
- Fresh listings
- Fast filters
- No subscription required
Create a free account and start exploring right away.

Principal Azure Platform – Cloud Operations Architect
Kastech Canada. Assess and improve cloud operations processes, including provisioning, deployments, incident response, and change management .
Core Competencies
Role fitCore Competencies
Use this summary to align your resume positioning with the role.
Demonstrates expertise in cloud operations, particularly with Azure Kubernetes Service and Terraform for infrastructure as code. Proficient in CI/CD pipeline management and observability tooling to enhance service reliability and incident response.
Highest-signal resume keywords
Azure Kubernetes Service (AKS)Infrastructure as Code (IaC) with TerraformCI/CD Pipeline Management (Azure DevOps, Jenkins, GitHub Actions)Azure Networking and Secure ConnectivityGitOps with Argo CD and Helm
ATS Keywords
Tailor your resumeApplicant Tracking System Keywords
Tip: use these terms in your resume and cover letter to boost ATS matches.
Hard Skills
Cloud Platform EngineeringDevOps/SREKubernetes ManagementScripting/Automation (Python, Bash, PowerShell)Incident Response and Root Cause Analysis
Soft Skills
Problem-SolvingCommunicationMentoring
Tools & Technologies
Azure MonitorLog AnalyticsApplication InsightsArgo CDHelm
Certifications & Qualifications
Bachelor's or Master's Degree in Computer Science, Engineering, or Related Field
Industry Keywords
Cloud OperationsProduction-Grade SystemsDeployment StandardizationOperational MaturityChange Management
Tech Stack
Tools & technologiesAzureCloudJenkinsKubernetesNode.jsPythonTerraform
About the role
Key responsibilities & impact- Assess and improve cloud operations processes, including provisioning, deployments, incident response, and change management
- Lead the design and operational maturity of Azure Kubernetes Service environments, including cluster topology, node pools, upgrades, scaling, resiliency, ingress/egress, workload identity, secrets, and runtime security
- Lead Azure networking architecture and troubleshoot complex connectivity, network, and performance issues
- Design and implement Terraform infrastructure as code with reusable modules and safe lifecycle management
- Strengthen GitOps and deployment standardization using Argo CD and Helm
- Enhance CI/CD pipelines using Azure DevOps, Jenkins, or GitHub Actions with quality gates, validation, security scanning, and automated delivery
- Improve monitoring, logging, alerting, and dashboards using Azure Monitor, Log Analytics, and Application Insights
- Promote production readiness through runbooks, readiness reviews, and operational checklists
- Provide L3/L4 escalation support and lead high-severity incident triage, recovery, root cause analysis, and corrective actions
- Mentor and unblock Cloud Operations engineers during complex technical challenges
- Collaborate with Engineering, SRE, Security, and Delivery teams on operational patterns, platform guardrails, and production readiness
Requirements
What you’ll need- Bachelor's or master's degree in Computer Science, Engineering, or a related field
- 8+ years of hands-on experience in cloud platform engineering, DevOps/SRE, or cloud operations, with ownership of production-grade systems
- Strong hands-on experience with Azure, particularly Azure Kubernetes Service (AKS), and deep experience running Kubernetes in production (upgrades, scaling, failure modes, troubleshooting)
- Deep expertise in Azure networking and secure connectivity patterns, with the ability to diagnose complex multi-layer issues across AKS + network + application boundaries
- Proven hands-on experience implementing IaC with Terraform (modules, state strategy, environment consistency, safe rollout practices)
- Strong experience with GitOps and deployment tooling, including Argo CD and Helm, and a strong understanding of release strategies and operational controls
- Proficiency managing CI/CD pipelines and automation (Azure Pipelines, Jenkins, GitHub Actions) and improving deployment reliability through automated checks and gates
- Hands-on experience with Azure observability tooling (Azure Monitor, Log Analytics, Application Insights) to improve service health visibility and incident response effectiveness
- Proficiency in scripting/automation with Python and/or Bash/PowerShell to build operational tooling and reduce repetitive manual work
- Strong problem-solving and communication skills, with the ability to operate calmly under pressure and guide teams through technical challenges