Apply

Ready to go for it?

AI Apply speeds things up—apply directly if you prefer.

FREE ACCESS
5,000–10,000 jobs/day
Scoutfield Logo

See all jobs on Scoutfield

Search thousands of fresh jobs every day.

Discover
  • Fresh listings
  • Fast filters
  • No subscription required
Create a free account and start exploring right away.
Azumo

Cloud DevOps Engineer

Azumo

. Own production infrastructure for AI systems, including clusters, deployment pipelines, and monitoring .

Posted 9/25/2026full-timeRemote • Argentina, Brazil, Uruguay, Chile, ColombiaMid-LevelSeniorWebsite

Core Competencies

Role fit
Core Competencies

Use this summary to align your resume positioning with the role.

Demonstrates expertise in managing production infrastructure for AI systems, including Kubernetes, CI/CD pipelines, and infrastructure as code with Terraform. Proficient in security hardening, cost management, and incident response within compliance frameworks such as SOC 2 and HIPAA.

Highest-signal resume keywords
Production Kubernetes ExperienceInfrastructure As Code Using TerraformCI/CD Pipeline Development Using GitHub ActionsMonitoring And Incident Response Using DatadogCloud Deployment Experience On AWS, Azure Or GCP

ATS Keywords

Tailor your resume
Applicant Tracking System Keywords

Tip: use these terms in your resume and cover letter to boost ATS matches.

Hard Skills
Linux AdministrationNetworkingScripting In BashPythonGoInfrastructure Cost ManagementSecurity Hardening Of Linux And KubernetesContainer Image ManagementService Mesh ExperienceDatabase Operations With PostgreSQL
Soft Skills
Clear Written And Spoken EnglishAbility To Explain Technical Trade-Offs
Tools & Technologies
TerraformGitHub ActionsDatadogCloudWatchAI-Assisted Coding ToolsKubernetesHelmKustomizeAWSAzure
Certifications & Qualifications
Cloud Certifications
Industry Keywords
SOC 2HIPAAAI WorkloadsInfrastructure DriftIncident ResponseCost ProfilesTraffic ManagementReproducible InfrastructureChange ProcessesClient Environments

Tech Stack

Tools & technologies
AWSAzureCloudDynamoDBGoogle Cloud PlatformKubernetesLinuxMongoDBPostgresPythonTerraformGo

About the role

Key responsibilities & impact
  • Own production infrastructure for AI systems, including clusters, deployment pipelines, and monitoring
  • Provision, upgrade, network, and configure production Kubernetes clusters
  • Build reproducible infrastructure as code with Terraform or equivalent and detect infrastructure drift
  • Build and maintain reliable CI/CD pipelines and standardized container image workflows
  • Implement monitoring and alerting, investigate incidents, determine root causes, and make preventive changes
  • Provide infrastructure for AI workloads, including inference services and their scaling and cost profiles
  • Analyze infrastructure costs and capacity needs, including projected costs at higher traffic levels
  • Secure Linux, Kubernetes, containers, and service meshes, including secrets, access, and audit trails
  • Develop internal tooling and documentation and support developers and QA during release cycles
  • Work within client environments, repositories, cloud accounts, and change processes when required
  • Operate within Azumo's SOC 2-certified environment and accommodate engagement requirements such as HIPAA
  • Use AI-assisted engineering tools and automated codebase audits to assess security, cost, and architecture findings

Requirements

What you’ll need
  • 5+ years as a DevOps, SRE or systems engineer running production infrastructure
  • Linux administration, networking, Git, and scripting in Bash plus Python or Go
  • Production Kubernetes experience, including provisioning, upgrading and debugging clusters
  • Infrastructure as code using Terraform or equivalent, including state, modules and environment parity
  • Experience building and maintaining CI/CD pipelines using GitHub Actions, GitLab or equivalent
  • Responsibility for container images
  • Cloud deployment experience on AWS, Azure or GCP, including managed Kubernetes services and associated cost models
  • Monitoring and incident response experience using Datadog, CloudWatch or equivalent
  • Experience taking incidents from alert through root cause analysis to preventive changes
  • Security hardening of Linux, containers and Kubernetes
  • Infrastructure cost and capacity management experience
  • Active use of AI-assisted coding tools such as Claude Code, Cursor, or GitHub Copilot in delivery work
  • Clear written and spoken English at C1 or above
  • Ability to explain technical trade-offs directly to a client
  • Bachelor's degree in Computer Science, a related field, or equivalent professional experience
  • Preferred: service mesh and traffic management experience with Istio, Linkerd or equivalent
  • Preferred: Helm, Kustomize or equivalent templating and environment configuration
  • Preferred: production-scale database operations with PostgreSQL, MongoDB, RDS, DynamoDB or equivalent
  • Preferred: running or scaling inference workloads and evaluating cost and latency trade-offs against hosted APIs
  • Preferred: delivery under compliance regimes such as SOC 2 or HIPAA
  • Preferred: cloud certifications, open-source infrastructure contributions, or published technical writing

Benefits

Comp & perks
  • 100% remote-first culture (work anywhere in Latin America)
  • Paid time off (PTO)
  • U.S. Holidays
  • Solid AI Training and certification
  • Mentored career development
  • Profit sharing
  • $US remuneration
  • Maternity coverage