Apply

Ready to go for it?

AI Apply speeds things up—apply directly if you prefer.

FREE ACCESS
5,000–10,000 jobs/day
Scoutfield Logo

See all jobs on Scoutfield

Search thousands of fresh jobs every day.

Discover
  • Fresh listings
  • Fast filters
  • No subscription required
Create a free account and start exploring right away.
Allstate

Cloud Platform Lead Consultant

Allstate

. Build, operate, and optimize data streaming infrastructure using Amazon MSK (Kafka) or Google Pub/Sub .

Posted 10/7/2026full-timeRemote • United StatesSenior💰 $100,000 - $170,500 per yearWebsite

Core Competencies

Role fit
Core Competencies

Use this summary to align your resume positioning with the role.

Demonstrates expertise in building and optimizing data streaming infrastructure, managing cloud environments, and implementing infrastructure-as-code practices. Proficient in monitoring and incident management, with a strong foundation in distributed systems principles.

Highest-signal resume keywords
Amazon MSK (Kafka) ExperienceTerraform for Infrastructure-as-CodePython and Shell ScriptingPostgreSQL Database ManagementCI/CD Pipeline Implementation

ATS Keywords

Tailor your resume
Applicant Tracking System Keywords

Tip: use these terms in your resume and cover letter to boost ATS matches.

Hard Skills
Data Streaming InfrastructureCloud Infrastructure ManagementDistributed Systems PrinciplesMonitoring and Alerting ImplementationRelational and NoSQL DatabasesIncident Response and Root Cause AnalysisAutomation with AnsibleContainer Orchestration with KubernetesNetworking FundamentalsSecurity Fundamentals
Soft Skills
Problem-SolvingCommunicationDocumentationOwnership in Incident Management
Tools & Technologies
AWSGoogle Cloud PlatformJenkinsGitStarburst GalaxyAWS AthenaApache KafkaConfluent KafkaCI/CD ToolsCloud Automation Tools
Industry Keywords
SREDevOpsPlatform EngineeringHigh AvailabilityFault ToleranceDisaster RecoveryCapacity PlanningCost OptimizationCloud AutomationIncident Management

Tech Stack

Tools & technologies
AnsibleApacheAWSCloudDistributed SystemsGoogle Cloud PlatformJenkinsKafkaKubernetesNoSQLPostgresPythonShell ScriptingTerraform

About the role

Key responsibilities & impact
  • Build, operate, and optimize data streaming infrastructure using Amazon MSK (Kafka) or Google Pub/Sub
  • Design, deploy, and manage highly available database and caching platforms across multi-cloud environments
  • Develop and maintain infrastructure-as-code, CI/CD pipelines, and cloud automation using Terraform and Python
  • Implement monitoring, alerting, and observability for data platform services
  • Partner with application development teams to troubleshoot, tune, and optimize application performance and data access layers
  • Administer and optimize Starburst Galaxy and AWS Athena
  • Participate in incident response, root cause analysis, and post-incident reviews
  • Share on-call rotation to support mission-critical data infrastructure
  • Review application source code and collaborate on targeted fixes and optimization guidance
  • Evaluate emerging tools and AI-assisted automation approaches
  • Contribute to capacity planning, disaster recovery, security hardening, and cost optimization

Requirements

What you’ll need
  • 3–5 years of software engineering or infrastructure experience, with at least 2 years in SRE, DevOps, or platform engineering operating production systems at scale
  • Hands-on experience designing, deploying, and managing cloud infrastructure on AWS and/or Google Cloud Platform, including networking, identity, and security fundamentals
  • Production experience operating data streaming platforms; hands-on work with Apache Kafka, including Amazon MSK or Confluent Kafka, and understanding of partitions, consumer groups, delivery semantics, and backpressure
  • Production experience with relational and NoSQL databases; PostgreSQL required
  • Strong Python and Shell scripting for automation and operational solutions
  • Strong experience with infrastructure-as-code such as Terraform, CI/CD such as Jenkins and Git, Ansible, and container orchestration such as Kubernetes in production environments
  • Experience implementing and automating monitoring, logging, and alerting for distributed systems
  • Proven ability to contribute to root cause analysis for production incidents
  • Strong problem-solving, communication, and documentation skills with ownership in on-call and incident management environments
  • Working understanding of distributed systems principles including high availability, fault tolerance, consistency models, and disaster recovery
  • Background investigation required
  • Dedicated, private workspace and reliable internet with minimum speeds of 50 MB download and 5 MB upload when working from home

Benefits

Comp & perks
  • Comprehensive technology setup including a laptop, monitors, headset, keyboard, and mouse
  • Monthly connectivity reimbursement to help offset internet costs for employees eligible to work from home