FREE ACCESS
5,000–10,000 jobs/day
See all jobs on Scoutfield
Search thousands of fresh jobs every day.
Discover
- Fresh listings
- Fast filters
- No subscription required
Create a free account and start exploring right away.
Core Competencies
Role fitCore Competencies
Use this summary to align your resume positioning with the role.
Demonstrates extensive experience in DevOps and Site Reliability Engineering, with a strong focus on AWS infrastructure, CI/CD pipeline design, and observability practices. Proficient in applying SRE principles to enhance service reliability and operational efficiency.
Highest-signal resume keywords
AWS Infrastructure ManagementCI/CD Pipeline DesignSRE Principles ApplicationObservability and MonitoringDocker and Containerization
ATS Keywords
Tailor your resumeApplicant Tracking System Keywords
Tip: use these terms in your resume and cover letter to boost ATS matches.
Hard Skills
DevOpsSite Reliability EngineeringInfrastructure EngineeringPlatform EngineeringCI/CD SystemsAWSDockerLinuxNetworkingSecurity
Soft Skills
Clear CommunicationCollaboration with Distributed Teams
Tools & Technologies
CircleCIGitHub ActionsJenkinsGitLab CIDatadogNew RelicDynatraceGrafana CloudSplunk
Industry Keywords
Cloud InfrastructureOperational ReadinessIncident ResponseRoot-Cause AnalysisCapacity Planning
Tech Stack
Tools & technologiesAWSCloudDockerGrafanaJenkinsLinuxSplunk
About the role
Key responsibilities & impact- Evolve infrastructure, delivery systems, and operational practices enabling engineering teams to build and run reliable software
- Establish scalable approaches to cloud infrastructure, CI/CD, observability, alerting, operational readiness, and maintenance
- Partner with software engineers to improve developer experience and strengthen reliability, security, performance, and cost efficiency of Karat’s hosted SaaS platform
- Own and evolve Karat’s AWS SaaS infrastructure
- Design, improve, and operate CI/CD pipelines using CircleCI and related tooling
- Build and maintain observability capabilities including metrics, logs, traces, dashboards, and actionable alerting using Datadog and related tools
- Apply SRE principles to service reliability, availability, performance, capacity planning, incident response, root-cause analysis, and operational learning
- Influence engineering-wide technical decisions and delivery practices through partnership and standards
Requirements
What you’ll need- 5+ years of experience in DevOps, Site Reliability Engineering, infrastructure engineering, platform engineering, or a closely related discipline
- Significant hands-on production experience with AWS; required qualification
- Experience designing, operating, and improving CI/CD systems using CircleCI, GitHub Actions, Jenkins, GitLab CI, or another major CI/CD platform
- Strong experience with Docker and containerized application environments
- Strong Linux, networking, security, and cloud-infrastructure fundamentals
- Practical experience applying SRE principles to production systems, including observability, alerting, incident response, root-cause analysis, capacity planning, and reliability improvement
- Hands-on experience with a leading telemetry and observability platform, such as Datadog, New Relic, Dynatrace, Grafana Cloud, or Splunk
- Experience designing monitoring and alerting systems
- Experience with cloud cost management and optimization
- Experience partnering with application-engineering teams
- Clear written and verbal English communication skills
- Comfort working with globally distributed teams and regularly collaborating with colleagues in the United States
- Must reside in Bengaluru (formerly known as Bangalore), India
- Schedule must overlap with U.S. business hours
- Application submissions must be 100% in English
Benefits
Comp & perks- 100% remote work
- Inclusive workplace and non-discrimination commitment
- Accommodation support for disabilities or special needs
