FREE ACCESS
5,000–10,000 jobs/day
See all jobs on Scoutfield
Search thousands of fresh jobs every day.
Discover
- Fresh listings
- Fast filters
- No subscription required
Create a free account and start exploring right away.
Core Competencies
Role fitCore Competencies
Use this summary to align your resume positioning with the role.
Demonstrates expertise in observability engineering, focusing on operational telemetry, SLI/SLO implementation, and dashboard creation to enhance business outcomes. Proficient in leveraging observability tools and AI/ML capabilities to drive insights and improve system reliability across hybrid environments.
Highest-signal resume keywords
7+ Years Experience In SREHands-On Experience With Observability ToolsDeep Understanding Of SLIs And SLOsExperience Building Dashboards For Critical FlowsStrong Interpersonal And Collaboration Skills
ATS Keywords
Tailor your resumeApplicant Tracking System Keywords
Tip: use these terms in your resume and cover letter to boost ATS matches.
Hard Skills
Observability EngineeringOperational TelemetrySLIsSLOsError BudgetsTelemetry Best PracticesTroubleshooting Integration IssuesDashboard CreationAI/ML CapabilitiesTrace Correlation
Soft Skills
Interpersonal SkillsCollaboration Skills
Tools & Technologies
GrafanaPrometheusOpenTelemetryELKSplunkAWSGCPECSKubernetes
Industry Keywords
High-Availability EnvironmentsObservability PracticesHybrid PlatformsIncident WorkflowsBusiness Outcomes
Tech Stack
Tools & technologiesAWSCloudGoogle Cloud PlatformGrafanaKubernetesPrometheusSplunk
About the role
Key responsibilities & impact- Define the roadmap for Engineering enablers for Project Orion aligned with enterprise reliability and SRE Services organization goals
- Translate organizational strategy into an actionable delivery plan with Services Products, Operations & Engineering
- Understand critical business services functional scope and translate it into end-to-end monitoring solutions
- Build scalable, reusable telemetry solutions for the Services Technology observability roadmap
- Review and analyze application monitoring TOIL and collaborate with stakeholders on remediation
- Create and maintain dashboards and visualizations for critical client journeys and real-time payment flows
- Guide line-of-business teams in implementing SLIs/SLOs, golden signals, and effective alerting
- Support observability tooling across on-premises, AWS/GCP public cloud, ECS, and Kubernetes environments
- Customize shared dashboards and observability components with CTI and central Engineering functions
- Provide technical support and implementation guidance to SREs and developers
- Manage the observability book of work and drive initiatives to reduce MTTD and improve recovery outcomes
- Connect line-of-business SREs with central infrastructure functions through feedback, issue escalation, and platform enhancement influence
- Assess AI/ML-driven insights, anomaly detection, and emerging open-source observability practices
- Advise teams on observability platform features and vendor offerings
- Build AI adoption use cases for Orion L1 Functions and remediation using the Citi AI tech stack
Requirements
What you’ll need- 7+ years of experience in SRE, Observability Engineering, or platform infrastructure roles focused on operational telemetry
- Hands-on experience with observability tools and stacks such as Grafana, Prometheus, OpenTelemetry, ELK, Splunk, and similar platforms
- Deep understanding of SLIs, SLOs, Error Budgets, and telemetry best practices in high-availability environments
- Proven ability to troubleshoot integration issues and support observability across hybrid platforms (on-prem, cloud, containers)
- Experience building dashboards aligned to business outcomes and incident workflows, especially in critical flows like payments
- Familiarity with modern observability tooling ecosystems, including AI/ML capabilities, trace correlation, baselining, and alert tuning
- Strong interpersonal and collaboration skills — able to operate across federated engineering teams and central infrastructure groups
- Experience in enablement or platform teams with a track record of scaling best practices across diverse business units
- Bachelor’s degree in Computer Science, Engineering, or a related technical field, or equivalent practical experience
Benefits
Comp & perks- Equal opportunity employment
- Reasonable accommodation for applicants with disabilities
