Apply

Ready to go for it?

AI Apply speeds things up—apply directly if you prefer.

FREE ACCESS
5,000–10,000 jobs/day
Scoutfield Logo

See all jobs on Scoutfield

Search thousands of fresh jobs every day.

Discover
  • Fresh listings
  • Fast filters
  • No subscription required
Create a free account and start exploring right away.
Citi

SRE Observability Specialist – Vice President

Citi

. Define the roadmap for Engineering enablers for Project Orion aligned with enterprise reliability and SRE Services organization goals .

Posted 10/8/2026full-timePune • IndiaLeadWebsite

Core Competencies

Role fit
Core Competencies

Use this summary to align your resume positioning with the role.

Demonstrates expertise in observability engineering, focusing on operational telemetry, SLI/SLO implementation, and dashboard creation to enhance business outcomes. Proficient in leveraging observability tools and AI/ML capabilities to drive insights and improve system reliability across hybrid environments.

Highest-signal resume keywords
7+ Years Experience In SREHands-On Experience With Observability ToolsDeep Understanding Of SLIs And SLOsExperience Building Dashboards For Critical FlowsStrong Interpersonal And Collaboration Skills

ATS Keywords

Tailor your resume
Applicant Tracking System Keywords

Tip: use these terms in your resume and cover letter to boost ATS matches.

Hard Skills
Observability EngineeringOperational TelemetrySLIsSLOsError BudgetsTelemetry Best PracticesTroubleshooting Integration IssuesDashboard CreationAI/ML CapabilitiesTrace Correlation
Soft Skills
Interpersonal SkillsCollaboration Skills
Tools & Technologies
GrafanaPrometheusOpenTelemetryELKSplunkAWSGCPECSKubernetes
Industry Keywords
High-Availability EnvironmentsObservability PracticesHybrid PlatformsIncident WorkflowsBusiness Outcomes

Tech Stack

Tools & technologies
AWSCloudGoogle Cloud PlatformGrafanaKubernetesPrometheusSplunk

About the role

Key responsibilities & impact
  • Define the roadmap for Engineering enablers for Project Orion aligned with enterprise reliability and SRE Services organization goals
  • Translate organizational strategy into an actionable delivery plan with Services Products, Operations & Engineering
  • Understand critical business services functional scope and translate it into end-to-end monitoring solutions
  • Build scalable, reusable telemetry solutions for the Services Technology observability roadmap
  • Review and analyze application monitoring TOIL and collaborate with stakeholders on remediation
  • Create and maintain dashboards and visualizations for critical client journeys and real-time payment flows
  • Guide line-of-business teams in implementing SLIs/SLOs, golden signals, and effective alerting
  • Support observability tooling across on-premises, AWS/GCP public cloud, ECS, and Kubernetes environments
  • Customize shared dashboards and observability components with CTI and central Engineering functions
  • Provide technical support and implementation guidance to SREs and developers
  • Manage the observability book of work and drive initiatives to reduce MTTD and improve recovery outcomes
  • Connect line-of-business SREs with central infrastructure functions through feedback, issue escalation, and platform enhancement influence
  • Assess AI/ML-driven insights, anomaly detection, and emerging open-source observability practices
  • Advise teams on observability platform features and vendor offerings
  • Build AI adoption use cases for Orion L1 Functions and remediation using the Citi AI tech stack

Requirements

What you’ll need
  • 7+ years of experience in SRE, Observability Engineering, or platform infrastructure roles focused on operational telemetry
  • Hands-on experience with observability tools and stacks such as Grafana, Prometheus, OpenTelemetry, ELK, Splunk, and similar platforms
  • Deep understanding of SLIs, SLOs, Error Budgets, and telemetry best practices in high-availability environments
  • Proven ability to troubleshoot integration issues and support observability across hybrid platforms (on-prem, cloud, containers)
  • Experience building dashboards aligned to business outcomes and incident workflows, especially in critical flows like payments
  • Familiarity with modern observability tooling ecosystems, including AI/ML capabilities, trace correlation, baselining, and alert tuning
  • Strong interpersonal and collaboration skills — able to operate across federated engineering teams and central infrastructure groups
  • Experience in enablement or platform teams with a track record of scaling best practices across diverse business units
  • Bachelor’s degree in Computer Science, Engineering, or a related technical field, or equivalent practical experience

Benefits

Comp & perks
  • Equal opportunity employment
  • Reasonable accommodation for applicants with disabilities