Apply

Ready to go for it?

AI Apply speeds things up—apply directly if you prefer.

FREE ACCESS
5,000–10,000 jobs/day
Scoutfield Logo

See all jobs on Scoutfield

Search thousands of fresh jobs every day.

Discover
  • Fresh listings
  • Fast filters
  • No subscription required
Create a free account and start exploring right away.
Nomura

Observability SRE, Site Reliability Engineer

Nomura

. Administer and support the production environment .

Posted 9/22/2026full-timePhiladelphia • Pennsylvania • United StatesJuniorMid-Level💰 $115,000 - $135,000 per yearWebsite

Core Competencies

Role fit
Core Competencies

Use this summary to align your resume positioning with the role.

Demonstrates expertise in administering and supporting production environments, with a strong focus on observability, incident management, and change management. Proficient in leveraging modern tools and technologies to enhance system reliability and operational efficiency.

Highest-signal resume keywords
Grafana AdministrationLinux OS TroubleshootingProduction Support ExperienceIncident ManagementObservability Best Practices

ATS Keywords

Tailor your resume
Applicant Tracking System Keywords

Tip: use these terms in your resume and cover letter to boost ATS matches.

Hard Skills
PythonAnsibleSQLOpenTelemetryKubernetesDockerCI/CD ToolsRDBMS ConceptsChange ManagementRelease Management
Soft Skills
Good Communication SkillsInterpersonal SkillsAnalytical SkillsTroubleshooting SkillsMature Judgment
Tools & Technologies
GrafanaConfluenceJIRAGitLabJenkinsNexusMiddlewareWeb ServersLoad BalancersDirectory Services
Certifications & Qualifications
ITIL
Industry Keywords
ObservabilityTelemetryMonitoringProduction EnvironmentGlobal Agile TeamScalable SystemsIncident ManagementProblem ManagementUser EngagementCloud Platforms

Tech Stack

Tools & technologies
AnsibleCloudDockerGrafanaJenkinsKubernetesLinuxMySQLPythonRDBMSSQL

About the role

Key responsibilities & impact
  • Administer and support the production environment
  • Engineer reliability into supported products and services, including the monitoring and observability platform
  • Shape future monitoring strategy and direction within the Nomura Group
  • Provide cross-functional engagement and support adoption of the Telemetry, Observability and Monitoring platform
  • Understand observability and notification tools and frameworks and assist development and production support teams with usage issues
  • Act as custodian of the production environment and build and maintain robust, scalable, highly available systems
  • Prevent production incidents and perform incident management, problem management, and root-cause analysis
  • Push changes and releases reliably through effective change and release management
  • Respond to alerts quickly and prevent recurrence
  • Triage alerts, requests, and emails according to priority
  • Engage with users and provide guidance for a good customer experience
  • Collaborate in a global agile team, participating in sprint planning, reviews, and continuous improvement
  • Build and maintain scalable, reliable monitoring solutions for Nomura’s global infrastructure
  • Contribute to architectural decisions affecting the observability platform
  • Champion observability best practices across the organization
  • Partner with engineers to optimize operational efficiency and system resilience
  • Use AI tools such as Claude and Copilot with appropriate guardrails
  • Mentor and guide other SREs

Requirements

What you’ll need
  • Minimum 2 years’ experience with Grafana or another modern observability tool in an administrative capacity for a medium/large-scale enterprise
  • At least 2 years’ exposure to Linux OS with general-purpose troubleshooting and day-to-day commands
  • Exposure to Python and/or Ansible
  • Production support experience, including request handling, incident management, problem management, change management, release management, on-call handling, user engagement, and responding to alerts
  • Good communication and interpersonal skills
  • Strong analytical and troubleshooting skills with mature judgment
  • Basic understanding of cloud platforms
  • Basic understanding of CI/CD tools such as GitLab, Jenkins, Ansible, and Nexus
  • Understanding of OpenTelemetry standards
  • Understanding of containerization technologies such as Kubernetes, EKS, and Docker
  • Experience supporting a medium/large-scale production environment
  • Knowledge of ITIL
  • Understanding of database platforms such as Sybase, MySQL, and MSSQL; general RDBMS concepts and SQL
  • Familiarity with Confluence and JIRA
  • Familiarity with infrastructure technologies such as middleware, web servers, load balancers, and directory services
  • Experience working with a globally dispersed team

Benefits

Comp & perks
  • Sign-on bonus
  • Restricted stock units
  • Discretionary awards
  • Medical benefits
  • Financial benefits
  • Other benefits
  • 401(k) eligibility
  • Vacation time
  • Sick time
  • Parental leave