Apply

Ready to go for it?

AI Apply speeds things up—apply directly if you prefer.

FREE ACCESS
5,000–10,000 jobs/day
Scoutfield Logo

See all jobs on Scoutfield

Search thousands of fresh jobs every day.

Discover
  • Fresh listings
  • Fast filters
  • No subscription required
Create a free account and start exploring right away.
Broadridge

Global Manager – Site Reliability Engineering

Broadridge

. Build and lead a high-performing SRE team, including hiring, coaching, performance management, career development, and succession planning .

Posted 9/30/2026full-timeNew York City • New York • United StatesSeniorLead💰 $235,000 - $250,000 per yearWebsite

Core Competencies

Role fit
Core Competencies

Use this summary to align your resume positioning with the role.

Demonstrates expertise in building and leading high-performing SRE teams, with a focus on AWS infrastructure, Java and Spring Boot development, and operational excellence. Proficient in implementing CI/CD practices, improving system reliability, and managing complex incident responses.

Highest-signal resume keywords
SRE Team LeadershipJava And Spring Boot DevelopmentAWS Infrastructure ManagementKafka Event-Driven DesignPostgreSQL Performance Optimization

ATS Keywords

Tailor your resume
Applicant Tracking System Keywords

Tip: use these terms in your resume and cover letter to boost ATS matches.

Hard Skills
JavaSpring BootAWSKafkaPostgreSQLCI/CDInfrastructure As CodeAutomated TestingPerformance TuningRelease Engineering
Soft Skills
People ManagementMentoringCommunicationIncident Response CoordinationStakeholder Influence
Tools & Technologies
KubernetesTerraformOpenTelemetryKafka ConnectECSEKSIAM
Industry Keywords
Financial ServicesRegulated SystemsDistributed SystemsService-Level ObjectivesOperational Troubleshooting

Tech Stack

Tools & technologies
AWSDistributed SystemsJavaKafkaKubernetesPostgresSpringSpring BootSpringBootTerraform

About the role

Key responsibilities & impact
  • Build and lead a high-performing SRE team, including hiring, coaching, performance management, career development, and succession planning
  • Own the SRE strategy and execution roadmap, translating platform and business priorities into measurable reliability, release, automation, and scalability improvements
  • Lead architecture and code reviews, complex troubleshooting, automation, and tooling decisions
  • Develop reusable AWS infrastructure-as-code patterns, environment configurations, automated provisioning, and recovery capabilities
  • Improve Kafka reliability and event-processing performance, including partitioning, consumer groups, schema evolution, delivery semantics, replay, and failure recovery
  • Strengthen PostgreSQL performance and resilience through schema design, indexing, query optimization, transaction management, connection pooling, migrations, and recovery testing
  • Lead release engineering and deployment readiness by improving CI/CD, automated testing, security checks, artifact traceability, production validation, and rollback or roll-forward patterns
  • Define service-level indicators, service-level objectives, and error-budget practices; use metrics, logs, and distributed traces to improve service health
  • Reduce operational toil through reusable tooling, self-service capabilities, and controlled remediation
  • Enable responsible AI adoption in code and test development, infrastructure review, knowledge retrieval, and incident investigation
  • Coordinate complex incident response, communicate impact and recovery progress, conduct blameless reviews, and drive corrective actions
  • Validate capacity, failover, backup restoration, and disaster recovery against agreed objectives
  • Influence senior stakeholders on technical risk, investment trade-offs, production readiness, secure delivery, and reliable service operation

Requirements

What you’ll need
  • 10+ years of software engineering or closely related production engineering experience
  • Substantial experience building and operating Java-based services in production
  • Engineering people-management experience, including hiring, performance management, talent development, resource planning, and accountability for complex technical delivery
  • Experience leading senior engineers and developing technical leads
  • Strong Java and Spring Boot skills, including API design, concurrency, performance tuning, automated testing, and secure coding
  • Hands-on AWS experience with IAM, networking, observability, infrastructure as code, and ECS, EKS, or Lambda
  • Production Kafka experience covering event-driven design, consumer groups, partitioning, schema evolution, delivery semantics, and failure recovery
  • Strong PostgreSQL skills in schema design, indexing, query optimization, transactions, migrations, and operational troubleshooting
  • Ability to build maintainable automation and operational tooling using version control, peer review, automated testing, and reusable design
  • Experience with CI/CD, release engineering, automated testing, monitoring, incident response, and reliable distributed-system design
  • Understanding of service-level objectives, capacity planning, production readiness, recovery strategies, and delivery and reliability measures
  • Ability to lead technical design reviews, mentor engineers, and communicate architectural, operational, and delivery trade-offs
  • Preferred: Kubernetes, Terraform, OpenTelemetry, Kafka Connect, and PostgreSQL replication or high availability
  • Preferred: Experience with financial-services platforms or other regulated, business-critical distributed systems

Benefits

Comp & perks
  • Bonus eligible
  • Comprehensive benefit offerings
  • Equal employment opportunities
  • Reasonable accommodations during the application and hiring process