Apply

Ready to go for it?

AI Apply speeds things up—apply directly if you prefer.

FREE ACCESS
5,000–10,000 jobs/day
Scoutfield Logo

See all jobs on Scoutfield

Search thousands of fresh jobs every day.

Discover
  • Fresh listings
  • Fast filters
  • No subscription required
Create a free account and start exploring right away.
PROS

Site Reliability Engineer II

PROS

. Administer, support, troubleshoot, and problem-solve complex systems and services .

Posted 10/5/2026full-timeSofia • BulgariaMid-LevelSeniorWebsite

Core Competencies

Role fit
Core Competencies

Use this summary to align your resume positioning with the role.

Demonstrates expertise in system administration, automation, and cloud environment optimization, with a strong focus on performance monitoring, incident response, and security best practices. Proficient in scripting and automation tools to enhance operational efficiency and reliability.

Highest-signal resume keywords
Advanced ScriptingCloud Environment OptimizationMonitoring and Alerting with Prometheus and GrafanaRESTful API Design and DevelopmentSystem Administration Experience

ATS Keywords

Tailor your resume
Applicant Tracking System Keywords

Tip: use these terms in your resume and cover letter to boost ATS matches.

Hard Skills
Operating SystemsNetworkingDatabase ManagementRubyGoJavaAutomation for DeploymentInfrastructure ManagementAPI Testing ToolsIT Security Best Practices
Soft Skills
Excellent CommunicationTime ManagementOrganizational SkillsCrisis ManagementProblem-Solving Skills
Tools & Technologies
PrometheusGrafanaPostmanAutomation ScriptsSelf-Service Tools
Certifications & Qualifications
Applicable IT Certifications
Industry Keywords
Cloud ServicesOpen-Source TechnologySystem EngineeringScripting LanguagesMultiple Cloud Provider Environments

Tech Stack

Tools & technologies
CloudGrafanaJavaPrometheusRubyGo

About the role

Key responsibilities & impact
  • Administer, support, troubleshoot, and problem-solve complex systems and services
  • Monitor service performance, reliability metrics, and infrastructure stability
  • Analyze system performance and identify improvement opportunities
  • Participate in disaster recovery testing and implement reliability enhancements
  • Define and maintain SLOs, visualizations, and alerts
  • Collaborate with product and development teams to resolve performance bottlenecks and reliability concerns
  • Implement and maintain automated deployments and self-service tools
  • Create and troubleshoot automation scripts for operational tasks
  • Use automation to improve system scalability and efficiency
  • Participate in follow-the-sun on-call rotations and respond to incidents
  • Troubleshoot production incidents, identify root causes, and create post-incident reports
  • Maintain documentation, user stories, and operational processes
  • Share knowledge through team sessions and contribute to continuous improvement
  • Implement security-auditing and vulnerability-mitigation automation
  • Collaborate with security teams to improve cloud security posture
  • Perform detailed post-incident analysis and documentation

Requirements

What you’ll need
  • Working knowledge of operating systems, networking, and database management
  • Advanced scripting and automation for deployment, scaling, and maintenance tasks
  • Proficiency in at least one high-level programming language: Ruby, Go, or Java
  • Knowledge of infrastructure and configuration management via automation
  • Advanced skills creating monitoring and alerting rules with Prometheus and Grafana
  • Ability to implement and optimize cloud environments
  • Knowledge of RESTful API design and development
  • Familiarity with API testing tools such as Postman
  • University degree in computer science or related field
  • Knowledge of IT security best practices and procedures
  • Excellent command of English
  • Applicable IT certifications preferred
  • System administrator experience preferred
  • Previous cloud services experience preferred, including open-source technology, software development, system engineering, scripting languages, and multiple cloud provider environments
  • Ability to work in a team and independently
  • Willingness to innovate, learn, and share knowledge
  • Excellent communication, time management, organizational, crisis management, and problem-solving skills

Benefits

Comp & perks
  • Flexible ways of working
  • Continuous learning
  • Opportunities to grow, innovate, and develop professionally