Manager, Site Reliability Engineer

Full Time
  • December 4, 2026
  • Employment Info

    Responsibilities: 

    • Manage Forge’s Site Reliability Engineering team responsible for keeping Forge systems highly available for customers.
    • Drive strong incident management practices in partnership with engineering teams, including response, mitigation, follow-up, and post-incident learning.
    • Build, improve, and manage observability infrastructure in partnership with Platform Engineering, including monitoring, alerting, dashboards, and operational metrics.
    • Improve monitoring coverage and alert quality to reduce noise, shorten time to detect, and support faster response and mitigation.
    • Champion reliability best practices across engineering, including service ownership, operational readiness, disaster recovery, and production support standards.
    • Contribute to technical design, architecture, automation, infrastructure, and overall team delivery.
    • Collaborate with engineering teams to troubleshoot production issues, identify recurring problems, and improve system reliability.
      • Hire, coach, mentor, and manage performance for SRE team members while supporting career development and team health.

      • Partner with Security, Compliance, and Risk partners to ensure reliability and infrastructure practices meet the needs of a regulated business.

    Qualifications:

    • 5+ years of experience leading a Site Reliability Engineering, DevOps, Cloud Operations, or similar reliability-focused function.
    • 10+ years of total software engineering, infrastructure, platform, cloud, or production operations experience.
    • Bachelor’s degree in Computer Science, Engineering, or a closely related field, or equivalent practical experience.
    • Experience building, operating, and maintaining large-scale cloud infrastructure and distributed systems.
    • Hands-on experience with observability, monitoring, alerting, incident response, troubleshooting, and production support.
    • Experience with CI/CD, infrastructure automation, cloud platforms, and operational tooling.
    • Strong technical judgment, communication skills, and ability to influence across engineering and non-engineering stakeholders.

     

    Are you interested in this position?

    Apply by clicking on the “Apply Now” button below!

    #FintechCareersGCC
    #GCCFintechJobs
    #TechOpportunitiesGCC
    #FinanceTechGulf
    #GCCJobSearch
    #FintechOpportunities
    #GulfCareerHub
    #TechJobsGCC
    #FintechGCC
    #CareerInFinanceGCC

    Featured jobs

    Get the latest news, updates and tips

    Quantitative Risk Analyst
    Full Time

    Quantitative Risk Analyst