EU remote
Engineering Team Leader (Site Reliability Engineering)
About this role
We are building XTB – a global investment company offering innovative technological solutions that allow our clients to effectively manage their finances in multiple ways. All of this within a single, intuitive XTB app already used by over one million users worldwide! We are a certified Great Place to Work company. We are looking for an Engineering Team Leader to drive the development and growth of the Site Reliability Engineering team.
In this role, you will have the opportunity to shape the technical and operational direction of SRE practices, lead a resilience strategy, and play a key role in defining and delivering solutions that ensure the reliability and scalability of XTB systems for millions of clients across a growing organization. Responsibilities Team Leadership: Shape and grow a high-performing Site Reliability Engineering team, fostering an environment of technical excellence, ownership, and continuous improvement.
Reliability Strategy: Define and drive the SRE platform strategy in close collaboration with infrastructure and development teams, ensuring alignment with organizational business objectives. Oversee reliability strategy across the organization, ensuring consistent architectural alignment, scalability, and the adoption of industry-standard SRE best practices. Incident Management: Own and oversee organization-wide 24/7 on-call and incident management processes.
Manage incident tooling, establish and maintain operational procedures, ensure compliance with regulations, oversee reporting, and drive continuous improvement of incident response strategy. Data-Driven Management: Define and track measurable objectives (KPIs) for team performance. Leverage data to drive improvements and build management metrics that provide clear visibility into operational health and team productivity.
Observability Engineering: Oversee the design, development, and evolution of the organization-wide observability ecosystem. Lead the strategy for implementing standardized telemetry, including structured logging, distributed tracing, and intelligent sampling. Requirements Professional Background: Several years of experience in SRE, Infrastructure, or DevOps roles managing high-scale, distributed environments. Leadership Experience: Proven track record in a formal management role, leading, mentoring, and developing high-performing SRE or DevOps engineering teams.