Unlock Full Resume Report
ATS Pass
Missing keywords
Tailored AI suggestions
Job match analysis
Interview-focused insights
New offer - be the first one to apply!
September 3, 2026
Engineering Team Leader (Site Reliability Engineering)
Senior • Remote
27,200 - 34,600 PLN/yr
Warsaw, MZ, Poland
Apply now
We are looking for an Engineering Team Leader to drive the development and growth of the Site Reliability Engineering team. In this role, you will shape the technical and operational direction of SRE practices, lead a resilience strategy, and define and deliver solutions that ensure the reliability and scalability of systems for millions of clients across a growing organization.
Responsibilities
- Team Leadership: Shape and grow a high-performing Site Reliability Engineering team, fostering technical excellence, ownership, and continuous improvement.
- Reliability Strategy: Define and drive the SRE platform strategy in close collaboration with infrastructure and development teams, ensuring alignment with business objectives. Oversee reliability strategy across the organization, ensuring consistent architectural alignment, scalability, and adoption of industry-standard SRE best practices.
- Incident Management: Own and oversee organization-wide 24/7 on-call and incident-management processes. Manage incident tooling, establish and maintain operational procedures, ensure regulatory compliance, oversee reporting, and continuously improve incident-response strategy.
- Data-Driven Management: Define and track measurable objectives (KPIs) for team performance. Use data to drive improvements and build management metrics that provide visibility into operational health and team productivity.
- Observability Engineering: Oversee the design, development, and evolution of the organization-wide observability ecosystem. Lead the strategy for standardized telemetry, including structured logging, distributed tracing, and intelligent sampling.
Requirements
- Several years of experience in SRE, Infrastructure, or DevOps roles managing high-scale, distributed environments.
- Proven track record in a formal management role leading, mentoring, and developing high-performing SRE or DevOps engineering teams.
- Extensive experience building and maintaining scalable, reliable, and observable infrastructure systems across Azure, Kubernetes, and on-premises environments.
- Demonstrated ability to deliver end-to-end reliability strategies, drive architectural improvements, and manage large-scale technical projects from design through production.
- Proven ability to drive cultural change, partner strategically with product engineering teams, and work effectively with distributed or remote teams.
Leadership & Management
- Execution: Break down complex projects into actionable tasks and deliver incremental value.
- Growth: Mentor and support the professional growth of team members.
- Collaboration: Facilitate workshops, build technical community, and resolve conflicts effectively.
- Operational Excellence: Drive operational excellence and reliability culture; lead incident management and champion post-mortems.
- Strategy: Proactively manage technical debt and align team output with organizational goals.
Technical skills we expect
- Programming & Scripting: Strong Python skills for scalable automation, internal tools, and scripts.
- Cloud & Orchestration: Expertise managing Kubernetes, configuration management with Ansible, and designing resilient infrastructure on Azure and on-premises.
- Observability Engineering: Deep proficiency in building standardized telemetry systems, including Prometheus, Grafana, OTEL, ELK, Tempo, Thanos, and similar tools.
- AI & Automation: Use AI/ML for AIOps, anomaly detection, log analysis, and reliability-workflow optimization.
Nice to have
- Experience with commercial APM platforms, such as Datadog, Splunk, and New Relic, and chaos-engineering tooling.
- Proficiency with cloud cost management and FinOps principles.
What we offer
- Real influence on company and product development.
- Work in an experienced team that shares knowledge.
- Clear development vision through regular feedback and career paths.
- Regular team-building meetings.
Benefits
- Training budget for courses and conferences.
- An extra day off on your birthday.
- An extra day off for parents.
- Equipment tailored to your needs.
- Private medical care and group insurance.
- Access to an e-learning platform for English learning and a benefits platform.
- Access to a wellbeing platform, workshops, and private therapy sessions.
- Remote work, work from the Warsaw office, or a coworking space in your city.
Similar jobs you might like
Site Reliability Engineer
XTB S.A.
Senior
Technology
Warsaw, Poland · Remote
18K zł - 23K zł/yr
1 days ago
Site Reliability Engineer
Transition Technologies MS
Senior
Technology
Warsaw, MZ, Poland · Remote
N/A
46m ago

Site Reliability Engineer
EPAM Systems
Mid
Technology
Krakow, MA, Poland · Remote
N/A
1 days ago

Site Reliability Engineer
EPAM Systems
Mid
Technology
Gdansk, PM, Poland · Remote
N/A
1 days ago

Site Reliability Engineer
EPAM Systems
Mid
Technology
Katowice, SL, Poland · Remote
N/A
1 days ago

Site Reliability Engineer
EPAM Systems
Mid
Technology
Warsaw, MZ, Poland · Remote
N/A
1 days ago

Site Reliability Engineer
EPAM Systems
Mid
Technology
Poznan, WP, Poland · Remote
N/A
1 days ago

Site Reliability Engineer
EPAM Systems
Mid
Technology
Wroclaw, DS, Poland · Remote
N/A
1 days ago
Senior AI Platform Engineer
XTB S.A.
Senior
Technology
Warsaw, MZ, Poland · Remote
20K zł - 26K zł/yr
1 days ago
Kubernetes Platform SRE Team Lead – Container Orchestration
ITDS
Senior
Technology
Krakow, MA, Poland · On-site
30K zł - 36K zł/yr
1 days ago