Unlock Full Resume Report

New offer - be the first one to apply!

September 14, 2026

Site Reliability Engineer

Senior • Remote

312,000 - 395,200 PLN/yr

Warsaw, Poland

Quick Facts

  • Role: Site Reliability Engineer (SRE)

  • Engagement: Remote (B2B), full-time

Description

The SRE ensures reliability, availability, performance, and operational continuity for an enterprise ecosystem of multiple highly integrated systems. The role covers technical stability of individual applications and end-to-end reliability of critical business processes spanning multiple platforms and services, working closely with DevOps teams, application owners, architects, support teams, and SPOCs.

Responsibilities

  • Ensure high availability, reliability, performance, and operational continuity of business-critical systems and services

  • Monitor and analyze end-to-end business processes across multiple applications, APIs, middleware components, messaging platforms, and external systems

  • Identify system dependencies and assess their impact on business service availability

  • Establish and maintain monitoring, observability, alerting, and dashboards for both technical and business-process metrics

  • Define and track SLIs, SLOs, availability targets, and other reliability metrics

  • Coordinate major incidents involving multiple systems and technical teams

  • Lead root cause analysis for production incidents, integration failures, performance degradation, and service interruptions

  • Coordinate troubleshooting and communication with SPOCs, system owners, vendors, infrastructure teams, and external providers

  • Manage and prioritize the work of the DevOps team responsible for deployment, monitoring, automation, infrastructure, and operational support

  • Drive automation of operational activities, deployments, health checks, recovery procedures, and system maintenance

  • Maintain operational runbooks, troubleshooting guides, escalation paths, and recovery procedures

  • Ensure appropriate backup, disaster recovery, failover, and business continuity mechanisms are implemented and validated

  • Support release planning, production readiness, risk assessment, dependency analysis, and rollback strategies

  • Proactively identify reliability risks, performance bottlenecks, single points of failure, and architectural weaknesses

  • Improve resilience, scalability, fault tolerance, retry mechanisms, and graceful degradation with development and architecture teams

  • Lead post-incident reviews and ensure corrective and preventive actions are implemented

Requirements

  • Strong experience in SRE, DevOps, Production Engineering, Application Operations, or similar

  • Experience working with complex, highly integrated enterprise architectures

  • Strong understanding of end-to-end business process monitoring and dependency management

  • Experience with incident management, root cause analysis, problem management, and service restoration

  • Practical knowledge of monitoring, logging, alerting, and observability platforms

  • Good understanding of SLI, SLO, SLA, availability, latency, throughput, and reliability concepts

  • Experience with CI/CD, release management, infrastructure automation, and deployment processes

  • Understanding of high availability, disaster recovery, failover, and resilience patterns

  • Ability to coordinate technical activities across multiple teams and system owners

  • Experience managing or coordinating a DevOps or operations-focused engineering team

  • Strong analytical, troubleshooting, and communication skills

Nice to have

  • Cloud platforms (AWS, Azure, GCP) in a production operations context

  • Containerization and orchestration (Docker, Kubernetes)

  • Scripting/automation (Python, Bash, or similar)

  • APM/observability tools (Datadog, Dynatrace, New Relic, Grafana, Prometheus)

  • Messaging/integration middleware (Kafka, MQ, ESB platforms)

  • ITIL or similar IT service management framework knowledge

Benefits

  • B2B contract (rate up to 190 PLN net/h + VAT)

  • Fully remote service delivery

  • Broad range of projects (internal, international)

  • Budget for skills development and certifications

  • Regular collaboration reviews and discussion of project scope

  • Additional benefits available as part of the collaboration

  • Networking and team-building events for project teams

Similar jobs you might like