June 16, 2026

Senior SRE/DevOps Technical Lead – Observability and Automation

Senior • Hybrid

25,200 - 29,820 PLN

Krakow, Poland

Empower uptime and reliability — lead the next wave of observability and automation excellence!

Krakow-based opportunity with hybrid work model, allowing up to 3 remote days per week.

As a Senior SRE/DevOps Technical Lead, you will be working for our client, a global leader in the banking and financial services industry. You will spearhead efforts to build and operate cutting-edge SRE and observability platform solutions, ensuring system reliability, automation, and performance across a highly regulated environment. This role offers a pivotal leadership position that drives innovation and engineering ownership within a diverse international team.

Your main responsibilities:

  • Lead and develop the delivery capability for SRE/observability platform solutions, fostering excellence in automation, reliability, and monitoring.
  • Build and maintain highly available, low-latency systems aligned with banking industry standards.
  • Drive automation and scripting initiatives utilizing Python, Go, Bash, and other technologies.
  • Manage and optimize CI/CD pipelines and observability/monitoring stacks such as AppDynamics, Grafana, Splunk, and OpenTelemetry.
  • Ensure optimal system performance and reliability across global operations.
  • Collaborate effectively with international teams across different time zones.
  • Provide technical leadership, mentorship, and guidance to team members.

You're ideal for this role if you have:

  • 8+ years of experience in SRE, DevOps, or similar leadership roles.
  • Strong automation and scripting skills (Python, Go, Bash, etc.).
  • Proficiency in CI/CD pipelines and observability/monitoring tools.
  • Proven experience maintaining highly available, low-latency systems in regulated industries such as banking, fintech, or insurance.
  • Ability to work in the Krakow office at least 6 days per month.
  • Fluent English communication skills for global team collaboration.

It is a strong plus if you have: (optional)

  • Certifications related to DevOps, SRE, or cloud platforms.

Language Required for the role:

  • Fluent English

Eligibility for the role:

  • Only candidates with an existing legal right to work in Europe will be considered for this role.

#MAKEYourCareerBETTER
Interested? Apply now and include your CV (preferably in English) along with a statement confirming your consent to the processing and storage of your personal data.

Similar jobs you might like

Technology

emagine Polska

Senior DevOps / SRE (Platform Reliability Engineer) - French fluent

Senior

Remote

Lisbon, Portugal

🏢 Summary: Senior DevOps / SRE role focused on ensuring reliability, scalability, security, and performance of a cloud-native AWS platform. The position centers on infrastructure automation, CI/CD, Kubernetes operations, observability, and implementing SRE best practices to support highly available production systems. You will lead incident management, optimize cloud costs, and drive continuous improvement of platform resilience. 🗂️ Requirements: 5+ years in DevOps/SRE/Cloud/Platform Engineering, Strong Linux administration and troubleshooting, Production experience with Kubernetes, Experience with CI/CD tools, Expertise in Infrastructure as Code, Hands-on experience with AWS, Strong networking fundamentals, Experience with monitoring and logging tools, Scripting skills (Bash or Python) 📃 Skills: AWS, Kubernetes, Docker, Helm, Terraform, Ansible, CloudFormation, Linux, GitLab, Jenkins, GitHub, Azure, Prometheus, Grafana, ELK, Datadog, Splunk, Bash, Python, TCP/IP, DNS 🏢 Description: We are looking for a Senior DevOps / Site Reliability Engineer (SRE) to ensure the reliability, scalability, performance, and security of our platform and cloud infrastructure. You will play a key role in building and operating cloud-native systems, improving observability, automating operations, implementing SRE best practices (SLOs/SLIs), and supporting development teams to deliver highly available services. Key Responsibilities Design, implement, and maintain highly available and scalable infrastructure on AWS. Own and improve the reliability of production systems using SRE principles (SLO, SLI, error budgets). Build and manage CI/CD pipelines to support fast and safe software delivery. Develop and maintain Infrastructure as Code (IaC) using Terraform, Ansible, CloudFormation, etc. Manage and optimize container orchestration platforms (Kubernetes, Docker, Helm). Implement and maintain monitoring, logging, and alerting solutions (Prometheus, Grafana, ELK, Datadog, Splunk). Lead incident response, perform root cause analysis, and write postmortems to drive continuous improvement. Improve system performance, capacity planning, scaling strategies, and disaster recovery processes. Collaborate closely with development teams to improve deployment strategies and system resilience. Implement security best practices (IAM, secret management, vulnerability scanning, patching). Define operational standards, runbooks, documentation, and best practices for platform reliability. Participate in on-call rotation and provide senior-level support for critical production issues. Key Responsibilities (5 Main Missions) The DevOps / SRE lead will be responsible for the stability and evolution of the platform. Your role is structured around five main areas: Mission 1: AWS Infrastructure Management (Build & Run) Mission 2: CI/CD and Deployment Automation Mission 3: Monitoring, Observability, and Alerting: Global Monitoring , Log Management , Application Monitoring , Business Analytics Mission 4: Incident Management, Resilience, and Security Mission 5: FinOps and AWS Cost Optimization Key Requirements 5+ years of experience in DevOps / SRE / Cloud Infrastructure / Platform Engineering. Strong expertise in Linux systems administration and troubleshooting. Proven experience with Kubernetes in production environments. Strong experience with CI/CD tools (GitLab CI, Jenkins, GitHub Actions, Azure DevOps). Solid knowledge of Infrastructure as Code (Terraform highly preferred). Experience with AWS cloud platforms. Strong understanding of networking fundamentals (TCP/IP, DNS, load balancing, reverse proxies). Experience with observability tools: monitoring, metrics, logging, tracing. Strong scripting skills (Bash, Python, or similar). French advanced level. Nice to Have Experience with additional cloud platforms (Azure, GCP). Strong understanding of networking fundamentals.

Technology

ITDS

Senior Tech Lead – .NET & Angular Solutions

Senior

Remote

Wroclaw, Poland

27,300 - 31,500 PLN

🏢 Summary: Senior Tech Lead role focused on designing and delivering enterprise solutions using .NET and Angular in a 100% remote model. The position combines hands-on architecture work with technical leadership, guiding teams and driving integration, CI/CD, and event-driven solutions. The role emphasizes scalable systems, workflow automation, and high engineering standards. 🗂️ Requirements: 8+ years of software development experience, Proven experience in senior or lead technical role, Strong expertise in .NET, Strong expertise in Angular, Strong expertise in SQL, Experience building REST APIs, Experience with GitLab CI/CD, Experience with OpenShift, Experience with event-driven architecture, Knowledge of integration methods, Knowledge of workflow or BPM systems, Communicative Polish, Functional English, Legal right to work in the EU 📃 Skills: .NET, Angular, SQL, REST, GitLab, CICD, OpenShift, EventDriven, BPM, Architecture, Integration 🏢 Description: Unleash innovative leadership — drive cutting-edge technology solutions and inspire teams to excellence! Wroclaw-based opportunity with a 100% remote work model. As a Senior Tech Lead – .NET & Angular Solutions , you will be working for our client in the technology sector, pioneering enterprise solutions that streamline workflows and enable seamless integrations. Your leadership will shape the future of digital transformation, empowering development teams to deliver impactful and scalable systems. Your main responsibilities: Guide and mentor a team of developers in designing and implementing high-quality software solutions. Make key technical decisions related to architecture, integration, and development processes. Collaborate with stakeholders to translate business requirements into robust technical specifications. Oversee the development lifecycle, ensuring adherence to best practices and quality standards. Support continuous integration and deployment pipelines using GitLab CI/CD and OpenShift. Lead efforts in implementing event-driven architectures and workflow/BPM processes. Facilitate code reviews and foster a culture of technical excellence and innovation. You're ideal for this role if you have: 8+ years of experience in software development, with proven leadership in a senior or lead role. Strong expertise in .NET, Angular, SQL, and REST API development. Experience with GitLab CI/CD, OpenShift, and event-driven architecture. Solid understanding of integration methods and workflow management systems. Ability to lead technical discussions and support team members in solving complex challenges. It is a strong plus if you have: (optional) Certification in Agile, Scrum, or related methodologies. Experience with BPM tools or workflow automation solutions. Language Required for the role: Communicative Polish and functional proficiency in English. Eligibility for the role: Only candidates with an existing legal right to work in the European Union will be considered for this role. #MAKEYourCareerBETTER Interested? Apply now and include your CV (preferably in English) along with a statement confirming your consent to the processing and storage of your personal data.

Technology

DCV Technologies

DevOps SRE Engineer | Hybrid from Warsaw

Mid

Hybrid

Warsaw, Poland

🏢 Summary: DevOps SRE Engineer role focused on production support and CI/CD automation for a high-impact cards and online payments project. The position involves maintaining cloud-based infrastructure, managing incidents, and performing deployments in a hybrid work model in Warsaw. Minimum 6-month engagement with potential extension. 🗂️ Requirements: 3–5 years of DevOps/SRE experience, Application production support experience, Experience building and maintaining CI/CD pipelines, Proficiency in Python, Proficiency in Java, Experience with cloud platforms, Working knowledge of containers, Experience in monitoring and incident management, Ability to perform deployments and patching 📃 Skills: Jenkins, GitHubActions, GitLabCI, AzureDevOps, Python, Java, AWS, Azure, GCP, Docker, Kubernetes, Monitoring, IncidentManagement, CICD 🏢 Description: We’re looking for an experienced DevOps SRE Engineer to join a fast-paced, high-impact project in the financial / payments domain. Duration of project: Minimum 6 months and will be extended Client location: Warsaw, Poland Working model : Hybrid working – 3 days from office per week Years of relevant exp needed: 3 to 5 yrs of experience Mandatory Skills: Application PROD Support Details about the project they would be working on: Cards and Online Payment Job description: Building and maintaining CI/CD pipelines (Jenkins, GitHub Actions, GitLab CI, Azure DevOps) Proficiency in Python, Java Experience with AWS / Azure / GCP Working knowledge of containers Monitoring and Incident Management Roles and Responsibilities post joining: Production Support Automation of the Support activities Monitoring and Incident Management Perform deployments and patching Desired Skills: Building and maintaining CI/CD pipelines (Jenkins, GitHub Actions, GitLab CI, Azure DevOps) Proficiency in Python, Java Experience with AWS / Azure / GCP Working knowledge of containers Monitoring and Incident Management Benefits: Collaborative and supportive team culture Exposure to innovative projects  in the biopharmaceutical industry Opportunity to work with global teams and stakeholders 📩 If you’re interested and meet the qualifications, please send your CV to Alina Pchelnikova at alina.pchelnikova@dcvtechnologies.co.uk

Technology

ALTER GPU CENTER

Lead DevOps Engineer

Senior

Remote

Łódź, Poland

🏢 Summary: Technical leadership role combining hands-on DevOps/SRE engineering with team management to build and operate large-scale GPU infrastructure for AI workloads. Focused on infrastructure automation, reliability, observability, and high-performance networking across complex production environments. Responsible for shaping IaC standards, CI/CD, and operational excellence for software-defined, GPU-based platforms. 🗂️ Requirements: 8+ years in DevOps, SRE, or Platform Engineering, 3+ years in technical leadership role, Experience with large-scale infrastructure automation, Proficiency in Infrastructure as Code tools, Experience with GitOps and CI/CD, Hands-on experience with Kubernetes, Experience with GPU technologies, Scripting or programming in Python, Go, or Bash, Experience with bare-metal provisioning, Knowledge of observability and monitoring tools, Understanding of distributed systems reliability, Experience with high-performance networking technologies, Ability to lead technical discussions and mentor engineers, English proficiency at communicative level 📃 Skills: Terraform, Ansible, Pulumi, Crossplane, GitOps, Kubernetes, NVIDIA, MIG, Python, Go, Bash, Prometheus, Grafana, Loki, OpenTelemetry, RDMA, InfiniBand, RoCE, CI/CD 🏢 Description: About the role We are looking for a Lead DevOps Engineer to provide technical leadership for DevOps and Site Reliability Engineering practices supporting large-scale GPU infrastructure used for AI training and inference workloads. This role combines hands-on engineering with team leadership. You will be responsible for shaping automation standards, improving platform reliability, and leading a team working on software-defined infrastructure, high-performance networking, observability, and operational excellence across complex production environments. Responsibilities Lead, mentor, and support a team of DevOps and SRE engineers working across the full lifecycle of GPU infrastructure platforms Design and implement Infrastructure as Code solutions for provisioning and managing bare-metal GPU servers, networking, storage, and cluster orchestration components Build and improve CI/CD pipelines for infrastructure, platform services, and internal tooling Develop and maintain monitoring, logging, alerting, and observability solutions for large-scale GPU environments Define and track SLIs/SLOs , improve incident response processes, and contribute to post-incident reviews and long-term reliability improvements Work closely with Infrastructure, Networking, Facilities, and AI/ML teams to ensure stable and scalable platform operations Automate operational processes such as cluster scaling, firmware and BIOS updates, hardware diagnostics, and capacity planning Support DevSecOps practices, including infrastructure hardening, vulnerability management, and compliance automation Identify operational inefficiencies and reduce repetitive manual work through automation Evaluate and introduce new tools and solutions related to GPU infrastructure, orchestration, and cloud-native operations Requirements 8+ years of experience in DevOps, SRE, Platform Engineering , or a similar area At least 3 years of experience in a technical lead, lead engineer, or team leadership role Strong practical experience with infrastructure automation in large-scale or complex production environments Very good knowledge of Terraform, Ansible, Pulumi, Crossplane , or similar Infrastructure as Code tools Experience with GitOps , configuration management, and CI/CD practices Hands-on experience with Kubernetes Experience working with GPU-related technologies such as NVIDIA GPU Operator, device plugins, MIG, or time-slicing Good scripting or programming skills in Python, Go, or Bash Experience with bare-metal provisioning, infrastructure automation, or data center environments Good knowledge of observability tools such as Prometheus, Grafana, Loki, and OpenTelemetry Good understanding of distributed systems reliability and production incident management Experience with high-performance networking technologies such as RDMA, InfiniBand, or RoCE will be a strong advantage Ability to lead technical discussions, support team development, and communicate effectively with both technical and business stakeholders English proficiency at least at a communicative level is required, as you will be working in an international team Nice to have Experience in AI infrastructure, HPC environments, hyperscale infrastructure, or data center operations Familiarity with orchestration and scheduling tools such as Slurm, Ray, Run:ai, KServe , or Kubernetes-based schedulers Experience integrating telemetry from power, cooling, or environmental systems Experience building internal platforms or self-service tools for engineering or research teams Understanding of security, compliance, and audit requirements in regulated or security-sensitive environments What we offer Benefits package Opportunity to shape the DevOps and SRE foundation for advanced GPU infrastructure supporting AI workloads Real impact on the scalability, reliability, and operational standards of next-generation compute environments Collaboration with experienced engineers across infrastructure, platform, and AI domains A dynamic environment with space for ownership, technical leadership, and professional growth

Technology

Relativity

Senior Engineer - Site Reliability Engineering

Senior

Remote

Krakow, Poland

208,000 - 312,000 PLN/yr

🏢 Summary: Remote Senior Software Engineer – SRE role focused on building and maintaining highly available, scalable, and observable cloud-native systems. The position emphasizes automation, CI/CD improvements, incident management, and implementation of reliability best practices across SaaS platforms. The engineer collaborates cross-functionally to enhance system resilience, performance, and operational excellence. 🗂️ Requirements: 5+ years in Software Engineering, SRE, or Cloud Infrastructure roles, Experience with DevOps tools and practices, Proficiency in Python, Go, Java, C#, or .Net, Experience with at least two: GitHub, Azure DevOps, GitLab, Jenkins, Hands-on experience with observability tools, Strong experience with CI/CD pipelines and automation, Experience with cloud-native distributed systems, Experience in high-availability SaaS environments, Knowledge of SLOs, SLIs, and error budgets, Experience with redundancy and disaster recovery, Participation in on-call rotations 📃 Skills: Python, Go, Java, C#, .Net, GitHub, Azure, GitLab, Jenkins, Prometheus, Grafana, OpenTelemetry, CI/CD, DevOps, SLO, SLI, SaaS, Automation, Cloud, Agile 🏢 Description: Posting Type Remote Job Overview As the Senior Software Engineer – SRE you will focus on implementing and maintaining reliability solutions across the platform. This role emphasizes hands-on engineering work, automation, and operational excellence. The Senior Software Engineer will work closely with other engineers to ensure systems are highly available, observable, and resilient. As a member of the engineering team, the Senior Software Engineer will work closely with Infrastructure, Engineering, and Product teams to develop highly resilient, observable, and automated solutions that enhance system availability and efficiency. The ideal candidate will bring deep technical expertise, strong problem-solving skills, and a passion for reliability engineering. Job Description and Requirements Job Responsibilities Implement, and advocate for best-in-class reliability, observability, and scalability practices across the platform. Develop automated solutions for system reliability, capacity planning, and incident response to minimize manual intervention. Participate in improving Service Level Objectives (SLOs), Service Level Indicators (SLIs), and error budgets to enhance system reliability. Contribute to CI/CD pipeline improvements and DevOps practices. Support root cause analysis (RCA) investigations, drive corrective actions, and advocate for a blameless postmortem culture. Participate in on-call rotations to ensure 24/7 availability of critical systems. Influence and mentor engineering teams on SRE principles, DevOps culture, and best practices. Stay ahead of industry trends, adopting new tools, frameworks, and methodologies to continually improve system reliability. Preferred Qualifications 5+ years of experience in software engineering, site reliability engineering, or cloud infrastructure roles. Experience with DevOps tooling and practices. Proficient in building service-oriented architectures and cloud-native distributed systems. Proficiency in programming languages such as Python, Go, Java, or C# or .Net. In-depth technical understanding and experience with at least two of the following DevOps platforms: GitHub, Azure DevOps, GitLab, or Jenkins. Hands-on experience with observability tools (e.g., Prometheus, Grafana, OpenTelemetry or others). Strong background in CI/CD pipelines, automation, and DevOps practices. Experience working in global, high-availability SaaS environments. Experience implementing redundancy and disaster recovery scenarios. Excellent teamwork and cross-group collaboration skills. Ability to collaborate with both technical and business professionals. Hands-on experience with Agile Project Development Methodologies. Experience delivering complex technical solutions. Excellent problem-solving, analytical, and communication skills. Nice to have: Experience with Chaos Engineering and/or AI Ops . Competencies and Skills Automation-First Mindset – Commitment to reducing toil through scripting and automation. Reliability Engineering – Expertise in SLOs, SLIs, error budgets, and high-availability architectures. Incident Management & Postmortems – Experience in handling production incidents and driving continuous improvement. Observability & Monitoring – Deep understanding of logging, monitoring, and alerting best practices. Practical knowledge of data structures and modern data engines. Collaboration & Communication – Ability to work across teams, influence stakeholders, and advocate for reliability improvements. Mentorship & Coaching – Passion for mentoring engineers and building an SRE culture within the organization. Additional Information This role offers a unique opportunity to shape the future of SRE in a cutting-edge SaaS company, ensuring the reliability and scalability of mission-critical applications for customers worldwide. If you are passionate about solving complex reliability challenges and driving technical excellence, we’d love to hear from you! Relativity is a diverse workplace with different skills and life experiences—and we love and celebrate those differences. We believe that employees are happiest when they're empowered to be their full, authentic selves, regardless how you identify. Benefit Highlights: Comprehensive health, dental, and vision plans Parental leave for primary and secondary caregivers Flexible work arrangements Two, week-long company breaks per year Additional time off Long-term incentive program Training investment program All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, or national origin, disability or protected veteran status, or any other legally protected basis, in accordance with applicable law. Relativity is committed to competitive, fair, and equitable compensation practices. This position is eligible for total compensation which includes a competitive base salary, an annual performance bonus, and long-term incentives. The expected salary range for this role is between following values: 208 000 and 312 000PLN The final offered salary will be based on several factors, including but not limited to the candidate's depth of experience, skill set, qualifications, and internal pay equity. Hiring at the top end of the range would not be typical, to allow for future meaningful salary growth in this position. Required Skills: Automation, Data Analysis, Database Management, Network Architecture, Performance Optimizations, Problem Solving, Project Management, Software Development, System Designs, Technical Leadership

Technology

Connectis

DevOps/SRE

Senior

Remote

Warsaw, Poland

143 - 209 PLN

🏢 Summary: DevOps / SRE role focused on scaling and standardizing observability across a large enterprise environment with over 160 applications. The position involves integrating systems with a central observability model, defining SRE standards, and supporting teams in monitoring, logging, and metrics across cloud and enterprise platforms. The role is horizontal and advisory, emphasizing enablement, architecture guidance, and end-to-end visibility. 🗂️ Requirements: Minimum 5 years experience as SRE or DevOps Engineer in enterprise environments, Strong knowledge of Microsoft Azure core services and cloud architecture, Hands-on experience with Prometheus and Grafana, Practical knowledge of distributed tracing, centralized logging, metrics and visualization, Basic experience with Kubernetes deployment and container management, Practical experience with OpenTelemetry instrumentation, Experience defining and monitoring SLO and SLI, Fluent English 📃 Skills: Azure, Prometheus, Grafana, Kubernetes, OpenTelemetry, SLO, SLI, Loki, Dynatrace, Datadog, NewRelic, AppDynamics, ServiceNow, Logscale, LQL, GCP, Python, PowerShell, SAP, Oracle, Salesforce 🏢 Description: Do zespołu Observability poszukujemy doświadczonej osoby na stanowisko DevOps / SRE , który odegra kluczową rolę w skalowaniu i standaryzacji rozwiązań observability w dużej organizacji o złożonym krajobrazie technologicznym. Projekt koncentruje się na ujednoliceniu monitoringu, logowania i metryk dla ponad 160 aplikacji działających w wielu obszarach technologicznych (kilka niezależnych domen / „towerów”), obejmujących zarówno środowiska chmurowe, jak i rozbudowane systemy klasy enterprise. Rola ma charakter horyzontalny i enablementowy, jej celem jest wspieranie zespołów produktowych i utrzymaniowych w integracji systemów z centralnym modelem observability oraz współtworzenie i promowanie wspólnych standardów SRE / DevOps w skali całej organizacji. 💡 TWOJA ROLA: Analiza lokalnych rozwiązań monitoringowych i mapowanie ich do wspólnego modelu SRE. Definiowanie i promowanie standardów observability (naming, schematy danych, konwencje). Udział w PoC / pilotach (zbieranie metryk, konfiguracja, testy wysyłki danych do Azure). Wsparcie zespołów w interpretacji danych observability (RCA, SLO/SLA, diagnostyka). Współpraca z zespołami produktowymi, zespołami utrzymaniowymi oraz vendorami. Budowanie widoczności end-to-end dla krytycznych procesów biznesowych. Integracja systemów dziedzinowych z centralnym modelem observability. Doradztwo architektoniczne. 🔍 CZEGO OCZEKUJEMY OD CIEBIE? Minimum 5-letnie doświadczenie w roli SRE / DevOps Engineer w środowiskach enterprise. Znajomość platformy Microsoft Azure w zakresie core services oraz podstaw architektury chmurowej. Bardzo dobra znajomość narzędzi monitoringu i observability, w szczególności Prometheus i Grafana. Praktyczna znajomość observability : distributed tracing, centralne logowanie, metryki i wizualizacja. Podstawowa znajomość Kubernetes w zakresie deploymentu oraz zarządzania kontenerami. Praktyczna znajomość OpenTelemetry (OTel) w zakresie instrumentacji aplikacji. Doświadczenie w definiowaniu, wdrażaniu oraz monitorowaniu SLO / SLI. Biegła znajomość języka angielskiego. Mile widziane: Doświadczenie w pracy z rozbudowanymi systemami vendorowymi (takimi jak SAP , Oracle , Salesforce lub innymi platformami klasy enterprise). Doświadczenie z narzędziami klasy enterprise observability: Dynatrace, Datadog, New Relic, AppDynamics. Podstawowa znajomość ServiceNow w zakresie zarządzania incydentami i zmianami. Doświadczenie z Logscale zarządzanie logami, zapytania i analiza (LQL). Praktyczne doświadczenie z Loki - centralne logowanie. Podstawowa znajomość GCP, Python i Powershell . ✨ OFERUJEMY: 🤖 Nowoczesny proces rekrutacji z AI Rekruterem (AIR) - podczas aplikacji możesz odbyć rozmowę z wirtualnym rekruterem 24/7, bez czekania na telefon, z natychmiastowym feedbackiem i możliwością powtórzenia rozmowy (liczy się ostatnia wersja). Finalną decyzję zawsze podejmuje Rekruter Connectis. Uczestnictwo w spotkaniach integracyjnych oraz meetupach technologicznych, umożliwiających dzielenie się wiedzą i doświadczeniem. Wsparcie dedykowanej osoby kontaktowej z Connectis, dostępnej w celu pomocy w sprawach związanych z projektem. Stabilne i długoterminowe zatrudnienie w firmie o ugruntowanej pozycji na rynku. 100% zdalnie. Pełna praca zdalna, bez konieczności dojazdów. Możliwość rozwoju w nowoczesnym, dynamicznym środowisku IT. 5000 PLN za polecenie znajomych do naszych projektów. Szybki, zdalny proces rekrutacyjny. Dziękujemy za wszystkie zgłoszenia. Pragniemy poinformować, że skontaktujemy się z wybranymi osobami. 12821/NS

Technology

COMARCH

Support & Platform Maintenance Manager

Senior

Hybrid

Krakow, Poland

🏢 Summary: Leadership role responsible for transforming technical support (L2/L3) and platform maintenance into a modern SRE-driven, observability-focused model. The position focuses on stabilizing and scaling distributed systems, implementing structured incident and problem management, and modernizing legacy processes. The role combines operational excellence, automation, and platform monitoring to ensure high service quality and business continuity. 🗂️ Requirements: Higher education degree, 4-6 years experience in technical support or IT service management roles, Minimum 3 years experience managing technical teams, Experience in IT process modernization and SRE implementation, Practical knowledge of ITIL framework, Experience with ticketing systems, Experience with observability and monitoring tools, Knowledge of SQL Server, Knowledge of Windows Server, Experience with Incident and Problem Management, Experience defining SLA and KPI metrics, Experience in process automation with DevOps collaboration, Very good English proficiency 📃 Skills: SRE, ITIL, ITSM, JIRA, Grafana, Prometheus, ELK, SQL, SQLServer, WindowsServer, DevOps, Observability, IncidentManagement, ProblemManagement, SLA, KPI 🏢 Description: Poszukujemy doświadczonego lidera, który przejmie odpowiedzialność za obszar wsparcia technicznego i przeprowadzi transformację modelu pracy zespołów L2 i L3 i observability platformy. W tej roli będziesz mieć możliwość wykorzystania swojej wiedzy w celu tworzenia i rozwijania naszego produktu, ewoluując rozwiązania legacy w stronę nowoczesnych standardów oraz budując płynną synergię między inżynierią a biznesem. Jako Support & Platform Maintenance Lead staniesz na czele doświadczonego zespołu, wdrażając kulturę SRE oraz zaawansowane narzędzia observability niezbędne do monitorowania masowej wymiany danych. Twoim zadaniem będzie rozwój obecnych procesów operacyjnych w stronę pełnej stabilności i przewidywalności systemów rozproszonych obsługujących krytyczne procesy biznesowe. Będziesz mieć realny wpływ na to, jak szybko skalujemy nasz produkt przy zachowaniu najwyższej jakości usług oraz dbałości o ciągłość operacyjną naszych klientów. Profil stanowiska Wykształcenie wyższe Minimum 4-6 lat doświadczenia zawodowego na stanowiskach związanych ze wsparciem technicznym, utrzymaniem systemów lub zarządzaniem usługami IT (np. Support Manager, Service Delivery Manager, L2/L3 Lead, Platform Operations Manager) Minimum 3 lata doświadczenia w zarządzaniu zespołem technicznym, w roli lidera lub managera zespołu Doświadczenie w modernizacji procesów IT i wdrażaniu nowoczesnych standardów monitorowania (SRE, Observability) Praktyczna znajomość ITIL, systemów ticketowych (np. JIRA Service Management), narzędzi observability (np. Grafana, Prometheus, ELK Stack), SQL Server i Windows Server Bardzo dobra znajomość języka angielskiego Wysokie zdolności analityczne oraz proaktywność w identyfikowaniu obszarów do optymalizacji Umiejętność swobodnej komunikacji z kadrą zarządzającą oraz zespołami inżynieryjnymi Ponadprzeciętne zdolności organizacyjne i strukturalne podejście do problemów, umożliwiające sprawne priorytetyzowanie zadań w konkretne plany działania Mile widziane certyfikaty: ITIL/ SRE/ DevOps/ ITSM Twoje zadania Wdrożenie Incident i Problem Management, budowa bazy wiedzy, definiowanie metryk operacyjnych oraz koordynacja codziennych spotkań klasyfikacyjnych oraz prognozowanie dostępności zespołu Definiowanie metryk biznesowych platformy, projektowanie dashboardów oraz wdrożenie monitoringu procesów biznesowych we współpracy z zespołem AI R&D Automatyzacja procesów utrzymaniowych we współpracy z zespołem DevOps Wdrożenie ustrukturyzowanego cyklu utrzymania platformy, w tym zdefiniowania stałych okien na aktualizacje i poprawki systemowe Zaprojektowanie modelu on-call dla L3 Nadzór nad jakością procesów ticketowych i realizacją założeń SLA/KPI Raportowanie statusu usług do zarządu oraz merytoryczne wsparcie zespołu operacyjnego Koordynacja pracy podległego zespołu wsparcia technicznego i utrzymania platformy w celu zapewnienia ciągłości działania systemów Operacyjne wsparcie projektów skalowania infrastruktury w odpowiedzi na potrzeby biznesowe Priorytetyzacja zadań i przekładanie problemów na konkretne plany działania Współpraca z zespołami wdrożeniowymi, QA, compliance oraz consultingu i sprzedaży Dla Ciebie Realny wpływ na tempo skalowania produktu i budowanie procesów od podstaw - to rola, w której definiujesz model od triażu przez eskalację po observability Rozwój kompetencji liderskich poprzez zarządzanie pracą zespołów oraz wpływ na rozwój merytoryczny i budowanie kultury wysokiej jakości usług Możliwość rozwoju w strukturach globalnego software house’u, dostęp do unikalnej wiedzy i projektów o dużej skali Współpraca z różnymi zespołami kompetencyjnymi, w tym AI R&D oraz Platform Core Dostęp do zaawansowanych narzędzi AI oraz Google Workspace Prywatna opieka medyczna w Allianz - dostęp do specjalistów i badań diagnostycznych dla Ciebie i Twoich najbliższych System kafeteryjny - możesz wybrać pełne dofinansowanie do karty Multisport czy przeznaczyć środki na kulturę, wypoczynek lub sport i zakupy Możliwość pracy w modelu hybrydowym po okresie wdrożenia (2 dni pracy zdalnej, 3 dni pracy z biura) Współpraca z różnymi zespołami kompetencyjnymi, która pozwoli Ci na poznanie nowych technologii i poszerzenie wiedzy w obszarze IT Profesjonalny rozwój zawodowy, realny wpływ na podejmowane decyzje biznesowe Udogodnienia dla rowerzystów/ rowerzystek (stojaki, szatnie, rowerownie, stacja naprawcza), a dla tych, co do pracy docierają samochodem – naziemny i podziemny parking

Technology

COMARCH

Support & Platform Maintenance Lead

Senior

Hybrid

Krakow, Poland

🏢 Summary: Leadership role responsible for transforming technical support (L2/L3) and platform maintenance into a modern, SRE-driven operating model with advanced observability. The position focuses on stabilizing and scaling distributed systems supporting critical business processes while modernizing legacy solutions. It includes ownership of incident management, monitoring strategy, automation, and operational excellence. 🗂️ Requirements: Higher education degree, 4–6 years experience in technical support, system maintenance or IT service management roles, Minimum 3 years experience managing technical teams, Experience implementing SRE and observability standards, Practical knowledge of ITIL framework, Experience with ticketing systems, Hands-on experience with monitoring and observability tools, Knowledge of SQL Server, Knowledge of Windows Server, Very good English skills 📃 Skills: SRE, Observability, ITIL, JIRA, Grafana, Prometheus, ELK, SQL, SQLServer, WindowsServer, ITSM, DevOps, AI 🏢 Description: Poszukujemy doświadczonego lidera, który przejmie odpowiedzialność za obszar wsparcia technicznego i przeprowadzi transformację modelu pracy zespołów L2 i L3 i observability platformy. W tej roli będziesz mieć możliwość wykorzystania swojej wiedzy w celu tworzenia i rozwijania naszego produktu, ewoluując rozwiązania legacy w stronę nowoczesnych standardów oraz budując płynną synergię między inżynierią a biznesem. Jako Support & Platform Maintenance Lead staniesz na czele doświadczonego zespołu, wdrażając kulturę SRE oraz zaawansowane narzędzia observability niezbędne do monitorowania masowej wymiany danych. Twoim zadaniem będzie rozwój obecnych procesów operacyjnych w stronę pełnej stabilności i przewidywalności systemów rozproszonych obsługujących krytyczne procesy biznesowe. Będziesz mieć realny wpływ na to, jak szybko skalujemy nasz produkt przy zachowaniu najwyższej jakości usług oraz dbałości o ciągłość operacyjną naszych klientów. Profil stanowiska Wykształcenie wyższe Minimum 4-6 lat doświadczenia zawodowego na stanowiskach związanych ze wsparciem technicznym, utrzymaniem systemów lub zarządzaniem usługami IT (np. Support Manager, Service Delivery Manager, L2/L3 Lead, Platform Operations Manager) Minimum 3 lata doświadczenia w zarządzaniu zespołem technicznym, w roli lidera lub managera zespołu Doświadczenie w modernizacji procesów IT i wdrażaniu nowoczesnych standardów monitorowania (SRE, Observability) Praktyczna znajomość ITIL, systemów ticketowych (np. JIRA Service Management), narzędzi observability (np. Grafana, Prometheus, ELK Stack), SQL Server i Windows Server Bardzo dobra znajomość języka angielskiego Wysokie zdolności analityczne oraz proaktywność w identyfikowaniu obszarów do optymalizacji Umiejętność swobodnej komunikacji z kadrą zarządzającą oraz zespołami inżynieryjnymi Ponadprzeciętne zdolności organizacyjne i strukturalne podejście do problemów, umożliwiające sprawne priorytetyzowanie zadań w konkretne plany działania Mile widziane certyfikaty: ITIL/ SRE/ DevOps/ ITSM Twoje zadania Wdrożenie Incident i Problem Management, budowa bazy wiedzy, definiowanie metryk operacyjnych oraz koordynacja codziennych spotkań klasyfikacyjnych oraz prognozowanie dostępności zespołu Definiowanie metryk biznesowych platformy, projektowanie dashboardów oraz wdrożenie monitoringu procesów biznesowych we współpracy z zespołem AI R&D Automatyzacja procesów utrzymaniowych we współpracy z zespołem DevOps Wdrożenie ustrukturyzowanego cyklu utrzymania platformy, w tym zdefiniowania stałych okien na aktualizacje i poprawki systemowe Zaprojektowanie modelu on-call dla L3 Nadzór nad jakością procesów ticketowych i realizacją założeń SLA/KPI Raportowanie statusu usług do zarządu oraz merytoryczne wsparcie zespołu operacyjnego Koordynacja pracy podległego zespołu wsparcia technicznego i utrzymania platformy w celu zapewnienia ciągłości działania systemów Operacyjne wsparcie projektów skalowania infrastruktury w odpowiedzi na potrzeby biznesowe Priorytetyzacja zadań i przekładanie problemów na konkretne plany działania Współpraca z zespołami wdrożeniowymi, QA, compliance oraz consultingu i sprzedaży Dla Ciebie Realny wpływ na tempo skalowania produktu i budowanie procesów od podstaw - to rola, w której definiujesz model od triażu przez eskalację po observability Rozwój kompetencji liderskich poprzez zarządzanie pracą zespołów oraz wpływ na rozwój merytoryczny i budowanie kultury wysokiej jakości usług Możliwość rozwoju w strukturach globalnego software house’u, dostęp do unikalnej wiedzy i projektów o dużej skali Współpraca z różnymi zespołami kompetencyjnymi, w tym AI R&D oraz Platform Core Dostęp do zaawansowanych narzędzi AI oraz Google Workspace Prywatna opieka medyczna w Allianz - dostęp do specjalistów i badań diagnostycznych dla Ciebie i Twoich najbliższych System kafeteryjny - możesz wybrać pełne dofinansowanie do karty Multisport czy przeznaczyć środki na kulturę, wypoczynek lub sport i zakupy Możliwość pracy w modelu hybrydowym po okresie wdrożenia (2 dni pracy zdalnej, 3 dni pracy z biura) Współpraca z różnymi zespołami kompetencyjnymi, która pozwoli Ci na poznanie nowych technologii i poszerzenie wiedzy w obszarze IT Profesjonalny rozwój zawodowy, realny wpływ na podejmowane decyzje biznesowe Udogodnienia dla rowerzystów/ rowerzystek (stojaki, szatnie, rowerownie, stacja naprawcza), a dla tych, co do pracy docierają samochodem – naziemny i podziemny parking

Technology

Link Group

DevOps / Site Reliability Engineer

Mid

Hybrid

Kraków, Poland

20,000 - 25,000 PLN

🏢 Summary: DevOps / Site Reliability Engineer role focused on building and maintaining scalable cloud infrastructure while improving platform reliability and automation. The position centers on Kubernetes-based environments, CI/CD pipeline development, and enhancing monitoring and observability. The engineer will support development teams through infrastructure as code and internal developer platform initiatives. 🗂️ Requirements: Experience with cloud platforms (Azure preferred), Strong experience with Kubernetes, Strong knowledge of Infrastructure as Code (Terraform), Hands-on experience with CI/CD tools, Experience with monitoring and observability tools, Understanding of scalability, reliability, and security best practices 📃 Skills: Azure, Kubernetes, Terraform, GitHubActions, ArgoCD, CI/CD, Datadog, Prometheus, Grafana, MongoDB, Rancher, Jenkins, PowerBI, Jira, Confluence 🏢 Description: DevOps / Site Reliability Engineer We’re looking for a DevOps / SRE to help build and maintain scalable cloud infrastructure and improve reliability across our platform. You’ll focus on automation, CI/CD, and supporting development teams with efficient tooling and processes. Key responsibilities Develop and manage cloud infrastructure (Azure preferred) Work with Kubernetes and containerized environments Build and maintain CI/CD pipelines (GitHub Actions, ArgoCD) Automate deployments and operational processes Contribute to Internal Developer Platform (IDP) development Improve monitoring and observability (e.g., Datadog, Prometheus, Grafana) Requirements Experience with cloud platforms and Kubernetes Strong knowledge of Infrastructure as Code (e.g., Terraform) Hands-on experience with CI/CD tools Understanding of scalability, reliability, and security best practices Experience with monitoring/observability tools Nice to have Experience with MongoDB Atlas, Rancher, Jenkins, Power BI Familiarity with Jira, Confluence

Technology

Yard Corporate

Site Reliability Engineer (SRE)

Senior

Hybrid

Warsaw, Poland

40,000 - 55,000 PLN

🏢 Summary: Senior Site Reliability Engineer role focused on building and standardizing SRE practices across a hybrid AWS and on-prem infrastructure. The position centers on ensuring scalability, resilience, and high availability of high-frequency, data-intensive platforms through observability, automation, and Kubernetes optimization. You will define SLOs, enhance monitoring architecture, and drive reliability culture across engineering teams. 🗂️ Requirements: 5+ years experience in SRE, DevOps, or Infrastructure Engineering supporting distributed production systems, Bachelor’s degree in Computer Science, Computer Engineering, or related field (or equivalent experience), Deep expertise in Grafana, Prometheus, Loki, and Tempo (OpenTelemetry), Strong production experience with Docker and Kubernetes, Experience managing hybrid infrastructure (AWS and on-premises), Proficiency in at least one language: Python, Go, or Bash, Hands-on experience with CI/CD pipelines and Infrastructure-as-Code, Experience defining and managing SLOs and SLAs, Willingness to participate in on-call rotation 📃 Skills: AWS, Kubernetes, Docker, Prometheus, Grafana, Loki, Tempo, OpenTelemetry, Python, Go, Bash, CI/CD, IaC, Git, Hypervisors 🏢 Description: About the Client Our client is a premier, global investment management firm operating at the intersection of finance and technology. Known for their sophisticated, data-intensive systems, they build and maintain high-performance platforms that process massive volumes of market and operational data. To support their expanding footprint, they are looking for a senior-level Site Reliability Engineer (SRE) who will take ownership of shaping, standardizing, and scaling their SRE frameworks and reliability culture from the ground up. The Role In this role, you will serve as a foundational force for SRE practices, partnering directly with Cloud, Infrastructure, and Software Engineering squads. You will work across a hybrid infrastructure (combining advanced AWS cloud environments and physical on-premises servers) to guarantee the scalability, resilience, and maximum uptime of critical, high-frequency transactional platforms. Core Responsibilities SRE Evangelism: Design, implement, and champion core reliability principles, helping technology teams adopt sustainable scaling practices. Observability Architecture: Implement, scale, and maintain end-to-end monitoring, telemetry, and distributed tracing systems utilizing Prometheus, Grafana, Loki, and Tempo (OpenTelemetry framework). Kubernetes Optimization: Establish best-practice configurations for containerized workloads, ensuring applications running on Kubernetes are highly resilient, cost-effective, and performant. Incident Management & Culture: Participate in a balanced, shared on-call rotation (averaging one week per month). Automation & Engineering: Build custom tooling and CI/CD pipelines to automate routine tasks, system health checks, and rapid disaster recovery workflows. SLO/SLA Definition: Partner with product and engineering teams to define, monitor, and enforce Service Level Objectives (SLOs) and Error Budgets. What We Look For Experience: 5+ years of hands-on experience in a dedicated SRE, DevOps, or Infrastructure Engineering role supporting complex, distributed production systems. Education: A Bachelor’s degree in Computer Science, Computer Engineering, or a related technical discipline (or equivalent practical experience). Observability Expertise: Deep, subject-matter knowledge of modern monitoring stacks, specifically Grafana, Prometheus, Loki, and Tempo (OTel). Orchestration & Containers: Strong, production-grade expertise in containerization (Docker) and orchestration (Kubernetes). Hybrid Infrastructure: Experience navigating hybrid models—managing both cloud services (AWS preferred) and physical on-premise hardware resources. Scripting/Coding: Proficiency in writing clean, maintainable code in at least one scripting or programming language (e.g., Python, Bash, or Go) to build reliable automation. Methodologies: Solid grounding in CI/CD concepts, infrastructure-as-code (IaC), and agile development processes. Soft Skills: Excellent verbal and written communication skills, with a proven ability to convey complex infrastructure and reliability concepts to both technical and non-technical stakeholders. What We Offer Stable Employment: Full-time employment contract ( Umowa o Pracę - UoP ). Tax Optimization: Eligibility for creative tax-deductible costs ( KUP - Koszty Uzyskania Przychodu). Financial Reward: Highly competitive base salary accompanied by a generous annual performance bonus . Comprehensive Health: Premium private medical care package that fully includes dental coverage (stomatologia) . Wellness & Lifestyle: MultiSport card to keep you active and healthy. Daily Perks: Pre-funded lunch card for your daily meals. Tech Stack at a Glance Cloud & Virtualization: AWS, Kubernetes, Docker, On-Premises Hypervisors Observability: Prometheus, Grafana, Loki, Tempo, OpenTelemetry (OTel) Languages: Python, Go, Bash CI/CD & Automation: Git-based pipelines, Configuration Management, IaC