June 16, 2026

Operations Engineer (Kafka Platform Support)

Mid • Remote

130 - 140 PLN

Warsaw, Poland

About Company: 

Team Connect is Poland’s leading nearshore and offshore IT provider. Since 2008 we successfully create and develop software for our clients. 
We specialize in Agile and DevOps-based software development. From the analysis stage through implementation. We develop backend, frontend, and mobile applications. 

For one of our clients, we are looking for an Operations Engineer (Kafka Platform Support)

This role is ideal for an IT Operations Engineer focused on maintaining and supporting a Kafka-based platform in production. The position emphasizes operational excellence, incident management, and user support rather than platform development, ensuring reliable and efficient system performance.

Key Responsibilities:

  • IT Operations & Platform Support

    • Operate, monitor, and maintain the Kafka-based messaging platform in a production environment

    • Ensure platform availability, stability, and performance in line with operational SLAs

    • Monitor system health using logs, metrics, and alerting tools

    • Perform routine operational checks and maintenance activities

  • Incident Management & Troubleshooting

    • Handle incidents and service requests via ticketing systems and internal support channels

    • Troubleshoot issues across Kafka components (brokers, producers, consumers, integrations)

    • Analyze logs, metrics, and system behavior to identify root causes

    • Escalate complex issues to engineering teams where necessary

  • Runbook Execution & Operational Processes

    • Execute operational procedures based on runbooks and standard operating procedures (SOPs)

    • Perform configuration changes (topics, access controls, settings) following established processes

    • Maintain and continuously improve operational documentation and runbooks

  • User Support & Communication

    • Act as a primary support contact for internal users of the Kafka platform

    • Provide technical support via collaboration tools (e.g., Slack, Teams)

    • Assist users with troubleshooting and best practices

    • Translate user-reported issues into actionable insights for technical teams

  • Collaboration & Continuous Improvement

    • Work closely with engineering and platform teams to resolve incidents

    • Identify recurring operational issues and suggest improvements or automation

    • Participate in incident reviews and post-mortems

    • Provide feedback to improve platform usability and support processes

Required Skills & Qualifications:

  • Technical Skills

    • Hands-on experience with Apache Kafka or similar event streaming platforms

    • Understanding of distributed systems (partitioning, replication, scaling)

    • Strong troubleshooting skills in production IT environments

    • Experience with monitoring, logging, and alerting tools

    • Knowledge of Git and version control practices

    • Familiarity with GitLab CI/CD and working with existing pipelines

  • IT Operations Experience

    • Experience in IT operations, production support, or platform support roles

    • Familiarity with incident management processes and tools

    • Experience working with runbooks, SOPs, and structured support models

  • Communication & Collaboration

    • Strong communication skills and ability to explain technical issues clearly

    • Experience working with internal customers and cross-functional teams

    • Customer-focused mindset with a proactive approach to support

    • Fluency in English (written and spoken)

Nice to Have:

  • Experience with AWS or other cloud platforms

  • Familiarity with Kubernetes and containerized environments

  • Experience with monitoring tools such as Grafana and Prometheus

Benefits:

  • Long-term cooperation

  • Multisport, private healthcare, life insurance

  • Training budget

  • English lessons

  • Support from a dedicated partnership consultant

Similar jobs you might like

Technology

Team Connect

Windows Engineer

Mid

Hybrid

Warsaw, MZ, Poland

160 - 180 PLN

🏢 Summary: The offer is for a Windows Engineer responsible for administering and maintaining Windows Server environments in virtualized infrastructures, supporting DevOps operations, and ensuring system security, stability, and performance. The role includes monitoring, troubleshooting, automation, and participation in architectural and change management processes. The position involves 2nd line support and collaboration within complex, distributed IT environments. 🗂️ Requirements: Experience administering Windows Server 2016 or later in virtualized environments, Experience managing complex and distributed IT systems, Knowledge of network protocols: SMTP, NTP, DNS, LDAP, DHCP, Hands-on experience with networking, security, and Windows Server services, Scripting experience in PowerShell or other scripting languages, Experience with ITIL-based change, configuration, and release management, Experience working in secure isolated or tiered environments, Experience provisioning Virtual Machines in VMware or Nutanix, Experience with DevOps operations in Azure, Strong troubleshooting and analytical skills 📃 Skills: Windows, WindowsServer, PowerShell, VMware, Nutanix, Azure, SMTP, NTP, DNS, LDAP, DHCP, ITIL, DevOps, Virtualization, Networking, Security 🏢 Description: Team Connect is Poland’s leading nearshore and offshore IT provider. Since 2008 we successfully create and develop software for our clients. We specialize in Agile and DevOps-based software development. From the analysis stage through implementation. We develop backend, frontend, and mobile applications. For one of our clients, we are looking for an Windows Engineer. RESPONSIBILITIES - Administering and monitoring of Windows Server operating systems and services in a virtualized environment - Monitoring and analysis of alerts - Handling of backups and restoring operations of systems - Installation, upgrades and configuration of system components and server software - Scripting and configurations in context of Windows Server environments - Participate in architectural design/reviews - Advising in areas such as capacity management, contingency planning, IT service continuity management, automation of repetitive tasks, security - Providing with recommendations and implementing OS best practices and security standards - Problem definition, analysis, and resolution: diagnosis of software and hardware problems - Defining and updating technical documentation and operating procedures - Implementing changes according to ICT change management procedures - Providing 2nd line support to other IT teams and coordinating with 3rd line of support - Participation in coordination and project meetings - Other specific duties aligned to profile as assigned by supervisor REQUIREMENTS - Experience in administration of Windows Server 2016 or later based systems in virtualized environment - Experience in administration of complex IT systems and distributed environments - Good knowledge of network protocols: SMTP, NTP, DNS, LDAP, DHCP - Hands-on experience in networking, security and systems services for Microsoft Windows Server OS - Solid scripting skills in PowerShell or other scripting languages - Experience with ticketing systems and ITIL based change management, configuration management and release management processes - Strong analytical and troubleshooting skills - A proactive attitude, team-work spirit, being self-motivated with a strong user orientation - Good communication skills - Able to cope with fast changing technologies - Experience working in secure isolated/tiered environment - Experience in provisioning of Virtual Machines in VMware or Nutanix environment - Experience in DEVOPS operations in Azure WE OFFER Long-term cooperation Benefits package – Multisport card, medical care & life insurance Training budget Assistance in accounting issues when setting up/running a business Free english lessons Individual support from a dedicated company supervisor

Technology

Xceedance Consulting Polska Sp. z o.o.

IT Engineer

Junior

On-site

Krakow, Poland

🏢 Summary: On-site IT Support Engineer role providing Level 1–2 technical support for end users in a Poland office, ensuring reliable local IT operations and alignment with global IT processes. The position focuses on hands-on troubleshooting, asset management, Microsoft 365 support, and coordination with regional and global IT teams. 🗂️ Requirements: Bachelor’s degree in Information Technology, Computer Science, or related field or equivalent experience, 1–2 years of experience in IT support, desktop support, or service desk, Experience with Windows 10 and 11 administration and basic macOS support, Hands-on support for mobile devices, printers, and peripherals, Experience with Microsoft 365 including Outlook, Teams, and SharePoint, Experience with ticketing systems and incident management aligned with SLAs, Knowledge of IMAC activities including workstation setup and troubleshooting, Experience managing IT assets and inventory lifecycle, Understanding of identity and access management concepts, Familiarity with basic ITIL concepts including incident, request, and change management, Foundational understanding of AI concepts and use of AI-powered tools 📃 Skills: Windows, macOS, Microsoft365, Outlook, Teams, SharePoint, OneDrive, ITIL, Ticketing, IMAC, ActiveDirectory, AV, Printers, Mobile, SLA, Inventory, AI 🏢 Description: Description We are looking for an On-site IT Support Engineer to provide day-to-day technical support for our Poland office. This role is focused on end-user support, local IT operations, and close cooperation with regional and global IT teams to ensure smooth and reliable technology services. The position combines hands-on technical troubleshooting with user support, asset management, and consistent execution of global IT processes. Requirements As the on-site IT support engineer for the Poland office, this role focuses on end-user support, local infrastructure coordination, and consistent execution of global IT processes. Provide Level 1–2 on-site and remote IT support for end users (Windows/Mac, mobile devices, printers, peripherals). Log, track, and resolve incidents and service requests in the ticketing system in line with SLAs. Support employee onboarding and offboarding, including device setup, access requests, and equipment returns. Manage local IT assets (laptops, accessories, meeting room equipment), ensuring accurate inventory and lifecycle tracking. Perform IMAC activities, including workstation setup, moves, changes, and local troubleshooting. Support Microsoft 365 and collaboration tools (Outlook, Teams, SharePoint, meeting room A/V). Coordinate with regional/global IT teams and vendors for escalations and advanced support. Maintain documentation, support compliance activities, and follow information security policies. Experience, and Certifications Bachelor’s degree in Information Technology, Computer Science, or a related field (or equivalent practical experience). 1–2 years of experience in IT support / desktop support / service desk (on-site and/or remote), ideally in an office environment. Good working knowledge of end-user computing (Windows 10/11, basic Mac support), mobile devices, printers, and standard troubleshooting practices. Familiarity with Microsoft 365 (Outlook, Teams, OneDrive/SharePoint) and meeting room / A/V basics; understanding of common identity/access concepts is a plus. Experience handling IT assets and inventory updates; comfortable following standard processes for device lifecycle, repairs, and replacements. Ability to coordinate with vendors and regional/global IT teams for escalations and scheduled on-site visits, providing clear information and follow-through. Well-organized and able to prioritize multiple tickets and tasks, documenting work clearly and escalating appropriately. Experience with ticketing tools and basic ITIL concepts (incident, request, change) is an advantage. Foundational understanding of AI concepts, including how machine learning models  work, common AI use cases, and basic principles of data handling. Ability to use AI‑powered tools and platforms to support daily tasks, with strong attention to detail and commitment to maintaining data accuracy. Fluent command of the Polish language (C1/C2) What you can expect from us: Unique professional and personal development at one of the pioneer companies in professional insurance support. Ongoing professional training – to onboard you for a good start and your further professional development. Your own growth training budget - use internal and external development opportunities and take advantage of your own budget. LUXMED medical cover for you with full dental care, oncological preventive program and additional mental health support with helpline & individual sessions with a therapist.. 8-hour work time with a lunch break already included - spend the rest of the day doing what is important to you as intended in the #2h4Family program. Private life insurance - 80% of the premium is covered by an employer. Employee referral program - we appreciate you recommending your friends to join us. Additional days off - to celebrate your birthday, moreover, if you want to do a volunteering work, feel free to do so with some extra days off. Integration events - monthly delicious breakfasts, movie nights, board game nights, outdoor events. Lively and modern office in the City Centre with parking space for employees. A supportive and friendly atmosphere created by passionate people.

Technology

Yard Corporate

Senior Backend Engineer (Java, Kafka Streams, AWS)

Senior

Hybrid

Krakow, MA, Poland

30,000 - 35,000 PLN/hr

🏢 Summary: Senior/Staff Backend Engineer role focused on building and evolving a large-scale, event-driven platform using Java, Kafka Streams, and AWS. The position involves hands-on work with high-volume data flows, real-time stream processing, distributed systems, and production reliability. You will take strong technical ownership in a small team, shaping architecture, standards, and production readiness. 🗂️ Requirements: 7+ years of backend engineering experience, Strong hands-on experience with Java, Strong hands-on experience with Spring Boot, Deep production experience with Kafka Streams (topologies, joins, aggregations, windowing, state stores), Experience with event-driven or distributed systems, Hands-on experience with AWS services, Experience with Kubernetes and containerized deployments, Experience with PostgreSQL, MongoDB, and Redis, Strong knowledge of monitoring, debugging, and performance optimization, Ability to work independently in evolving technical environments, Professional English communication skills 📃 Skills: Java, Spring, Kafka, KafkaStreams, AWS, EKS, S3, SQS, SNS, Kinesis, Lambda, Kubernetes, PostgreSQL, MongoDB, Redis, Terraform, Helm, Prometheus, Grafana, Micrometer, Kotlin 🏢 Description: Senior / Staff Backend Engineer (Java, Kafka Streams, AWS) Location: Kraków, hybrid Contract: B2B or UoP B2B budget: up to 210 PLN/h + VAT Work model: around 3 days per week from the office We are looking for a Senior or Staff Backend Engineer to join a small, high-impact engineering team building a large-scale, event-driven platform. This is a hands-on role for someone who enjoys working close to real production systems: high-volume data flows, asynchronous processing, external integrations, reliability, observability and technical ownership. You will join a team where the codebase and domain are evolving, so this is a good opportunity for an engineer who does not need every problem to be fully defined before getting started. You will be expected to investigate, create structure, improve standards and make pragmatic technical decisions. What you will do Build and evolve backend services using Java and Spring Boot Design and maintain real-time event-processing flows using Kafka Streams Work with stream-processing concepts such as topologies, transformations, joins, aggregations, stateful processing and data routing Integrate external systems and infrastructure into a scalable internal platform Improve reliability, observability and production readiness of services handling high-volume data Investigate production issues, performance bottlenecks and data-flow problems Work with AWS, Kubernetes, PostgreSQL, MongoDB, Redis, Terraform and Helm Contribute to code reviews, architecture decisions, documentation and engineering standards Collaborate with product, platform and engineering stakeholders across an international environment What we are looking for 7+ years of backend engineering experience Strong, current hands-on experience with Java and Spring Boot Deep practical experience with Kafka Streams in production This should go beyond Kafka producers/consumers and include work with stream-processing logic, joins, aggregations, windowing, state stores or topology design Background in event-driven, distributed or real-time systems Hands-on AWS experience, ideally with services such as EKS, S3, SQS, SNS, Kinesis or Lambda Experience with Kubernetes , containers and production deployment environments Good understanding of databases such as PostgreSQL, MongoDB and Redis Strong production mindset: monitoring, debugging, reliability, performance and incident investigation Ability to work independently in an evolving and not always fully documented technical environment Clear English communication skills and confidence explaining technical topics to both engineers and non-technical stakeholders Nice to have Kotlin experience Experience with high-throughput data processing or streaming platforms Terraform, Helm, Prometheus, Grafana, Micrometer or similar observability tooling Experience in a small product team, scale-up or ownership-heavy engineering environment Familiarity with AI-assisted development tools What you can expect A technically challenging role focused on real production systems rather than isolated feature delivery Strong ownership Small team setup with room to influence architecture, standards and ways of working International engineering environment Clear recruitment process focused on practical engineering discussions Recruitment process Introductory conversation with the Hiring Manager Technical interview focused on Kafka/use cases and Java/code quality Final discussion with senior technical leadership The full company and product context will be shared during the first recruiter call.

Technology

Aidoc

Technical Support Engineer

Mid

Remote

Denver, CO

🏢 Summary: Shift-based Technical Support Engineer role supporting a global clinical AI SaaS platform, acting as the bridge between customers and engineering to resolve complex production issues. The position focuses on diagnosing system behavior, coordinating incident response, and ensuring platform reliability through deep troubleshooting and clear customer communication. Ideal for candidates experienced in production support within fast-paced, customer-facing technical environments. 🗂️ Requirements: 2+ years in Technical Support, Production Support, Technical Operations, or similar customer-facing technical role, Strong troubleshooting and analytical skills in complex production environments, Proficiency in SQL for data investigation and root cause analysis, Ability to analyze distributed systems and application layers, Experience collaborating with Engineering teams on escalations, Ability to handle production incidents, alerts, and time-sensitive issues, Availability for shift-based work including evenings, weekends, and holidays, Ability to conduct live troubleshooting sessions with customers 📃 Skills: SQL, APIs, REST, Datadog, Grafana, NewRelic, Python, Bash, AWS, GCP, Azure, Salesforce, Jira, Atlassian, Monday, Slack 🏢 Description: About this role We are looking for a customer-obsessed Technical Support Engineer who thrives at the intersection of customers, production systems, and engineering. In this role, you will act as a critical bridge between our customers and internal engineering teams diagnosing complex technical issues, coordinating production incident response, and helping ensure the reliability of our platform for customers around the world. You will investigate technical problems, analyze system behavior, respond to production alerts, and drive issues through to resolution while maintaining clear and proactive communication with customers and internal stakeholders. This role is ideal for someone who enjoys deep technical troubleshooting, working directly with production systems, and advocating for customers, while building strong expertise in a fast-paced SaaS environment. This is a shift-based role supporting global customers, which may include Pacific Time business hours, evenings, weekends, and some holidays. Quick Facts - Shift-based support role - Global customer-facing production support - Focus on incident response and system reliability Responsibilities Own Customer Impacting Issues - Serve as the primary technical owner for complex customer issues and escalations. - Investigate and resolve technical problems spanning multiple systems and services. - Provide clear, proactive communication to customers throughout the lifecycle of an issue. Coordinate Production Incident Response - Monitor and triage production alerts impacting customers or system reliability. - Coordinate incident response efforts across engineering and internal teams. - Ensure incidents are properly documented, communicated, and followed through to resolution. Troubleshoot Systems and Data - Diagnose issues using logs, system metrics, and SQL queries. - Analyze system behavior to identify root causes of production problems. - Escalate and partner with engineering teams to drive long-term fixes. Improve Reliability and Operational Excellence - Develop and maintain troubleshooting documentation, runbooks, and operational processes. - Identify recurring patterns and contribute to systemic improvements. - Help strengthen incident response and operational best practices as the organization scales. Enable the Broader Support Team - Share technical insights and best practices with colleagues. - Act as a technical resource within the Support organization. Impact You'll Have - Ensure platform reliability for customers worldwide by diagnosing and resolving complex production issues. - Serve as a critical bridge between customers and engineering, bringing technical insight and real-world production context to accelerate resolutions. - Strengthen production support capabilities by improving troubleshooting processes, documentation, and incident response practices. - Help scale a modern, high-impact Support organization, contributing ideas that improve how issues are detected, investigated, and resolved. Requirements - Exceptional communication skills with a customer-first mindset. - Ability to troubleshoot live with customers during critical situations. - 2+ years of experience in Technical Support, Production Support, Technical Operations, or similar roles. - Strong troubleshooting and analytical skills. - Proficiency in SQL for production data investigation and root cause analysis. - Ability to analyze distributed systems and application layers. - Experience collaborating with Engineering teams on escalations. - Comfort operating in fast-paced production environments handling incidents and alerts. Nice to Have - Experience with incident response, production alerting, or on-call rotations. - Familiarity with observability and monitoring tools. - Experience supporting SaaS or cloud-based platforms. - Working knowledge of Salesforce, Jira / Atlassian, Monday.com, or Slack. Bonus - Experience debugging API integrations, REST endpoints, request/response payloads, and authentication mechanisms. - Experience analyzing application logs, system metrics, and traces. - Familiarity with Datadog, Grafana, or New Relic. - Basic scripting or automation skills (Python, Bash, or similar). - Exposure to AWS, GCP, or Microsoft Azure. Benefits - Medical, dental, and vision benefits. - Stock options for full-time employees. - Flexible time off without accrual limits, coordinated with team and manager. - 401(k) plan with company match, life insurance, and long- and short-term disability coverage. - Opportunity to directly improve medical care and impact patient outcomes.

Technology

Team Connect

SCCM Administrator

Mid

On-site

Warsaw, Poland

140 - 160 PLN

🏢 Summary: The offer is for an SCCM Administrator responsible for managing and maintaining Microsoft SCCM environments, Windows OS patching, and software distribution across on-prem and cloud infrastructures. The role focuses on deploying updates, creating task sequences, troubleshooting issues, and supporting IT operations in complex enterprise environments. 🗂️ Requirements: Experience administering complex SCCM environments (on-prem and Cloud), Experience with Azure IaaS updates, PowerShell scripting on Windows Servers, Administration of Windows Server and Workstation systems, Experience with ticketing systems, Knowledge of ITIL-based change, configuration and release management, Strong troubleshooting and analytical skills 📃 Skills: SCCM, Azure, IaaS, PowerShell, Windows, WindowsServer, ITIL, Ticketing 🏢 Description: About Company: Team Connect is Poland’s leading nearshore and offshore IT provider. Since 2008 we successfully create and develop software for our clients. We specialize in Agile and DevOps-based software development. From the analysis stage through implementation. We develop backend, frontend, and mobile applications. For one of our clients, we are looking for an SCCM Administrator. RESPONSIBILITIES Deploying, upgrading, tuning and administering of Microsoft SCCM or other patch management solutions Monitoring, updating and upgrading Windows OS (on-prem and Cloud distributions) Managing monthly patching campaigns for Windows Server and Workstation OS Managing software distribution (creating packages for new applications’ installations and updates) Creating and updating Microsoft SCCM task sequences for Windows OS images Creating periodical patching reports Updating technical documentation and operating procedures Troubleshooting Microsoft SCCM related issues and problems Providing 2nd line support to other IT teams and coordinating with 3rd line of support Other specific duties assigned by supervisor REQUIREMENTS Experience in administering complex SCCM environments (on-prem and Cloud) Experience in Azure cloud IaaS updates Experience in scripting on Windows Servers (PowerShell) Hands-on experience in administration of MS Windows Servers and workstations Experience with ticketing systems and ITIL based change management, configuration management and release management processes Strong analytical and troubleshooting skills A proactive attitude, team-work spirit, being self-motivated with a strong user orientation Good communication skills Able to cope with fast changing technologies WE OFFER Long-term cooperation Benefits package – Multisport card, medical care & life insurance Training budget Assistance in accounting issues when setting up/running a business Free english lessons Individual support from a dedicated company supervisor

Technology

LeoVegas Mobile Gaming Group

Staff Backend Engineer - Sports

Senior

Hybrid

Warsaw, Poland

🏢 Summary: Staff Engineer role focused on leading the design, evolution, and reliability of a high-scale distributed backend platform that generates and distributes real-time customer offers. The position combines deep hands-on backend engineering with architectural leadership in event-driven systems. It emphasizes scalability, performance, observability, and operational excellence in a complex production environment. 🗂️ Requirements: Experience owning complex backend systems end-to-end in scalable production environments, Proven experience designing and operating distributed systems, Strong backend engineering expertise in Java, Solid understanding of JVM performance, Hands-on experience with Apache Kafka, Hands-on experience with Spring Boot and/or Vert.x, Hands-on experience with Redis, Hands-on experience with OpenSearch or Elasticsearch, Experience with observability tools such as Datadog, Grafana, or Kibana, Strong system design skills in event-driven and asynchronous architectures, Knowledge of concurrency and non-blocking system design, Experience with scalability, resiliency, and fault-tolerance strategies, Experience debugging production incidents in distributed environments 📃 Skills: Java, JVM, Kafka, Spring, Vert.x, Redis, OpenSearch, Elasticsearch, Datadog, Grafana, Kibana, Kubernetes 🏢 Description: ABOUT THE ROLE We are looking for a passionate and highly experienced Staff Backend Engineer to help lead the design, development, reliability, and evolution of a complex distributed backend platform. This role is ideal for someone who enjoys going deep into understanding both the technical and business context behind problems, not just implementing solutions mechanically. We value engineers who are curious, proactive, and driven by learning - people who enjoy experimenting, challenging assumptions, and troubleshooting complex systems until the root cause is fully understood. As a Staff Backend Engineer, you will provide technical leadership across architecture, engineering practices, and platform evolution while remaining hands-on when needed. You will help teams make better technical decisions, raise engineering standards, simplify complex systems, and create solutions that balance business needs with long-term platform health. We're looking for someone who is comfortable working in ambiguity, treats failure as an opportunity to learn, welcomes constructive feedback, and leads through collaboration, ownership, and execution rather than authority. ABOUT YOUR FUTURE PLATFORM TEAM We build and operate a critical backend platform that generates, validates, and distributes customer-facing offers across multiple markets in real time. The system sits at the intersection of pricing, configuration, compliance, and localization - where milliseconds matter, correctness is non-negotiable, and small mistakes can have large-scale impact. Our systems apply complex configuration and licensing rules, handle market and event-level blocking, manage multi-language localization, and transform raw pricing data into customer-ready offers consumed across the ecosystem. We process high-throughput event streams, deal with constantly changing data, and solve problems that require both strong engineering fundamentals and deep domain understanding. Success means no incorrect or missing offers in production, near real-time reflection of pricing and configuration changes, fully compliant offers across jurisdictions, and seamless multi-language experiences for customers. Our ecosystem is built around: Event streaming and backbone: Apache Kafka Backend services: Java 23+ with Spring Boot and Vert.x (we are continuously evolving toward highly reactive, asynchronous, and event-driven architecture) State and performance optimization: Redis Search and analytics: OpenSearch / Elasticsearch Observability and operations: Datadog, Grafana, Kibana Runtime and scalability: Kubernetes-based containerized infrastructure We are actively evolving our platform toward more reactive, asynchronous, and event-driven architectures, making this an exciting environment for engineers who enjoy distributed systems, performance, concurrency, and real-time data processing challenges. WHAT YOU'LL DO Technical Leadership Leading architecture and design discussions for scalable, resilient distributed systems Collaborating with Architects and senior engineers to define solution approaches and evaluate new technologies through spikes and PoCs Improving and simplifying existing system designs while maintaining high engineering standards for performance, correctness, and maintainability Maintaining a strong engineering bar across quality, scalability, and maintainability Platform Reliability & Engineering Excellence Proactively identifying scalability, reliability, and performance risks across the platform Leading debugging and root cause analysis for complex production incidents Improving observability, monitoring, alerting, and operational readiness Driving technical debt reduction and continuous platform improvement Promoting best practices across event-driven systems, resiliency patterns, performance optimization, testing, automation, and CI/CD Team Enablement & Collaboration Supporting day-to-day technical coordination within the team and contributing to incident handling and engineering discussions Mentoring engineers through design reviews, code reviews, and hands-on technical guidance Supporting cross-team collaboration and helping troubleshoot integration and platform issues Encouraging a culture of collaboration, clarity, and shared ownership Actively participating in engineering communication channels and helping unblock others Delivery & Execution Working closely with Product, Engineering Managers, and Architects to translate ideas into executable technical solutions Helping define and break down EPICs and technical initiatives in Jira Driving clarity in ambiguous or complex technical initiative Ensuring alignment between short-term delivery needs and long-term platform evolution WHAT WE'RE LOOKING FOR Required Experience: Experience owning complex backend systems end-to-end in scalable production environments Proven experience designing and operating distributed systems Strong backend engineering expertise in Java, with solid understanding of JVM performance Hands-on experience with: Apache Kafka and event-driven architectures Spring Boot and/or Vert.x Redis for low-latency and stateful use cases OpenSearch / Elasticsearch for querying and analytics Observability tools such as Datadog, Grafana, or Kibana Strong system design skills, including: Event-driven and asynchronous architectures Concurrency and non-blocking system design Scalability, resiliency, and fault-tolerance strategies Evaluating trade-offs between latency, throughput, and consistency Experience debugging production incidents in distributed environments Strong communication skills and ability to influence technical decisions across teams Proactive mindset with strong ownership and accountability Comfortable working in complex, fast-paced engineering environments Collaborative approach and willingness to support and mentor others Nice to Have: Kubernetes and cloud-native infrastructure experience Experience building internal platforms or shared engineering frameworks SRE mindset and operational excellence practices Experience with real-time or high-throughput data systems Experience leading cross-team technical initiatives WHAT SUCCESS LOOKS LIKE Teams rely on you for technical direction and practical problem-solving The platform becomes more reliable, scalable, and easier to operate over time Production incidents decrease due to proactive engineering and improved system design Cross-team dependencies are better coordinated and easier to manage Engineering quality and consistency improve across services and squads Engineers grow technically through your mentorship and guidance WHAT WE OFFER Competitive salary aligned with your experience, expertise, and impact A collaborative, inclusive, and fast-paced engineering environment where ideas are welcomed, ownership is encouraged, and continuous learning is valued Hybrid work policy 4 weeks of Workation (T&C apply) Benefits package - 800 PLN net (B2B)/gross (UOP) monthly with access to the cafeteria platform (private medical care, insurances, sports card) Training budget ICAS assistance program that can provide help and guidance during challenging moments. Our office provides complimentary snacks and drinks; on Mondays, we serve complimentary breakfast. Team and office social events throughout the year Free parking for bicycles and motorcycles Free gym in the building Equipment: MacBook Pro + smartphone (iPhone 16/17 or Samsung Galaxy S25)

Technology

Aidoc

Technical Support Engineer

Mid

Remote

New York, NY

🏢 Summary: Shift-based Technical Support Engineer role focused on diagnosing and resolving complex production issues in a global SaaS clinical AI platform. The position bridges customers and engineering, leading incident response, troubleshooting distributed systems, and ensuring platform reliability. Ideal for candidates experienced in production support, SQL-based investigation, and working in fast-paced, customer-facing technical environments. 🗂️ Requirements: 2+ years in Technical Support, Production Support, Technical Operations, or similar customer-facing technical role, Strong troubleshooting and analytical skills in production environments, Proficiency in SQL for data investigation and root cause analysis, Ability to analyze system behavior across distributed systems and application layers, Experience collaborating with Engineering teams to escalate and resolve issues, Ability to handle production incidents, alerts, and time-sensitive customer-impacting issues, Availability for shift-based work including evenings, weekends, and holidays, Ability to conduct live troubleshooting sessions with customers 📃 Skills: SQL, APIs, REST, Datadog, Grafana, NewRelic, Python, Bash, AWS, GCP, Azure, Salesforce, Jira, Atlassian, Monday, Slack 🏢 Description: About this role We are looking for a customer-obsessed Technical Support Engineer who thrives at the intersection of customers, production systems, and engineering. In this role, you will act as a critical bridge between our customers and internal engineering teams diagnosing complex technical issues, coordinating production incident response, and helping ensure the reliability of our platform for customers around the world. You will investigate technical problems, analyze system behavior, respond to production alerts, and drive issues through to resolution while maintaining clear and proactive communication with customers and internal stakeholders. This role is ideal for someone who enjoys deep technical troubleshooting, working directly with production systems, and advocating for customers, while building strong expertise in a fast-paced SaaS environment. This is a shift-based role supporting our global customers, which may include Pacific Time business hours, evenings, weekends, and some holidays. Responsibilities Own Customer Impacting Issues - Serve as the primary technical owner for complex customer issues and escalations. - Investigate and resolve technical problems spanning multiple systems and services. - Provide clear, proactive communication to customers throughout the lifecycle of an issue. Coordinate Production Incident Response - Monitor and triage production alerts impacting customers or system reliability. - Coordinate incident response efforts across engineering and internal teams. - Ensure incidents are properly documented, communicated, and followed through to resolution. Troubleshoot Systems and Data - Diagnose issues using logs, system metrics, and SQL queries. - Analyze system behavior to identify root causes of production problems. - Escalate and partner with engineering teams to drive long-term fixes. Improve Reliability and Operational Excellence - Develop and maintain troubleshooting documentation, runbooks, and operational processes. - Identify recurring patterns and contribute to systemic improvements. - Help strengthen incident response and operational best practices as the organization scales. Enable the Broader Support Team - Share technical insights and best practices with colleagues. - Act as a technical resource within the Support organization. Impact You'll Have - Ensure platform reliability for customers worldwide by diagnosing and resolving complex production issues. - Serve as a critical bridge between customers and engineering, bringing technical insight and real-world production context to accelerate resolutions. - Strengthen our production support capabilities by improving troubleshooting processes, documentation, and incident response practices. - Help scale a modern, high-impact Support organization, contributing ideas that improve how we detect, investigate, and resolve issues as the company grows. Requirements - Exceptional communication skills with a customer-first mindset, capable of translating complex technical issues into clear and actionable insights. - A demonstrated ability to get on a call or have a remote session with a customer during less-than-optimal times, to troubleshoot and put them at ease. - 2+ years of experience in Technical Support, Production Support, Technical Operations, or a similar customer-facing technical role. - Strong troubleshooting and analytical skills, with the ability to quickly diagnose and resolve complex technical issues. - Proficiency in SQL for data investigation, troubleshooting, and root cause analysis within production environments. - Ability to analyze system behavior, investigate anomalies, and debug issues across distributed systems and application layers. - Experience partnering closely with Engineering teams to escalate issues, provide technical context, and drive timely resolution. - Comfortable operating in fast-paced production environments, including handling incidents, alerts, and time-sensitive customer-impacting issues. Nice to Have - Experience participating in incident response, production alerting, or on-call rotations in a live production environment. - Familiarity with observability, monitoring, and debugging tools used to investigate system performance and reliability issues. - Experience supporting or operating within SaaS or cloud-based platforms. - Working knowledge of support, collaboration, and ticketing tools such as Salesforce, Jira / Atlassian, Monday.com, or Slack. Bonus - Experience debugging API integrations, working with REST endpoints, request/response payloads, and authentication mechanisms. - Comfort analyzing application logs, system metrics, and traces to diagnose production issues. - Familiarity with modern observability and monitoring platforms such as Datadog, Grafana, or New Relic. - Basic scripting or automation skills (Python, Bash, or similar) to streamline investigations or repetitive support workflows. - Exposure to cloud infrastructure platforms such as Amazon Web Services, Google Cloud Platform, or Microsoft Azure. What we offer: - A range of medical, dental and vision benefits - Stock options for all full-time employees - Flexible time off to enjoy the autonomy to take time off as needed to rest and recharge without vacation accrual limits, while coordinating with your manager and team to ensure business continuity and project goals are met. - A 401(k) plan with company match, life insurance, plus long- and short-term disability - The opportunity to directly improve medical care and impact patient outcomes

Technology

Aidoc

Technical Support Engineer

Mid

Remote

Chicago, IL

🏢 Summary: Shift-based Technical Support Engineer role focused on diagnosing and resolving complex production issues in a SaaS clinical AI platform, acting as the bridge between customers and engineering. The position involves incident response, system troubleshooting, SQL-based investigations, and improving operational reliability for global customers. Ideal for candidates experienced in production support within fast-paced, cloud-based environments. 🗂️ Requirements: 2+ years in Technical Support, Production Support, Technical Operations, or similar customer-facing technical role, Strong troubleshooting and analytical skills in complex production environments, Proficiency in SQL for data investigation and root cause analysis, Ability to analyze distributed systems and application layers, Experience collaborating with Engineering teams on escalations and resolutions, Experience handling production incidents, alerts, and customer-impacting issues, Availability for shift-based work including evenings, weekends, and holidays 📃 Skills: SQL, SaaS, REST, API, Datadog, Grafana, NewRelic, Python, Bash, AWS, GCP, Azure, Salesforce, Jira, Atlassian, Monday.com, Slack 🏢 Description: About this role We are looking for a customer-obsessed Technical Support Engineer who thrives at the intersection of customers, production systems, and engineering. In this role, you will act as a critical bridge between our customers and internal engineering teams diagnosing complex technical issues, coordinating production incident response, and helping ensure the reliability of our platform for customers around the world. You will investigate technical problems, analyze system behavior, respond to production alerts, and drive issues through to resolution while maintaining clear and proactive communication with customers and internal stakeholders. This role is ideal for someone who enjoys deep technical troubleshooting, working directly with production systems, and advocating for customers, while building strong expertise in a fast-paced SaaS environment. This is a shift-based role supporting our global customers, which may include Pacific Time business hours, evenings, weekends, and some holidays. Responsibilities Own Customer Impacting Issues Serve as the primary technical owner for complex customer issues and escalations. Investigate and resolve technical problems spanning multiple systems and services. Provide clear, proactive communication to customers throughout the lifecycle of an issue. Coordinate Production Incident Response Monitor and triage production alerts impacting customers or system reliability. Coordinate incident response efforts across engineering and internal teams. Ensure incidents are properly documented, communicated, and followed through to resolution. Troubleshoot Systems and Data Diagnose issues using logs, system metrics, and SQL queries. Analyze system behavior to identify root causes of production problems. Escalate and partner with engineering teams to drive long-term fixes. Improve Reliability and Operational Excellence Develop and maintain troubleshooting documentation, runbooks, and operational processes. Identify recurring patterns and contribute to systemic improvements. Help strengthen incident response and operational best practices as the organization scales. Enable the Broader Support Team Share technical insights and best practices with colleagues. Act as a technical resource within the Support organization. Impact You'll Have Ensure platform reliability for customers worldwide by diagnosing and resolving complex production issues. Serve as a critical bridge between customers and engineering, bringing technical insight and real-world production context to accelerate resolutions. Strengthen our production support capabilities by improving troubleshooting processes, documentation, and incident response practices. Help scale a modern, high-impact Support organization, contributing ideas that improve how we detect, investigate, and resolve issues as the company grows. Requirements Exceptional communication skills with a customer-first mindset, capable of translating complex technical issues into clear and actionable insights. A demonstrated ability to get on a call or have a remote session with a customer during less-than-optimal times, to troubleshoot and put them at ease. 2+ years of experience in Technical Support, Production Support, Technical Operations, or a similar customer-facing technical role. Strong troubleshooting and analytical skills, with the ability to quickly diagnose and resolve complex technical issues. Proficiency in SQL for data investigation, troubleshooting, and root cause analysis within production environments. Ability to analyze system behavior, investigate anomalies, and debug issues across distributed systems and application layers. Experience partnering closely with Engineering teams to escalate issues, provide technical context, and drive timely resolution. Comfortable operating in fast-paced production environments, including handling incidents, alerts, and time-sensitive customer-impacting issues. Nice to Have Experience participating in incident response, production alerting, or on-call rotations in a live production environment. Familiarity with observability, monitoring, and debugging tools used to investigate system performance and reliability issues. Experience supporting or operating within SaaS or cloud-based platforms. Working knowledge of support, collaboration, and ticketing tools such as Salesforce, Jira / Atlassian, Monday.com, or Slack. Bonus Experience debugging API integrations, working with REST endpoints, request/response payloads, and authentication mechanisms. Comfort analyzing application logs, system metrics, and traces to diagnose production issues. Familiarity with modern observability and monitoring platforms such as Datadog, Grafana, or New Relic. Basic scripting or automation skills (Python, Bash, or similar) to streamline investigations or repetitive support workflows. Exposure to cloud infrastructure platforms such as Amazon Web Services, Google Cloud Platform, or Microsoft Azure. What we offer: A range of medical, dental and vision benefits Stock options for all full-time employees Flexible time off to enjoy the autonomy to take time off as needed to rest and recharge without vacation accrual limits, while coordinating with your manager and team to ensure business continuity and project goals are met. A 401(k) plan with company match, life insurance, plus long- and short-term disability The opportunity to directly improve medical care and impact patient outcomes Aidoc is deeply committed to creating an inclusive workplace, and to the principle of equal opportunity for all individuals. We prohibit discrimination and harassment based on race, color, religion, sex, sexual orientation, gender identity, national origin, age, disability, veteran status, or any other status protected by law.

Technology

Aidoc

Technical Support Engineer

Mid

Remote

San Francisco, CA

🏢 Summary: Shift-based Technical Support Engineer role focused on diagnosing and resolving complex production issues in a global SaaS clinical AI platform. The position acts as a bridge between customers and engineering, owning customer-impacting incidents, troubleshooting distributed systems, and coordinating incident response. Ideal for candidates experienced in SQL-driven investigations, production environments, and cross-team technical collaboration. 🗂️ Requirements: 2+ years in Technical Support, Production Support, Technical Operations, or similar customer-facing technical role, Strong troubleshooting and analytical skills in production environments, Proficiency in SQL for data investigation and root cause analysis, Ability to debug issues across distributed systems and application layers, Experience collaborating with Engineering teams on escalations and resolutions, Ability to handle production incidents, alerts, and time-sensitive customer issues, Willingness to work shift-based schedule including evenings, weekends, and holidays, Ability to conduct live troubleshooting sessions with customers 📃 Skills: SQL, SaaS, REST, APIs, Datadog, Grafana, NewRelic, Python, Bash, AWS, GCP, Azure, Salesforce, Jira, Atlassian, Monday, Slack 🏢 Description: About this role We are looking for a customer-obsessed Technical Support Engineer who thrives at the intersection of customers, production systems, and engineering. In this role, you will act as a critical bridge between customers and internal engineering teams, diagnosing complex technical issues, coordinating production incident response, and ensuring the reliability of the platform for customers around the world. You will investigate technical problems, analyze system behavior, respond to production alerts, and drive issues through to resolution while maintaining clear and proactive communication with customers and internal stakeholders. This is a shift-based role supporting global customers, including Pacific Time business hours, evenings, weekends, and some holidays. Responsibilities Own Customer Impacting Issues - Serve as the primary technical owner for complex customer issues and escalations. - Investigate and resolve technical problems spanning multiple systems and services. - Provide clear, proactive communication to customers throughout the lifecycle of an issue. Coordinate Production Incident Response - Monitor and triage production alerts impacting customers or system reliability. - Coordinate incident response efforts across engineering and internal teams. - Ensure incidents are properly documented, communicated, and followed through to resolution. Troubleshoot Systems and Data - Diagnose issues using logs, system metrics, and SQL queries. - Analyze system behavior to identify root causes of production problems. - Escalate and partner with engineering teams to drive long-term fixes. Improve Reliability and Operational Excellence - Develop and maintain troubleshooting documentation, runbooks, and operational processes. - Identify recurring patterns and contribute to systemic improvements. - Help strengthen incident response and operational best practices as the organization scales. Enable the Broader Support Team - Share technical insights and best practices with colleagues. - Act as a technical resource within the Support organization. Impact You'll Have - Ensure platform reliability for customers worldwide by diagnosing and resolving complex production issues. - Serve as a bridge between customers and engineering, bringing technical insight and production context to accelerate resolutions. - Strengthen production support capabilities by improving troubleshooting processes, documentation, and incident response practices. - Help scale a modern Support organization, contributing ideas that improve detection, investigation, and resolution of issues. Requirements - Exceptional communication skills with a customer-first mindset. - Ability to conduct remote troubleshooting sessions with customers. - 2+ years of experience in Technical Support, Production Support, Technical Operations, or similar role. - Strong troubleshooting and analytical skills. - Proficiency in SQL for production data investigation and root cause analysis. - Ability to debug distributed systems and application layers. - Experience partnering with Engineering teams to drive timely resolutions. - Comfort operating in fast-paced production environments handling incidents and alerts. Nice to Have - Experience with incident response, production alerting, or on-call rotations. - Familiarity with observability, monitoring, and debugging tools. - Experience supporting SaaS or cloud-based platforms. - Working knowledge of tools such as Salesforce, Jira / Atlassian, Monday.com, or Slack. Bonus - Experience debugging API integrations, REST endpoints, request/response payloads, and authentication mechanisms. - Experience analyzing application logs, system metrics, and traces. - Familiarity with Datadog, Grafana, or New Relic. - Basic scripting or automation skills (Python, Bash, or similar). - Exposure to cloud platforms such as Amazon Web Services, Google Cloud Platform, or Microsoft Azure. What We Offer - Medical, dental, and vision benefits. - Stock options for full-time employees. - Flexible time off without accrual limits, coordinated with team needs. - 401(k) plan with company match, life insurance, and long- and short-term disability coverage. - Opportunity to directly improve medical care and patient outcomes.

Technology

emagine Polska

Application Support Engineer

Senior

Hybrid

Gdansk, Poland

130 - 130 PLN/hr

🏢 Summary: Hybrid B2B role for an Application Support Engineer focused on production support, DevOps practices, and CI/CD pipeline management in complex environments. The position involves leading incident resolution, optimizing service performance, managing containerized deployments, and driving automation and Infrastructure as Code initiatives. The role also supports AI/ML workloads and continuous improvement of operational processes. 🗂️ Requirements: Experience with Docker containerization, Experience with Kubernetes orchestration, Proficiency with CI/CD tools (Jenkins, Maven), Advanced troubleshooting using Splunk, Experience with relational databases (Oracle, DB2), Strong knowledge of DevOps practices, Experience managing production and pre-production environments, Ability to lead RCA and incident management processes 📃 Skills: Docker, Kubernetes, Jenkins, Maven, Splunk, Oracle, DB2, DevOps, CI/CD, RCA, IaC 🏢 Description: B2B contract rate 130 pln/h + VAT work mode: hybrid, Gdańsk We are seeking an experienced Application Support Engineer to join the team. The ideal candidate will have a deep understanding of DevOps practices, innovative solutions, and the ability to manage complex environments effectively. Main Responsibilities Lead RCA processes and implement preventive measures for incident & problem management. Monitor, analyze, and optimize service performance and stability in production support. Build, maintain, and enhance deployment pipelines for multiple APIs in CI/CD management. Maintain Pre-Production and Production environments through effective environment management. Collaborate closely with Platform Operations teams and the delivery factory to foster cross-team collaboration. Create comprehensive runbooks and knowledge bases for operational procedures. Evaluate and implement new DevOps tools, practices, and AI-powered solutions to drive innovation leadership. Implement Infrastructure as Code practices and drive automation for self-healing infrastructure solutions. Support AI/ML workload deployments and operations. Lead the adoption of emerging technologies and best practices. Key Requirements Experience with Docker containerization and Kubernetes orchestration. Proficient in CI/CD tools such as Jenkins and Maven. Advanced troubleshooting skills using tools like Splunk. Proficient in relational databases, specifically Oracle and DB2. Nice to Have Experience with Infrastructure as Code using Ansible. Basic knowledge of Spring, Hibernate, XML, JSON, and Kafka. Experience with modern observability tools such as Prometheus, Grafana, and ELK stack. Experience in AI/ML operations (MLOps) and AI agent deployments. Knowledge of AI model serving platforms and ML workload containerization. Strong understanding of AI agent orchestration and workflow automation. Ability to evaluate and integrate new DevOps tools and practices. Software development knowledge in Java/J2EE. Experience with cloud platforms (AWS, Azure, or GCP). Familiarity with IBM Datastage.