June 20, 2026

Lead C++ Developer

Senior • Remote

Warsaw, Poland

We are seeking a talented Lead C++ Developer to join the highly interdisciplinary CoreML team, where you will drive the performance and optimization of both training and serving, delivering massive impact for customers. In this role, you will have exposure to the newest Tensor Processing Unit (TPU), Graphics Processing Unit (GPU) hardware, the latest ML models, and the advanced toolchains that bridge them. Your work will directly enable AI research, production deployments and the broader open-source ecosystem, addressing complex technical issues that directly impact the efficiency and scalability of AI across the industry.

Responsibilities

  • Design and optimize high-performance kernels (using languages like Pallas, Mosaic and Triton) targeting Tensor Processing Unit (TPU) and Graphics Processing Unit (GPU) architectures for critical Machine Learning (ML) operations, redefining what's possible from massive training runs to high-speed inference
  • Architecture of infrastructure such as benchmarking suites, autotuning frameworks, performance analysis tools, regression testing and documentation, transforming how the developer community interacts with increasingly critical custom kernels in key Open-Source Software (OSS) libraries
  • Track the latest advancements in hardware architectures, compiler technologies and AI models to identify new opportunities for performance optimization through custom kernels
  • Engagement with ML researchers, framework developers (Just After eXecution (JAX), PyTorch) and compiler engineers (Accelerated Linear Algebra (XLA)) to enhance adoption, identify new requirements and address bottlenecks by providing appropriate solutions

Requirements

  • Bachelor's degree or equivalent practical experience
  • Overall 7+ years of industry experience
  • 5 years of experience with software development in C++ or Python
  • 3 years of experience testing, maintaining or launching software products, and 1 year of experience with software design and architecture
  • Expertise in performance optimization at the kernel level
  • English proficiency at B2 level or higher

Nice to have

  • Skills in optimizing TPU/GPU code, using low-level kernel languages like Pallas, Compute Unified Device Architecture (CUDA) or Triton
  • Knowledge of ML Frameworks (JAX/PyTorch), common operations like attention and Mixture of Experts (MoEs), including model optimization and low-precision formats
  • Understanding of modern accelerators (e.g., data movement, pipelining, heterogeneous compute and scale-out)
  • Understanding of compiler principles (optimization, code generation) and toolchains such as MLIR and OpenXLA
  • Showcase of building developer infrastructure, including Open-Source Software (OSS) libraries, flexible high-performance APIs and easy-to-consume documentation to empower the community

We offer

  • We gather like-minded people:
    • Engineering community of industry professionals
    • Friendly team and enjoyable working environment
    • Flexible schedule and opportunity to work remotely within Poland
    • Chance to work abroad for up to 60 days annually
    • Business-driven relocation opportunities
  • We provide growth opportunities:
    • Outstanding career roadmap
    • Leadership development, career advising, soft skills, and well-being programs
    • Certification (GCP, Azure, AWS)
    • Unlimited access to LinkedIn Learning, Get Abstract, Cloud Guru
    • English classes
  • We cover it all:
    • Stable income (Employment Contract or B2B)
    • Participation in the Employee Stock Purchase Plan
    • Benefits package (health insurance, multisport, shopping vouchers)
    • Strategically located offices featuring entertainment and relaxation zones, table tennis and football, free snacks, fantastic coffee, and more
    • Referral bonuses
    • Corporate, social and well-being events
  • Please, note:
    • The set of bonuses might vary based on the role you apply for – specifics will be discussed with our recruiter during the general interview.
    • We will reach out to selected candidates exclusively.

EPAM is a leading global provider of digital platform engineering and development services. We are committed to having a positive impact on our customers, our employees, and our communities. We embrace a dynamic and inclusive culture. Here you will collaborate with multi-national teams, contribute to a myriad of innovative projects that deliver the most creative and cutting-edge solutions, and have an opportunity to continuously learn and grow. No matter where you are located, you will join a dedicated, creative, and diverse community that will help you discover your fullest potential.

Similar jobs you might like

Technology

EPAM Systems

Lead C++ Developer

Senior

Remote

Krakow, Poland

🏢 Summary: Lead C++ Developer role within a CoreML team focused on optimizing high-performance kernels for TPU and GPU architectures to improve large-scale ML training and inference. The position involves designing developer infrastructure, enhancing benchmarking and autotuning frameworks, and collaborating with ML and compiler teams to drive AI performance and scalability. The role directly impacts AI research, production systems, and open-source ecosystems through advanced performance engineering. 🗂️ Requirements: Bachelor’s degree or equivalent practical experience, 7+ years of industry experience, 5+ years of software development experience in C++ or Python, 3+ years of experience testing, maintaining or launching software products, 1+ year of experience in software design and architecture, Expertise in kernel-level performance optimization, English proficiency at B2 level or higher 📃 Skills: C++, Python, TPU, GPU, Pallas, Mosaic, Triton, CUDA, JAX, PyTorch, XLA, MLIR, OpenXLA 🏢 Description: We are seeking a talented Lead C++ Developer to join the highly interdisciplinary CoreML team, where you will drive the performance and optimization of both training and serving, delivering massive impact for customers. In this role, you will have exposure to the newest Tensor Processing Unit (TPU), Graphics Processing Unit (GPU) hardware, the latest ML models, and the advanced toolchains that bridge them. Your work will directly enable AI research, production deployments and the broader open-source ecosystem, addressing complex technical issues that directly impact the efficiency and scalability of AI across the industry. Responsibilities Design and optimize high-performance kernels (using languages like Pallas, Mosaic and Triton) targeting Tensor Processing Unit (TPU) and Graphics Processing Unit (GPU) architectures for critical Machine Learning (ML) operations, redefining what's possible from massive training runs to high-speed inference Architecture of infrastructure such as benchmarking suites, autotuning frameworks, performance analysis tools, regression testing and documentation, transforming how the developer community interacts with increasingly critical custom kernels in key Open-Source Software (OSS) libraries Track the latest advancements in hardware architectures, compiler technologies and AI models to identify new opportunities for performance optimization through custom kernels Engagement with ML researchers, framework developers (Just After eXecution (JAX), PyTorch) and compiler engineers (Accelerated Linear Algebra (XLA)) to enhance adoption, identify new requirements and address bottlenecks by providing appropriate solutions Requirements Bachelor's degree or equivalent practical experience Overall 7+ years of industry experience 5 years of experience with software development in C++ or Python 3 years of experience testing, maintaining or launching software products, and 1 year of experience with software design and architecture Expertise in performance optimization at the kernel level English proficiency at B2 level or higher Nice to have Skills in optimizing TPU/GPU code, using low-level kernel languages like Pallas, Compute Unified Device Architecture (CUDA) or Triton Knowledge of ML Frameworks (JAX/PyTorch), common operations like attention and Mixture of Experts (MoEs), including model optimization and low-precision formats Understanding of modern accelerators (e.g., data movement, pipelining, heterogeneous compute and scale-out) Understanding of compiler principles (optimization, code generation) and toolchains such as MLIR and OpenXLA Showcase of building developer infrastructure, including Open-Source Software (OSS) libraries, flexible high-performance APIs and easy-to-consume documentation to empower the community We offer We gather like-minded people: Engineering community of industry professionals Friendly team and enjoyable working environment Flexible schedule and opportunity to work remotely within Poland Chance to work abroad for up to 60 days annually Business-driven relocation opportunities We provide growth opportunities: Outstanding career roadmap Leadership development, career advising, soft skills, and well-being programs Certification (GCP, Azure, AWS) Unlimited access to LinkedIn Learning, Get Abstract, Cloud Guru English classes We cover it all: Stable income (Employment Contract or B2B) Participation in the Employee Stock Purchase Plan Benefits package (health insurance, multisport, shopping vouchers) Strategically located offices featuring entertainment and relaxation zones, table tennis and football, free snacks, fantastic coffee, and more Referral bonuses Corporate, social and well-being events Please, note: The set of bonuses might vary based on the role you apply for – specifics will be discussed with our recruiter during the general interview. We will reach out to selected candidates exclusively. EPAM is a leading global provider of digital platform engineering and development services. We are committed to having a positive impact on our customers, our employees, and our communities. We embrace a dynamic and inclusive culture. Here you will collaborate with multi-national teams, contribute to a myriad of innovative projects that deliver the most creative and cutting-edge solutions, and have an opportunity to continuously learn and grow. No matter where you are located, you will join a dedicated, creative, and diverse community that will help you discover your fullest potential.

Technology

EPAM Systems

Lead C++ Developer

Senior

Remote

Lodz, Poland

🏢 Summary: Lead C++ Developer role focused on building and optimizing high-performance ML kernels for TPU and GPU architectures within a CoreML team. The position drives performance improvements across training and inference by developing custom kernels, benchmarking infrastructure and performance tooling while collaborating with ML researchers and compiler engineers. The role directly impacts AI scalability, efficiency and open-source ML ecosystems. 🗂️ Requirements: Bachelor’s degree or equivalent practical experience, 7+ years of industry experience, 5+ years of software development experience in C++ or Python, 3+ years of experience testing, maintaining or launching software products, 1+ year of experience in software design and architecture, Expertise in kernel-level performance optimization, English proficiency at B2 level or higher 📃 Skills: C++, Python, TPU, GPU, Pallas, Mosaic, Triton, CUDA, JAX, PyTorch, XLA, MLIR, OpenXLA 🏢 Description: We are seeking a talented Lead C++ Developer to join the highly interdisciplinary CoreML team, where you will drive the performance and optimization of both training and serving, delivering massive impact for customers. In this role, you will have exposure to the newest Tensor Processing Unit (TPU), Graphics Processing Unit (GPU) hardware, the latest ML models, and the advanced toolchains that bridge them. Your work will directly enable AI research, production deployments and the broader open-source ecosystem, addressing complex technical issues that directly impact the efficiency and scalability of AI across the industry. Responsibilities Design and optimize high-performance kernels (using languages like Pallas, Mosaic and Triton) targeting Tensor Processing Unit (TPU) and Graphics Processing Unit (GPU) architectures for critical Machine Learning (ML) operations, redefining what's possible from massive training runs to high-speed inference Architecture of infrastructure such as benchmarking suites, autotuning frameworks, performance analysis tools, regression testing and documentation, transforming how the developer community interacts with increasingly critical custom kernels in key Open-Source Software (OSS) libraries Track the latest advancements in hardware architectures, compiler technologies and AI models to identify new opportunities for performance optimization through custom kernels Engagement with ML researchers, framework developers (Just After eXecution (JAX), PyTorch) and compiler engineers (Accelerated Linear Algebra (XLA)) to enhance adoption, identify new requirements and address bottlenecks by providing appropriate solutions Requirements Bachelor's degree or equivalent practical experience Overall 7+ years of industry experience 5 years of experience with software development in C++ or Python 3 years of experience testing, maintaining or launching software products, and 1 year of experience with software design and architecture Expertise in performance optimization at the kernel level English proficiency at B2 level or higher Nice to have Skills in optimizing TPU/GPU code, using low-level kernel languages like Pallas, Compute Unified Device Architecture (CUDA) or Triton Knowledge of ML Frameworks (JAX/PyTorch), common operations like attention and Mixture of Experts (MoEs), including model optimization and low-precision formats Understanding of modern accelerators (e.g., data movement, pipelining, heterogeneous compute and scale-out) Understanding of compiler principles (optimization, code generation) and toolchains such as MLIR and OpenXLA Showcase of building developer infrastructure, including Open-Source Software (OSS) libraries, flexible high-performance APIs and easy-to-consume documentation to empower the community We offer We gather like-minded people: Engineering community of industry professionals Friendly team and enjoyable working environment Flexible schedule and opportunity to work remotely within Poland Chance to work abroad for up to 60 days annually Business-driven relocation opportunities We provide growth opportunities: Outstanding career roadmap Leadership development, career advising, soft skills, and well-being programs Certification (GCP, Azure, AWS) Unlimited access to LinkedIn Learning, Get Abstract, Cloud Guru English classes We cover it all: Stable income (Employment Contract or B2B) Participation in the Employee Stock Purchase Plan Benefits package (health insurance, multisport, shopping vouchers) Strategically located offices featuring entertainment and relaxation zones, table tennis and football, free snacks, fantastic coffee, and more Referral bonuses Corporate, social and well-being events Please, note: The set of bonuses might vary based on the role you apply for – specifics will be discussed with our recruiter during the general interview. We will reach out to selected candidates exclusively. EPAM is a leading global provider of digital platform engineering and development services. We are committed to having a positive impact on our customers, our employees, and our communities. We embrace a dynamic and inclusive culture. Here you will collaborate with multi-national teams, contribute to a myriad of innovative projects that deliver the most creative and cutting-edge solutions, and have an opportunity to continuously learn and grow. No matter where you are located, you will join a dedicated, creative, and diverse community that will help you discover your fullest potential.

Technology

EPAM Systems

Lead C++ Developer

Senior

Remote

Poznan, Poland

🏢 Summary: Lead C++ Developer role within a CoreML team focused on designing and optimizing high-performance kernels for TPU and GPU architectures to enhance machine learning training and inference. The position involves building performance infrastructure, collaborating with ML researchers and compiler engineers, and driving large-scale AI efficiency through advanced tooling and custom kernel development. Offers flexible work options, professional growth programs, and comprehensive benefits. 🗂️ Requirements: Bachelor's degree or equivalent practical experience, 7+ years of industry experience, 5+ years of software development experience in C++ or Python, 3+ years of experience testing, maintaining or launching software products, 1+ year of experience in software design and architecture, Expertise in kernel-level performance optimization, English proficiency at B2 level or higher 📃 Skills: C++, Python, TPU, GPU, Pallas, Mosaic, Triton, CUDA, JAX, PyTorch, XLA, MLIR, OpenXLA 🏢 Description: We are seeking a talented Lead C++ Developer to join the highly interdisciplinary CoreML team, where you will drive the performance and optimization of both training and serving, delivering massive impact for customers. In this role, you will have exposure to the newest Tensor Processing Unit (TPU), Graphics Processing Unit (GPU) hardware, the latest ML models, and the advanced toolchains that bridge them. Your work will directly enable AI research, production deployments and the broader open-source ecosystem, addressing complex technical issues that directly impact the efficiency and scalability of AI across the industry. Responsibilities Design and optimize high-performance kernels (using languages like Pallas, Mosaic and Triton) targeting Tensor Processing Unit (TPU) and Graphics Processing Unit (GPU) architectures for critical Machine Learning (ML) operations, redefining what's possible from massive training runs to high-speed inference Architecture of infrastructure such as benchmarking suites, autotuning frameworks, performance analysis tools, regression testing and documentation, transforming how the developer community interacts with increasingly critical custom kernels in key Open-Source Software (OSS) libraries Track the latest advancements in hardware architectures, compiler technologies and AI models to identify new opportunities for performance optimization through custom kernels Engagement with ML researchers, framework developers (Just After eXecution (JAX), PyTorch) and compiler engineers (Accelerated Linear Algebra (XLA)) to enhance adoption, identify new requirements and address bottlenecks by providing appropriate solutions Requirements Bachelor's degree or equivalent practical experience Overall 7+ years of industry experience 5 years of experience with software development in C++ or Python 3 years of experience testing, maintaining or launching software products, and 1 year of experience with software design and architecture Expertise in performance optimization at the kernel level English proficiency at B2 level or higher Nice to have Skills in optimizing TPU/GPU code, using low-level kernel languages like Pallas, Compute Unified Device Architecture (CUDA) or Triton Knowledge of ML Frameworks (JAX/PyTorch), common operations like attention and Mixture of Experts (MoEs), including model optimization and low-precision formats Understanding of modern accelerators (e.g., data movement, pipelining, heterogeneous compute and scale-out) Understanding of compiler principles (optimization, code generation) and toolchains such as MLIR and OpenXLA Showcase of building developer infrastructure, including Open-Source Software (OSS) libraries, flexible high-performance APIs and easy-to-consume documentation to empower the community We offer We gather like-minded people: Engineering community of industry professionals Friendly team and enjoyable working environment Flexible schedule and opportunity to work remotely within Poland Chance to work abroad for up to 60 days annually Business-driven relocation opportunities We provide growth opportunities: Outstanding career roadmap Leadership development, career advising, soft skills, and well-being programs Certification (GCP, Azure, AWS) Unlimited access to LinkedIn Learning, Get Abstract, Cloud Guru English classes We cover it all: Stable income (Employment Contract or B2B) Participation in the Employee Stock Purchase Plan Benefits package (health insurance, multisport, shopping vouchers) Strategically located offices featuring entertainment and relaxation zones, table tennis and football, free snacks, fantastic coffee, and more Referral bonuses Corporate, social and well-being events Please, note: The set of bonuses might vary based on the role you apply for – specifics will be discussed with our recruiter during the general interview. We will reach out to selected candidates exclusively. EPAM is a leading global provider of digital platform engineering and development services. We are committed to having a positive impact on our customers, our employees, and our communities. We embrace a dynamic and inclusive culture. Here you will collaborate with multi-national teams, contribute to a myriad of innovative projects that deliver the most creative and cutting-edge solutions, and have an opportunity to continuously learn and grow. No matter where you are located, you will join a dedicated, creative, and diverse community that will help you discover your fullest potential.

Technology

EPAM Systems

Lead C++ Developer

Senior

Remote

Katowice, SL, Poland

🏢 Summary: Lead C++ Developer role focused on designing and optimizing high-performance ML kernels for TPU and GPU architectures within a CoreML team. The position drives performance improvements across training and inference, builds benchmarking and autotuning infrastructure, and collaborates with ML researchers and compiler engineers. It offers exposure to cutting-edge hardware, AI models and toolchains, impacting large-scale AI research and production deployments. 🗂️ Requirements: Bachelor’s degree or equivalent practical experience, 7+ years of industry experience, 5+ years of software development experience in C++ or Python, 3+ years of experience testing, maintaining or launching software products, 1+ year of experience in software design and architecture, Expertise in kernel-level performance optimization, English proficiency at B2 level or higher 📃 Skills: C++, Python, TPU, GPU, Pallas, Mosaic, Triton, CUDA, JAX, PyTorch, XLA, MLIR, OpenXLA, OSS 🏢 Description: We are seeking a talented Lead C++ Developer to join the highly interdisciplinary CoreML team, where you will drive the performance and optimization of both training and serving, delivering massive impact for customers. In this role, you will have exposure to the newest Tensor Processing Unit (TPU), Graphics Processing Unit (GPU) hardware, the latest ML models, and the advanced toolchains that bridge them. Your work will directly enable AI research, production deployments and the broader open-source ecosystem, addressing complex technical issues that directly impact the efficiency and scalability of AI across the industry. Responsibilities Design and optimize high-performance kernels (using languages like Pallas, Mosaic and Triton) targeting Tensor Processing Unit (TPU) and Graphics Processing Unit (GPU) architectures for critical Machine Learning (ML) operations, redefining what's possible from massive training runs to high-speed inference Architecture of infrastructure such as benchmarking suites, autotuning frameworks, performance analysis tools, regression testing and documentation, transforming how the developer community interacts with increasingly critical custom kernels in key Open-Source Software (OSS) libraries Track the latest advancements in hardware architectures, compiler technologies and AI models to identify new opportunities for performance optimization through custom kernels Engagement with ML researchers, framework developers (Just After eXecution (JAX), PyTorch) and compiler engineers (Accelerated Linear Algebra (XLA)) to enhance adoption, identify new requirements and address bottlenecks by providing appropriate solutions Requirements Bachelor's degree or equivalent practical experience Overall 7+ years of industry experience 5 years of experience with software development in C++ or Python 3 years of experience testing, maintaining or launching software products, and 1 year of experience with software design and architecture Expertise in performance optimization at the kernel level English proficiency at B2 level or higher Nice to have Skills in optimizing TPU/GPU code, using low-level kernel languages like Pallas, Compute Unified Device Architecture (CUDA) or Triton Knowledge of ML Frameworks (JAX/PyTorch), common operations like attention and Mixture of Experts (MoEs), including model optimization and low-precision formats Understanding of modern accelerators (e.g., data movement, pipelining, heterogeneous compute and scale-out) Understanding of compiler principles (optimization, code generation) and toolchains such as MLIR and OpenXLA Showcase of building developer infrastructure, including Open-Source Software (OSS) libraries, flexible high-performance APIs and easy-to-consume documentation to empower the community We offer We gather like-minded people: Engineering community of industry professionals Friendly team and enjoyable working environment Flexible schedule and opportunity to work remotely within Poland Chance to work abroad for up to 60 days annually Business-driven relocation opportunities We provide growth opportunities: Outstanding career roadmap Leadership development, career advising, soft skills, and well-being programs Certification (GCP, Azure, AWS) Unlimited access to LinkedIn Learning, Get Abstract, Cloud Guru English classes We cover it all: Stable income (Employment Contract or B2B) Participation in the Employee Stock Purchase Plan Benefits package (health insurance, multisport, shopping vouchers) Strategically located offices featuring entertainment and relaxation zones, table tennis and football, free snacks, fantastic coffee, and more Referral bonuses Corporate, social and well-being events Please, note: The set of bonuses might vary based on the role you apply for – specifics will be discussed with our recruiter during the general interview. We will reach out to selected candidates exclusively. EPAM is a leading global provider of digital platform engineering and development services. We are committed to having a positive impact on our customers, our employees, and our communities. We embrace a dynamic and inclusive culture. Here you will collaborate with multi-national teams, contribute to a myriad of innovative projects that deliver the most creative and cutting-edge solutions, and have an opportunity to continuously learn and grow. No matter where you are located, you will join a dedicated, creative, and diverse community that will help you discover your fullest potential.

Technology

EPAM Systems

Lead C++ Developer

Senior

Remote

Wroclaw, DS, Poland

🏢 Summary: Lead C++ Developer role focused on designing and optimizing high-performance ML kernels for TPU and GPU architectures within a CoreML team. The position drives performance improvements across training and inference, builds developer infrastructure, and collaborates with ML researchers and compiler engineers. It offers exposure to cutting-edge hardware, ML models, and toolchains impacting large-scale AI systems. 🗂️ Requirements: Bachelor’s degree or equivalent practical experience, 7+ years of industry experience, 5+ years of software development in C++ or Python, 3+ years of experience testing, maintaining or launching software products, 1+ year of software design and architecture experience, Expertise in kernel-level performance optimization, English proficiency at B2 level or higher 📃 Skills: C++, Python, TPU, GPU, Pallas, Mosaic, Triton, CUDA, JAX, PyTorch, XLA, MLIR, OpenXLA, Kernel, Benchmarking, Autotuning, Compilers 🏢 Description: We are seeking a talented Lead C++ Developer to join the highly interdisciplinary CoreML team, where you will drive the performance and optimization of both training and serving, delivering massive impact for customers. In this role, you will have exposure to the newest Tensor Processing Unit (TPU), Graphics Processing Unit (GPU) hardware, the latest ML models, and the advanced toolchains that bridge them. Your work will directly enable AI research, production deployments and the broader open-source ecosystem, addressing complex technical issues that directly impact the efficiency and scalability of AI across the industry. Responsibilities Design and optimize high-performance kernels (using languages like Pallas, Mosaic and Triton) targeting Tensor Processing Unit (TPU) and Graphics Processing Unit (GPU) architectures for critical Machine Learning (ML) operations, redefining what's possible from massive training runs to high-speed inference Architecture of infrastructure such as benchmarking suites, autotuning frameworks, performance analysis tools, regression testing and documentation, transforming how the developer community interacts with increasingly critical custom kernels in key Open-Source Software (OSS) libraries Track the latest advancements in hardware architectures, compiler technologies and AI models to identify new opportunities for performance optimization through custom kernels Engagement with ML researchers, framework developers (Just After eXecution (JAX), PyTorch) and compiler engineers (Accelerated Linear Algebra (XLA)) to enhance adoption, identify new requirements and address bottlenecks by providing appropriate solutions Requirements Bachelor's degree or equivalent practical experience Overall 7+ years of industry experience 5 years of experience with software development in C++ or Python 3 years of experience testing, maintaining or launching software products, and 1 year of experience with software design and architecture Expertise in performance optimization at the kernel level English proficiency at B2 level or higher Nice to have Skills in optimizing TPU/GPU code, using low-level kernel languages like Pallas, Compute Unified Device Architecture (CUDA) or Triton Knowledge of ML Frameworks (JAX/PyTorch), common operations like attention and Mixture of Experts (MoEs), including model optimization and low-precision formats Understanding of modern accelerators (e.g., data movement, pipelining, heterogeneous compute and scale-out) Understanding of compiler principles (optimization, code generation) and toolchains such as MLIR and OpenXLA Showcase of building developer infrastructure, including Open-Source Software (OSS) libraries, flexible high-performance APIs and easy-to-consume documentation to empower the community We offer We gather like-minded people: Engineering community of industry professionals Friendly team and enjoyable working environment Flexible schedule and opportunity to work remotely within Poland Chance to work abroad for up to 60 days annually Business-driven relocation opportunities We provide growth opportunities: Outstanding career roadmap Leadership development, career advising, soft skills, and well-being programs Certification (GCP, Azure, AWS) Unlimited access to LinkedIn Learning, Get Abstract, Cloud Guru English classes We cover it all: Stable income (Employment Contract or B2B) Participation in the Employee Stock Purchase Plan Benefits package (health insurance, multisport, shopping vouchers) Strategically located offices featuring entertainment and relaxation zones, table tennis and football, free snacks, fantastic coffee, and more Referral bonuses Corporate, social and well-being events Please, note: The set of bonuses might vary based on the role you apply for – specifics will be discussed with our recruiter during the general interview. We will reach out to selected candidates exclusively. EPAM is a leading global provider of digital platform engineering and development services. We are committed to having a positive impact on our customers, our employees, and our communities. We embrace a dynamic and inclusive culture. Here you will collaborate with multi-national teams, contribute to a myriad of innovative projects that deliver the most creative and cutting-edge solutions, and have an opportunity to continuously learn and grow. No matter where you are located, you will join a dedicated, creative, and diverse community that will help you discover your fullest potential.

Technology

EPAM Systems

Lead C++ Developer

Senior

Remote

Gdansk, PM, Poland

🏢 Summary: Lead C++ Developer role focused on building and optimizing high-performance ML kernels for TPU and GPU architectures within a CoreML team. The position drives performance improvements across training and inference, collaborating with ML researchers and compiler engineers to enhance scalability and efficiency. You will design kernel-level optimizations, benchmarking infrastructure and tooling that power large-scale AI workloads. 🗂️ Requirements: Bachelor’s degree or equivalent practical experience, 7+ years of industry experience, 5+ years of software development in C++ or Python, 3+ years of experience testing, maintaining or launching software products, 1+ year of software design and architecture experience, Expertise in kernel-level performance optimization, English proficiency at B2 level or higher 📃 Skills: C++, Python, TPU, GPU, Pallas, Mosaic, Triton, CUDA, JAX, PyTorch, XLA, MLIR, OpenXLA, Benchmarking, Autotuning, Regression, Compilers 🏢 Description: We are seeking a talented Lead C++ Developer to join the highly interdisciplinary CoreML team, where you will drive the performance and optimization of both training and serving, delivering massive impact for customers. In this role, you will have exposure to the newest Tensor Processing Unit (TPU), Graphics Processing Unit (GPU) hardware, the latest ML models, and the advanced toolchains that bridge them. Your work will directly enable AI research, production deployments and the broader open-source ecosystem, addressing complex technical issues that directly impact the efficiency and scalability of AI across the industry. Responsibilities Design and optimize high-performance kernels (using languages like Pallas, Mosaic and Triton) targeting Tensor Processing Unit (TPU) and Graphics Processing Unit (GPU) architectures for critical Machine Learning (ML) operations, redefining what's possible from massive training runs to high-speed inference Architecture of infrastructure such as benchmarking suites, autotuning frameworks, performance analysis tools, regression testing and documentation, transforming how the developer community interacts with increasingly critical custom kernels in key Open-Source Software (OSS) libraries Track the latest advancements in hardware architectures, compiler technologies and AI models to identify new opportunities for performance optimization through custom kernels Engagement with ML researchers, framework developers (Just After eXecution (JAX), PyTorch) and compiler engineers (Accelerated Linear Algebra (XLA)) to enhance adoption, identify new requirements and address bottlenecks by providing appropriate solutions Requirements Bachelor's degree or equivalent practical experience Overall 7+ years of industry experience 5 years of experience with software development in C++ or Python 3 years of experience testing, maintaining or launching software products, and 1 year of experience with software design and architecture Expertise in performance optimization at the kernel level English proficiency at B2 level or higher Nice to have Skills in optimizing TPU/GPU code, using low-level kernel languages like Pallas, Compute Unified Device Architecture (CUDA) or Triton Knowledge of ML Frameworks (JAX/PyTorch), common operations like attention and Mixture of Experts (MoEs), including model optimization and low-precision formats Understanding of modern accelerators (e.g., data movement, pipelining, heterogeneous compute and scale-out) Understanding of compiler principles (optimization, code generation) and toolchains such as MLIR and OpenXLA Showcase of building developer infrastructure, including Open-Source Software (OSS) libraries, flexible high-performance APIs and easy-to-consume documentation to empower the community We offer We gather like-minded people: Engineering community of industry professionals Friendly team and enjoyable working environment Flexible schedule and opportunity to work remotely within Poland Chance to work abroad for up to 60 days annually Business-driven relocation opportunities We provide growth opportunities: Outstanding career roadmap Leadership development, career advising, soft skills, and well-being programs Certification (GCP, Azure, AWS) Unlimited access to LinkedIn Learning, Get Abstract, Cloud Guru English classes We cover it all: Stable income (Employment Contract or B2B) Participation in the Employee Stock Purchase Plan Benefits package (health insurance, multisport, shopping vouchers) Strategically located offices featuring entertainment and relaxation zones, table tennis and football, free snacks, fantastic coffee, and more Referral bonuses Corporate, social and well-being events Please, note: The set of bonuses might vary based on the role you apply for – specifics will be discussed with our recruiter during the general interview. We will reach out to selected candidates exclusively. EPAM is a leading global provider of digital platform engineering and development services. We are committed to having a positive impact on our customers, our employees, and our communities. We embrace a dynamic and inclusive culture. Here you will collaborate with multi-national teams, contribute to a myriad of innovative projects that deliver the most creative and cutting-edge solutions, and have an opportunity to continuously learn and grow. No matter where you are located, you will join a dedicated, creative, and diverse community that will help you discover your fullest potential.

Technology

EPAM Systems

Senior C++ Software Engineer with CUDA/GPU/TPU

Senior

Remote

Wroclaw, Poland

🏢 Summary: Senior C++ Software Engineer role focused on building AI and ML compiler and neural network stacks for next-generation computing systems. The position involves developing and optimizing C++ and CUDA kernels for LLMs, CNNs and transformer architectures targeting GPU and AI accelerators. Candidates will collaborate with hardware and AI experts while contributing to high-performance ML infrastructure and models. 🗂️ Requirements: 5+ years of software engineering experience focused on ML/AI, 5+ years of experience with C++, 5+ years of experience with CUDA and kernel programming, Knowledge of GPU and AI-accelerator architectures, Experience building ML models with PyTorch or TensorFlow, Understanding of transformer and modern ML architectures, English proficiency at B2 level or higher 📃 Skills: C++, CUDA, PyTorch, TensorFlow, Python, GPU, ASIC, LLM, CNN, Transformers 🏢 Description: We are looking for a Senior C++ Software Engineer to join a team building the next generation of computers for artificial intelligence. The company brings together experts in computer architecture, ASIC design, advanced systems and neural network compilers, with the team focused on the compiler and neural network stack, including LLMs and CNNs. This is a key developer role within the ML & AI team. Responsibilities Write kernels in C++ for various operations used across AI and ML workloads Develop and optimize the compiler and neural network stack supporting LLMs and CNNs Collaborate with experts in computer architecture, ASIC design and advanced systems to build next-generation computing solutions Design and implement AI/ML kernel operations tailored to GPU/AI-accelerator architectures Build and maintain ML models using PyTorch or TensorFlow Ensure kernel-level performance and correctness across modern neural network architectures such as transformers Take ownership of key technical decisions as a lead developer within the ML & AI team Requirements 5+ years of experience in software engineering with a focus on ML/AI 5+ years of experience in C++ and CUDA/kernel programming Knowledge in GPU/AI-accelerator architectures Experience in building and working with ML models in PyTorch or TensorFlow Understanding of modern ML model architectures such as transformers English proficiency at B2 level or higher Nice to have Proficiency in Python programming with hands-on experience in developing, training and fine-tuning deep learning models Skills in designing, implementing and iterating on neural network architectures to achieve optimal performance on diverse tasks Capability to investigate and troubleshoot model performance issues including gaps in compilation steps and kernels Demonstrated experience in designing, training and deploying neural networks for various applications Solid understanding of machine learning fundamentals including supervised and unsupervised learning techniques We offer We gather like-minded people: Top tech minds driving innovation in AI, cloud and digital platform modernization Supportive team and agile, startup-like culture Hybrid by design mode and opportunity to work remotely within Poland Chance to work abroad for up to 60 days annually Business-driven relocation opportunities We provide growth opportunities: Career development programs Thought leadership, mentoring, soft skills and well-being programs Certification (Anthropic, Gemini, GCP, Azure, AWS) English classes We cover it all: Stable pay Participation in the Employee Stock Purchase Plan with a 15% discount Benefits package (health insurance, multisport, shopping vouchers) Referral bonuses up to $2,000 Offices featuring entertainment and relaxation zones, table tennis and football, free snacks, coffee and more Corporate, social and well-being events Please, note: Benefits listed above are available to employees only We are open for working with Contractors. Terms of B2B cooperation agreements are agreed individually We will reach out to selected candidates exclusively EPAM is a leading global provider of digital platform engineering and development services. We are committed to having a positive impact on our customers, our employees, and our communities. We embrace a dynamic and inclusive culture. Here you will collaborate with multi-national teams, contribute to a myriad of innovative projects that deliver the most creative and cutting-edge solutions, and have an opportunity to continuously learn and grow. No matter where you are located, you will join a dedicated, creative, and diverse community that will help you discover your fullest potential.

Technology

EPAM Systems

Senior C++ Software Engineer with CUDA/GPU/TPU

Senior

Remote

Lodz, Poland

🏢 Summary: Senior C++ Software Engineer role focused on developing and optimizing AI/ML compiler and neural network stacks for next-generation AI computing systems. The position involves C++ and CUDA kernel development for LLMs, CNNs and transformer architectures, collaborating with experts in GPU and AI-accelerator technologies. The offer includes hybrid or remote work flexibility, professional development programs and comprehensive employee benefits. 🗂️ Requirements: 5+ years of software engineering experience focused on ML/AI, 5+ years of experience with C++, 5+ years of experience with CUDA and kernel programming, Knowledge of GPU and AI-accelerator architectures, Experience with PyTorch or TensorFlow, Understanding of transformer and modern ML model architectures, English proficiency at B2 level or higher 📃 Skills: C++, CUDA, PyTorch, TensorFlow, Python, GPU, ASIC, LLM, CNN, Transformers 🏢 Description: We are looking for a Senior C++ Software Engineer to join a team building the next generation of computers for artificial intelligence. The company brings together experts in computer architecture, ASIC design, advanced systems and neural network compilers, with the team focused on the compiler and neural network stack, including LLMs and CNNs. This is a key developer role within the ML & AI team. Responsibilities Write kernels in C++ for various operations used across AI and ML workloads Develop and optimize the compiler and neural network stack supporting LLMs and CNNs Collaborate with experts in computer architecture, ASIC design and advanced systems to build next-generation computing solutions Design and implement AI/ML kernel operations tailored to GPU/AI-accelerator architectures Build and maintain ML models using PyTorch or TensorFlow Ensure kernel-level performance and correctness across modern neural network architectures such as transformers Take ownership of key technical decisions as a lead developer within the ML & AI team Requirements 5+ years of experience in software engineering with a focus on ML/AI 5+ years of experience in C++ and CUDA/kernel programming Knowledge in GPU/AI-accelerator architectures Experience in building and working with ML models in PyTorch or TensorFlow Understanding of modern ML model architectures such as transformers English proficiency at B2 level or higher Nice to have Proficiency in Python programming with hands-on experience in developing, training and fine-tuning deep learning models Skills in designing, implementing and iterating on neural network architectures to achieve optimal performance on diverse tasks Capability to investigate and troubleshoot model performance issues including gaps in compilation steps and kernels Demonstrated experience in designing, training and deploying neural networks for various applications Solid understanding of machine learning fundamentals including supervised and unsupervised learning techniques We offer We gather like-minded people: Top tech minds driving innovation in AI, cloud and digital platform modernization Supportive team and agile, startup-like culture Hybrid by design mode and opportunity to work remotely within Poland Chance to work abroad for up to 60 days annually Business-driven relocation opportunities We provide growth opportunities: Career development programs Thought leadership, mentoring, soft skills and well-being programs Certification (Anthropic, Gemini, GCP, Azure, AWS) English classes We cover it all: Stable pay Participation in the Employee Stock Purchase Plan with a 15% discount Benefits package (health insurance, multisport, shopping vouchers) Referral bonuses up to $2,000 Offices featuring entertainment and relaxation zones, table tennis and football, free snacks, coffee and more Corporate, social and well-being events Please, note: Benefits listed above are available to employees only We are open for working with Contractors. Terms of B2B cooperation agreements are agreed individually We will reach out to selected candidates exclusively EPAM is a leading global provider of digital platform engineering and development services. We are committed to having a positive impact on our customers, our employees, and our communities. We embrace a dynamic and inclusive culture. Here you will collaborate with multi-national teams, contribute to a myriad of innovative projects that deliver the most creative and cutting-edge solutions, and have an opportunity to continuously learn and grow. No matter where you are located, you will join a dedicated, creative, and diverse community that will help you discover your fullest potential.

Technology

EPAM Systems

Senior Software Engineer

Senior

Remote

🏢 Summary: Fully remote Senior Software Developer role focused on low-level C++ development and kernel optimization for machine learning and HPC applications. The position involves designing and optimizing tensor compute and data movement kernels, improving performance-critical code, and integrating optimized components into ML pipelines. The role requires strong expertise in performance profiling, debugging, and low-level software optimization. 🗂️ Requirements: 3+ years of experience in low-level programming and optimization, Proficiency in C/C++ with focus on kernel optimization, Experience in kernel development for machine learning or HPC applications, Expertise in tensor compute and data movement optimization, Experience with performance profiling and debugging tools, Experience in kernel-level software testing and debugging, Understanding of machine learning pipelines, Ability to identify and resolve performance bottlenecks, English proficiency at B2 level or higher 📃 Skills: C++, C, AI, ML, HPC, CUDA, OpenCL, Tensor, Profiling, Debugging 🏢 Description: We are seeking a highly skilled and hands-on Senior Software Developer with expertise in C++, AI and ML . It is a fully remote position offering you the flexibility to work from any location in Poland, whether it's your home or one of our well-equipped offices in Gdansk, Katowice, Krakow, Lodz, Warsaw, or Wroclaw. Responsibilities Develop and optimize low-level workloads and kernels to enhance the performance of software for machine learning applications Implement tensor compute and tensor data movement optimization kernels Design, develop, and maintain kernel-level software components for the client’s machine learning and HPC applications Perform in-depth analysis and optimization of low-level code, focusing on improving tensor optimization and efficiency Collaborate with machine learning engineers to integrate optimized kernels and routines into frameworks and pipelines Identify performance bottlenecks through profiling and apply strategies to resolve inefficiencies Conduct comprehensive testing, unit test development, and debugging to ensure the stability and reliability of kernel-level code Communicate and problem-solve effectively to analyze and debug complex software issues Leverage industry tools for performance profiling and optimization Requirements 3+ years of experience in relevant roles involving low-level programming and optimization Proficiency in C/C++ and low-level programming with a focus on kernel optimization Background in kernel development, including the implementation of efficient kernels and libraries for machine learning and HPC Expertise in low-level optimization techniques, specifically in tensor compute and data movement Experience with performance profiling, debugging tools, and strategies for software optimization Proven skills in kernel-level software testing, debugging, and development, ensuring efficiency and stability Understanding of machine learning pipelines and collaboration between software engineers and data scientists Strong problem-solving skills and the ability to resolve bottlenecks in performance-critical environments B2 level of English or higher, with an emphasis on technical communication skills Nice to have Familiarity with machine learning frameworks and an understanding of related concepts Knowledge of operating system internals Understanding and experience with GPU programming such as CUDA or OpenCL We offer/Benefits We gather like-minded people: Engineering community of industry professionals Friendly team and enjoyable working environment Flexible schedule and opportunity to work remotely within Poland Chance to work abroad for up to 60 days annually Business-driven relocation opportunities We provide growth opportunities: Outstanding career roadmap Leadership development, career advising, soft skills, and well-being programs Certification (GCP, Azure, AWS) Unlimited access to LinkedIn Learning, Get Abstract, Cloud Guru English classes We cover it all: Stable income (Employment Contract or B2B) Participation in the Employee Stock Purchase Plan Benefits package (health insurance, multisport, shopping vouchers) Strategically located offices featuring entertainment and relaxation zones, table tennis and football, free snacks, fantastic coffee, and more Referral bonuses Corporate, social and well-being events Please, note: The set of bonuses might vary based on the role you apply for – specifics will be discussed with our recruiter during the general interview. We will reach out to selected candidates exclusively. EPAM is a leading global provider of digital platform engineering and development services. We are committed to having a positive impact on our customers, our employees, and our communities. We embrace a dynamic and inclusive culture. Here you will collaborate with multi-national teams, contribute to a myriad of innovative projects that deliver the most creative and cutting-edge solutions, and have an opportunity to continuously learn and grow. No matter where you are located, you will join a dedicated, creative, and diverse community that will help you discover your fullest potential.

Technology

EPAM Systems

Senior C++ Software Engineer

Senior

Remote

🏢 Summary: Fully remote Senior Software Developer role focused on developing and optimizing low-level C++ kernels for machine learning and HPC applications. The position involves tensor compute optimization, performance profiling, and integrating optimized components into ML pipelines. The role emphasizes kernel-level development, debugging, and efficiency improvements in performance-critical environments. 🗂️ Requirements: 3+ years in low-level programming and optimization roles, Proficiency in C/C++, Experience in kernel development for ML or HPC, Expertise in tensor compute and data movement optimization, Experience with performance profiling and debugging tools, Experience in kernel-level testing and optimization, Understanding of machine learning pipelines, Experience identifying and resolving performance bottlenecks 📃 Skills: C++, C, AI, ML, HPC, Kernels, Tensor, Profiling, Debugging, CUDA, OpenCL 🏢 Description: We are seeking a highly skilled and hands-on Senior Software Developer with expertise in C++, AI and ML . It is a fully remote position offering you the flexibility to work from any location in Poland, whether it's your home or one of our well-equipped offices in Gdansk, Katowice, Krakow, Lodz, Warsaw, or Wroclaw. Responsibilities Develop and optimize low-level workloads and kernels to enhance the performance of software for machine learning applications Implement tensor compute and tensor data movement optimization kernels Design, develop, and maintain kernel-level software components for the client’s machine learning and HPC applications Perform in-depth analysis and optimization of low-level code, focusing on improving tensor optimization and efficiency Collaborate with machine learning engineers to integrate optimized kernels and routines into frameworks and pipelines Identify performance bottlenecks through profiling and apply strategies to resolve inefficiencies Conduct comprehensive testing, unit test development, and debugging to ensure the stability and reliability of kernel-level code Communicate and problem-solve effectively to analyze and debug complex software issues Leverage industry tools for performance profiling and optimization Requirements 3+ years of experience in relevant roles involving low-level programming and optimization Proficiency in C/C++ and low-level programming with a focus on kernel optimization Background in kernel development, including the implementation of efficient kernels and libraries for machine learning and HPC Expertise in low-level optimization techniques, specifically in tensor compute and data movement Experience with performance profiling, debugging tools, and strategies for software optimization Proven skills in kernel-level software testing, debugging, and development, ensuring efficiency and stability Understanding of machine learning pipelines and collaboration between software engineers and data scientists Strong problem-solving skills and the ability to resolve bottlenecks in performance-critical environments B2 level of English or higher, with an emphasis on technical communication skills Nice to have Familiarity with machine learning frameworks and an understanding of related concepts Knowledge of operating system internals Understanding and experience with GPU programming such as CUDA or OpenCL We offer/Benefits We gather like-minded people: Engineering community of industry professionals Friendly team and enjoyable working environment Flexible schedule and opportunity to work remotely within Poland Chance to work abroad for up to 60 days annually Business-driven relocation opportunities We provide growth opportunities: Outstanding career roadmap Leadership development, career advising, soft skills, and well-being programs Certification (GCP, Azure, AWS) Unlimited access to LinkedIn Learning, Get Abstract, Cloud Guru English classes We cover it all: Stable income (Employment Contract or B2B) Participation in the Employee Stock Purchase Plan Benefits package (health insurance, multisport, shopping vouchers) Strategically located offices featuring entertainment and relaxation zones, table tennis and football, free snacks, fantastic coffee, and more Referral bonuses Corporate, social and well-being events Please, note: The set of bonuses might vary based on the role you apply for – specifics will be discussed with our recruiter during the general interview. We will reach out to selected candidates exclusively. EPAM is a leading global provider of digital platform engineering and development services. We are committed to having a positive impact on our customers, our employees, and our communities. We embrace a dynamic and inclusive culture. Here you will collaborate with multi-national teams, contribute to a myriad of innovative projects that deliver the most creative and cutting-edge solutions, and have an opportunity to continuously learn and grow. No matter where you are located, you will join a dedicated, creative, and diverse community that will help you discover your fullest potential.