Unlock Full Resume Report

New offer - be the first one to apply!

September 17, 2026

Data Engineer / Data Scientist (AI / NLP)

Mid • On-site

374,400 - 457,600 PLN/yr

Warsaw, Poland

Quick Facts

  • Contract: B2B

  • Duration: until 30 September 2026 (possible extensions)

  • Start date: ASAP

  • Rate: up to 220 PLN/h (net) + VAT

  • Work model: hybrid (3 days onsite from Warsaw or Krakow, 2 days remote)

  • Onboarding: 2 weeks in Malmö (fully covered)

  • FTE: full-time (part-time can be discussed)

Description

You will work on an AI / Data Engineering project focused on building and improving AI-driven data solutions for large-scale web content processing, attribute extraction, and market expansion. The role combines data engineering, machine learning, and applied AI: you will design, build, and evaluate data pipelines, fine-tune lightweight ML models, and support internal AI research agents used across geographic markets and data domains.

Responsibilities

  • Build and optimize Spark pipelines for large-scale web content ingestion and processing

  • Use Python (including Polars and/or Pandas) for data processing, analysis, and pipeline development

  • Fine-tune lightweight ML models for task-specific attribute extraction

  • Prepare training data, manage data quality, and evaluate model performance end-to-end

  • Apply NLP techniques to extract, classify, and reason over information from web content

  • Expand an internal AI research agent to new geographic markets and adapt logic to local data conditions

  • Support evidence collection and reasoning logic for new place-related attributes

  • Evaluate ML systems across locales, domains, and data sources

  • Work on pipeline orchestration, optimization, and multi-source ingestion processes

  • Potentially use Scala and Spark in data engineering tracks

Requirements

  • Strong Python skills, especially with Polars and/or Pandas

  • Experience with NLP and fine-tuning lightweight ML models

  • Practical experience designing, building, and evaluating data pipelines

  • Experience with Spark and ideally Scala

  • Familiarity with agent frameworks, especially LangGraph

  • Understanding of data quality, model evaluation, and performance measurement

  • Ability to adapt ML/data solutions to different countries, languages, and data domains

  • Experience with pipeline orchestration and optimization for large-scale data ingestion

  • Hands-on, problem-solving mindset and ability to work in a fast-moving environment

Benefits

  • Hybrid work from Warsaw or Krakow (3 days onsite, 2 days remote)

  • Two-week onboarding in Malmö (fully covered)

  • Rate up to 220 PLN/h (net) + VAT

  • Full-time B2B contract until 30 September 2026 (possible extensions)

Similar jobs you might like