Unlock Full Resume Report
ATS Pass
Missing keywords
Tailored AI suggestions
Job match analysis
Interview-focused insights
New offer - be the first one to apply!
September 17, 2026
Data Engineer / Data Scientist (AI / NLP)
Mid • On-site
374,400 - 457,600 PLN/yr
Warsaw, Poland
Apply now
Quick Facts
Contract: B2B
Duration: until 30 September 2026 (possible extensions)
Start date: ASAP
Rate: up to 220 PLN/h (net) + VAT
Work model: hybrid (3 days onsite from Warsaw or Krakow, 2 days remote)
Onboarding: 2 weeks in Malmö (fully covered)
FTE: full-time (part-time can be discussed)
Description
You will work on an AI / Data Engineering project focused on building and improving AI-driven data solutions for large-scale web content processing, attribute extraction, and market expansion. The role combines data engineering, machine learning, and applied AI: you will design, build, and evaluate data pipelines, fine-tune lightweight ML models, and support internal AI research agents used across geographic markets and data domains.
Responsibilities
Build and optimize Spark pipelines for large-scale web content ingestion and processing
Use Python (including Polars and/or Pandas) for data processing, analysis, and pipeline development
Fine-tune lightweight ML models for task-specific attribute extraction
Prepare training data, manage data quality, and evaluate model performance end-to-end
Apply NLP techniques to extract, classify, and reason over information from web content
Expand an internal AI research agent to new geographic markets and adapt logic to local data conditions
Support evidence collection and reasoning logic for new place-related attributes
Evaluate ML systems across locales, domains, and data sources
Work on pipeline orchestration, optimization, and multi-source ingestion processes
Potentially use Scala and Spark in data engineering tracks
Requirements
Strong Python skills, especially with Polars and/or Pandas
Experience with NLP and fine-tuning lightweight ML models
Practical experience designing, building, and evaluating data pipelines
Experience with Spark and ideally Scala
Familiarity with agent frameworks, especially LangGraph
Understanding of data quality, model evaluation, and performance measurement
Ability to adapt ML/data solutions to different countries, languages, and data domains
Experience with pipeline orchestration and optimization for large-scale data ingestion
Hands-on, problem-solving mindset and ability to work in a fast-moving environment
Benefits
Hybrid work from Warsaw or Krakow (3 days onsite, 2 days remote)
Two-week onboarding in Malmö (fully covered)
Rate up to 220 PLN/h (net) + VAT
Full-time B2B contract until 30 September 2026 (possible extensions)
Similar jobs you might like

Data Engineer
Awareson Sp. z o.o.
Senior
Technology
Warsaw, MZ, Poland · Remote
2080K zł - 2746K zł/yr
11 days ago

Data Engineer (Databricks)
Cyclad
Mid
Technology
Warsaw, Poland · On-site
291K zł - 318K zł/yr
28m ago

ML Engineer
emagine Polska
Mid
Technology
Warsaw, MZ, Poland · On-site
333K zł - 458K zł/yr
12 days ago

Data & Cloud Engineer (CAD/GIS)
dmTECH Polska
Mid
Technology
Wrocław, DS, Poland · Remote
291K zł - 374K zł/yr
1 days ago
Data Engineer (AI & Big Data, ETL/Integration Consultant)
CloverDX Labs s.r.o.
Junior
Technology
Brno, Czechia · Remote
60K Kč - 120K Kč/yr
21h ago

Data Engineer
Connectis
Senior
Technology
Warsaw, Poland · Remote
24K zł - 27K zł/yr
28m ago
Data Engineer
McGregor Boyall
Senior
Technology
Krakow, MA, Poland · On-site
N/A
14 days ago

Senior Data Engineer (Java/Scala, Spark, AWS)
EPAM Systems
Senior
Technology
Warsaw, Poland · On-site
N/A
12 days ago
Data Engineer (Mid to Senior)
DataSentics a.s.
Mid
Technology
Prague, Czechia · Hybrid
60K Kč - 120K Kč/yr
4h ago

IT Developer LangChain and RAG
emagine Polska
Mid
Technology
Gdansk, PM, Poland · On-site
$322K - $322K/yr
14 days ago