Data Engineer (PySpark)
Offer summary

(Summary generated by AI based on the full job description)

The project involves production ETL on AWS using PySpark, AWS Glue, Amazon S3, and CloudWatch. The main responsibilities include implementing and maintaining ETL pipelines, diagnosing and resolving production issues, and optimizing Spark performance. Practical knowledge of Spark, SQL and experience with ETL, data migration, and error handling are required. The role offers remote work and benefits such as Luxmed Gold medical care and Multisport Plus.

you can start ASAP

Data Engineer (PySpark)

Company: YOUR ITEAMS sp. z o.o.

from: 27 August 2026
to: 26 September 2026
120 - 150net (+ VAT)/ hr.B2B contract (full-time)
Offer parameters
level:mid
working mode:remote
location:Warszawa, Masovian
Warszawa, Masovian

Requirements

Expected technologies

PySpark
Apache Spark
AWS
SQL

Operating system

Windows

Our requirements

  • 2–4 years of experience as a Data Engineer or in a similar role.
  • Practical, hands-on experience with PySpark / Apache Spark.
  • Solid understanding of key Spark concepts, including partitioning, shuffle, repartition vs coalesce, broadcast joins, and data skew impact.
  • Practical experience with AWS services, especially AWS Glue, Amazon S3, and Amazon CloudWatch.
  • Experience building, maintaining, or enhancing ETL/data processing pipelines.
  • Ability to diagnose common Glue and PySpark issues from logs and basic metrics.
  • Experience handling ETL error scenarios: retries, reprocessing, partial failures, and preventing duplicate processing.
  • Basic understanding of idempotency and safe restart strategies after failures.
  • Experience processing large data volumes or many files and familiarity with batch migration processes.
  • Strong SQL skills and familiarity with joins, aggregations, filtering, grouping, and ranking.

Your responsibilities

  • Implement data transformations and ETL logic using PySpark.
  • Maintain and develop existing AWS Glue jobs and related pipelines.
  • Investigate and resolve pipeline failures using logs, error messages, and basic metrics.
  • Diagnose and mitigate Spark performance issues (partitioning, shuffle, skew, joins).
  • Choose appropriate partitioning and join strategies for given workloads.
  • Handle error scenarios in ETL processes, including retries, reprocessing, and partial failures to avoid duplicate processing.
  • Work with large volumes of data and many files without creating uncontrolled parallelism.
  • Follow established architectural patterns and implement solutions that integrate with the current system design.
  • Support data migration or batch processing tasks as needed.

About the project

We are seeking a Regular Data Engineer (PySpark) to join a team working with production ETL pipelines on AWS. The person in this role will apply strong PySpark skills within an existing architecture, independently deliver tasks, and diagnose common production issues. This position focuses on practical implementation and operational excellence rather than end-to-end architecture design.

This is how we organize our work

This is how we work

at the client's siteyou focus on a single project at a timeyou focus on product development

This is how we work on a project

  • team-level deployment
  • documentation
  • issue tracking tools
  • testing environments
We encourage qualified candidates to apply via the application form. Applicants will be considered on the basis of their skills and experience.
Company

What we offer

  • Remote working.
  • Unique TEAL culture, relationship- and respect-driven community, non-corporate atmosphere.
  • Agile approach and no bureaucracy.
  • Outstanding integration trips to various places in Europe.
  • Activities to support your well-being and health.
  • Luxmed Gold Extended medical care and Multisport Plus benefit.

Benefits

  • sharing the costs of sports activities
  • private medical care
  • remote work opportunities
  • integration events

Recruitment stages

  • 1.
    Screening Call
  • 2.
    Cultural Fit Interview
  • 3.
    Technical Interview
  • 4.
    Final Round Interview

YOUR ITEAMS sp. z o.o.

Unique TEAL culture, relationship- and respect-driven community, non-corporate atmosphere.
Agile approach and no bureaucracy.

This is how we work

Data Engineer (PySpark)
120–150 zł / hr. (B2B)
I apply to:
YOUR ITEAMS sp. z o.o.
Warszawa, Masovian
Pracodawca zbiera zgłoszenia przez swój system.
Przejdziesz na zewnętrzny formularz.

By clicking "Aplikuj" you confirm that you've read and accepted our Terms and Conditions.



This site is protected by reCAPTCHA and the Google Privacy Policy and Terms of Service apply.

Need more information?

  • Make sure the body of the offer doesn’t already include what you’re looking for.
  • Ask a question if you need more information you’re interested in.
  • We’ll forward your question to the employer and aim to provide a response within 3 business days.

Share this offer