Data Engineer (PySpark)
120 - 150 zł / stawka godzinowaYOUR ITEAMS sp. z o.o.
Expected, PySpark, Apache Spark, AWS, SQL
Operating system, Windows About the project, We are seeking a Regular Data Engineer (PySpark) to join a team working with production ETL pipelines on AWS. The person in this role will apply strong PySpark skills within an existing architecture, independently deliver tasks, and diagnose common production issues. This position focuses on practical implementation and operational excellence rather than end-to-end architecture design. This is how we work, at the client's site, you focus on a single project at a time, you focus on product development Your responsibilities, Implement data transformations and ETL logic using PySpark., Maintain and develop existing AWS Glue jobs and related pipelines., Investigate and resolve pipeline failures using logs, error messages, and basic metrics., Diagnose and mitigate Spark performance issues (partitioning, shuffle, skew, joins)., Choose appropriate partitioning and join strategies for given workloads., Handle error scenarios in ETL processes, including retries, reprocessing, and partial failures to avoid duplicate processing., Work with large volumes of data and many files without creating uncontrolled parallelism., Follow established architectural patterns and implement solutions that integrate with the current system design., Support data migration or batch processing tasks as needed. 2–4 years of experience as a Data Engineer or in a similar role., Practical, hands-on experience with PySpark / Apache Spark., Solid understanding of key Spark concepts, including partitioning, shuffle, repartition vs coalesce, broadcast joins, and data skew impact., Practical experience with AWS services, especially AWS Glue, Amazon S3, and Amazon CloudWatch., Experience building, maintaining, or enhancing ETL/data processing pipelines., Ability to diagnose common Glue and PySpark issues from logs and basic metrics., Experience handling ETL error scenarios: retries, reprocessing, partial failures, and preventing duplicate processing., Basic understanding of idempotency and safe restart strategies after failures., Experience processing large data volumes or many files and familiarity with batch migration processes., Strong SQL skills and familiarity with joins, aggregations, filtering, grouping, and ranking. What we offer, Remote working., Unique TEAL culture, relationship- and respect-driven community, non-corporate atmosphere., Agile approach and no bureaucracy., Outstanding integration trips to various places in Europe., Activities to support your well-being and health., Luxmed Gold Extended medical care and Multisport Plus benefit. This is how we work on a project, team-level deployment, documentation, issue tracking tools, testing environments Benefits, sharing the costs of sports activities, private medical care, remote work opportunities, integration events Recruitment stages, Screening Call, Cultural Fit Interview, Technical Interview, Final Round Interview We encourage qualified candidates to apply via the application form. Applicants will be considered on the basis of their skills and experience. YOUR ITEAMS sp. z o.o., Unique TEAL culture, relationship- and respect-driven community, non-corporate atmosphere., Agile approach and no bureaucracy. This is how we work,Oferta pracy dodana 2 dni temu
Powiązane wyszukiwania