Utwórz profil, aby pracodawcy mogli Cię znaleźć, otrzymywać lepiej dopasowane oferty pracy i szybciej aplikować.
  • Wyszukiwanie ofert pracy
  • Zapisane
  • Stwórz CV
    Nowe
  • Wynagrodzenia
  • Subskrypcje

Senior Data Engineer

160 - 220 zł / stawka godzinowa

KUBO

Senior Data Engineer

Miejsce pracy: Gdańsk

Technologies we use

Expected

  • Databricks
  • Apache Spark
  • Python
  • PyTorch
  • Microsoft Azure

Optional

  • Pytest

About the project

We are looking for a skilled Data Engineer to join an international technology and data team on a 6-month engagement, starting in the first or second week of October.

This is a hands-on role focused on building data solutions for the insurance domain. You will use Databricks AI capabilities and Apache Spark to extract, process and structure information from both typed and handwritten insurance documents. A key part of the role will be transforming and modelling this data so it can be reliably used for machine learning and analytics use cases. You will be working 100% remotely in international team.

This is how we organize our work

This is how we work

  • you focus on a single project at a time
  • you have influence on the technological solutions applied
  • you focus on product development
  • agile

Team members

  • backend developer
  • technical leader
  • data scientist
  • product owner

Your responsibilities

  • Design and develop scalable data pipelines using Azure, Databricks, Python and PySpark.
  • Use Databricks AI tooling and Spark-based processing to extract and process data from typed and handwritten insurance documentation.
  • Clean, validate, transform and model data for machine learning, analytics and BI use cases.
  • Build reliable datasets that can support downstream reporting, analytical models and AI/ML solutions.
  • Develop and maintain data transformations using Python and PySpark.
  • Support end-to-end data engineering and analytics initiatives, from data ingestion through processing, quality validation and delivery.
  • Apply clean-code principles, write maintainable and well-documented code, and contribute to code reviews.
  • Create and maintain unit tests to ensure the quality, reliability and stability of data pipelines.
  • Collaborate with data, analytics and business stakeholders to understand requirements and translate them into practical technical solutions.
  • Identify data-quality issues and contribute to improvements in data governance, consistency and usability.

Our requirements

  • Commercial experience as a Data Engineer or in a similar data-focused technical role.
  • Strong hands-on experience with Microsoft Azure and Databricks.
  • Very good knowledge of Python and practical experience with PySpark.
  • Experience building and maintaining data pipelines using Apache Spark and Databricks.
  • Experience extracting, processing, cleansing and transforming structured and unstructured or semi-structured data.
  • Understanding of data modelling principles and the ability to prepare data for machine learning, advanced analytics and BI/reporting use cases.
  • Experience writing unit tests and validating data pipeline logic.
  • Strong focus on clean code, code quality, maintainability and technical documentation.
  • Fluent English at least B2
  • Availability to start in the first or second week of October - must have
  • Availability for a 6-month contract.

What we offer

  • Project for 6 months - possible to extend

This is how we work on a project

  • Clean Code
  • code review
  • Continuous Integration
  • documentation
  • test automation
  • unit tests

Benefits

  • remote work opportunities

Recruitment stages

  • Online meeting or phone call with Recruiter (20-30 min.)
  • Technical verification call focused on your experience and project requirements
  • Technical interview with the client, including practical questions and potentially live coding or a technical task
  • Feedback and decision
Oferta pracy dodana 5 dni temu