DevOps Engineer
15552 - 19052 złDevsData LLC
To stanowisko wymaga obecności na miejscu. Zobacz podobne oferty poniżej.
DevOps Engineer $4000-$4900/month Remote (on-site once every two weeks, Warsaw) Full-time B2B We are looking for a DevOps Engineer (Kubernetes/OpenShift) to join a Polish IT consulting company and work on a new data platform for one of its enterprise clients. The platform runs on on-premises OpenShift clusters connected to a hybrid cloud environment. The client's data engineering teams will use it to deploy and run their own batch, streaming, and orchestration jobs. Shared templates and a standard onboarding process mean they won't need to build infrastructure for each new project. The platform is built from open-source components, including Apache Spark, Apache Flink, Apache Airflow, and Trino. You will own how the platform is installed, upgraded, and extended. You will write the Helm charts and GitLab CI templates that teams deploy with, set up GitOps delivery with ArgoCD and Flux, and decide how new teams and workloads get onto the platform. Part of the role is hands-on support: you will work daily with the client's data engineers to get their pipelines running and teach them to use the platform on their own. The ideal candidate enjoys building standards and tooling for other engineers, works well without close supervision on a greenfield setup, and takes design decisions through to production. Requirements 3+ years of hands-on production experience with Kubernetes or OpenShift, including operators, CRDs, networking, and storage Experience with on-premises or bare-metal clusters, beyond managed services such as EKS, AKS, or GKE Ability to write non-trivial Helm charts from scratch, including templating, dependencies, values design, and versioning Experience building and hardening Docker images Experience with GitLab CI, including designing reusable pipeline templates Practical GitOps experience with ArgoCD or Flux Experience deploying and operating data tools on Kubernetes, such as Apache Spark (Spark Operator), Apache Airflow, or Trino Hands-on experience with authentication and access control in Kubernetes: EntraID (Azure AD), OIDC or LDAP, RBAC, secrets, and TLS certificates Good Linux (RHEL) skills, including Bash or Python scripting Ability to design processes, write documentation, and teach other engineers Based in Poland, with the right to work there, and able to travel to the Warsaw office once every two weeks Strong English communication skills for daily work with the client and its stakeholders Nice to Have Experience with OKDP (Open Kubernetes Data Platform) Experience with Apache Flink and the Flink Kubernetes Operator Experience with JFrog Artifactory Experience with Apache Iceberg, Apache Polaris, or S3-compatible object storage such as MinIO or Ceph Observability with Prometheus, Grafana, or OpenTelemetry CKA, CKAD, CKS, or Red Hat OpenShift certifications (e.g. EX280) Responsibilities Install, configure, and upgrade the data platform on on-premises OpenShift clusters across dev, test, and production Integrate new operators and data processing components into the platform, such as the Apache Flink Kubernetes Operator Design and maintain Helm charts for platform components and ETL workloads Build, harden, and publish container images for Spark, Flink, and Airflow Integrate the platform with the hybrid cloud environment and EntraID single sign-on, including RBAC and multi-tenancy Build reusable GitLab CI templates for building, testing, and releasing ETL workloads Automate deployments with Flux and ArgoCD, with promotion from dev to production Design the onboarding process for new teams and ETL workloads, covering namespaces, quotas, access, secrets, data contract validation, and catalog registration Define reference patterns for Spark batch jobs, Flink streaming jobs, Airflow orchestration, and Trino access Write documentation, runbooks, and onboarding guides, and run workshops for platform users Support data engineering teams with troubleshooting and pipeline reviews Monitor platform health, plan capacity, and resolve incidents with infrastructure teams