Senior Data Platform Engineer
Etraveligroup Tripstack
| Company | Etraveligroup Tripstack |
| Category | Engineering |
| Location | — |
| Remote | — |
| Employment | Not stated |
| Level | Senior |
| Salary | Not stated by the employer |
| Posted | 27 Apr 2026 |
| Last verified | 30 Jul 2026 |
| Source | Employer career page (teamtailor) |
Description
About Tripstack Founded in Toronto, Canada in 2016, Tripstack has been part of Etraveli Group since 2019. It is a B2B Flights as a Service provider and a world leader in virtual interlining. Operating from offices in Canada, India, and Poland, Tripstack is the gateway into Etraveli Group’s world leading tech platform - giving partners access to global flight content, virtual interlining, and a full suite of services including payments, fraud prevention, pricing, and customer support. As a world leader in virtual interlining technology, Tripstack connects non-partner, low-cost and full-service carriers, enabling the creation of unique and flexible itineraries through a simple, cost-effective API. Its technology ingest over 30B price points and handles over 240 million searches daily. Through partnerships with airlines, OTAs, and other distribution channels across the globe, Tripstack expands networks, drives new revenue streams, and offers more choice at competitive prices, all backed by robust technology and traveler protection. For more information, visit: www.tripstack.com www.etraveligroup.com We're hiring for this role in Kraków, Poland or Toronto, Canada. It's a hybrid position, with two days a week in office at the hiring location. The role Tripstack is moving its entire data stack from bare-metal VMs to Kubernetes on OpenStack in a new data centre. We are looking for a senior infrastructure engineer who has done stateful migrations before, who can plan, codify, and execute this one safely, and who will own the day-to-day operational health of the data platform once we are there. This is a hands-on platform and SRE role with a clear, time-bounded mission. You will partner closely with our SRE team on networking, hardware, and Kubernetes fundamentals. You will own the data end to end on the data applications Druid, Spark, Redpanda, Airflow, PostgreSQL, and Elasticsearch. This is not a machine learning role. We have a separate plan for evolving our MLFlow platform, and the right hire here may grow into more of that work over time, but day-one impact is the migration and the operational health of the platform. Responsibilities Lead the data-stack migration Plan and execute the migration of Druid, Spark, Redpanda, and our orchestration layer from bare-metal VMs to Kubernetes on OpenStack, with no downtime on stateful workloads. Design StatefulSet, PVC, pod-disruption-budget, and rolling-upgrade patterns that are safe for production data systems. Codify the migration with Infrastructure as Code — Terraform for OpenStack, Helm or Kustomize for Kubernetes, GitOps via ArgoCD or Flux — so the result is reproducible and supportable by the whole team . Operate the data platform Own the operational health of our production data systems, including Druid, Spark, Redpanda, Airflow, PostgreSQL, and Elasticsearch. You'll handle segment lifecycle, JVM tuning, ingestion specs, broker/coordinator/overlord internals, partition design, consumer lag, and replication tuning Build the KPIs, alerting, dashboards, and runbooks that let us see cluster exhaustion before it becomes an incident, and diagnose it quickly when it does. Own the query, report, segment, and tiering optimisations that keep our analytics cost-effective and responsive under load. Raise the bar on observability and reliability Build the Prometheus, Grafana, and distributed-tracing coverage our data systems need. Treat SLOs, error budgets, and post-incident discipline as table stakes. Partner with SRE on hardware, networking, and Kubernetes fundamentals, while owning the data applications themselves end to end. Requirements Strong Kubernetes experience with stateful workloads — StatefulSets, PVCs, pod disruption budgets, and rolling upgrades for data systems. You have done a real stateful migration before and can talk through what went wrong. Infrastructure as Code
991,236 openings. Erioun finds yours.Scored against your own profile, every hour.Try the radar →