
Backend Engineer - Data Platform
- 🇸🇪 Sweden
- Hybrid
- 5 days ago
- Apache Iceberg
- BigQuery
- Dataflow
- Trino
- Parquet
- MCP
- Kubernetes
- GCP
- IAM
- Java
- Scala
- JVM
- Beam
- Flink
- Snowflake
- Databricks
- AI
LilEngine is leading the Lakehouse, a company-strategic effort to make Apache Iceberg a first-class citizen in Spotify's data ecosystem. We're building it in close collaboration with Google, on open standards, so that different kinds of data are queryable the same way and teams can bring the right engine, from BigQuery to Dataflow to Trino, to the job.
The work has two parts. Hecate enables Iceberg for the Parquet datasets that already exist across Spotify, so teams get the benefits without rewriting their data. Alongside it, we're building the foundations of a new Lakehouse architecture that lets workflows write Iceberg tables natively. We're shipping the MVP this cycle, and you'll help shape what comes next.
We also own BigQuery for all of Spotify. Thousands of engineers, data scientists and pipelines depend on it every day, which puts us on the critical path for storing, finding and querying data. We're judged on reliability, cost efficiency and how quickly we unblock users when something breaks. This backend role sits where open table formats, cloud data warehouses and processing engines meet.
Â
What You'll Do
- Build and evolve the lakehouse: Help build the foundations that let data workflows read and write Iceberg tables directly, from this cycle's MVP to a production-grade platform.
- Drive adoption:Â Help migrate data pipelines from across Spotify onto the new stack, making the move as painless as possible through tooling, standards and hands-on support.
- Own BigQuery at Spotify scale: Manage reservations and capacity, build monitoring and anti-pattern detection, and develop tooling like the BigQuery MCP server so teams can use BigQuery well.
- Keep the platform healthy: Take part in a weekly support rotation and shared on-call. You'll debug production issues that span storage formats, catalog APIs, Kubernetes reconciliation, GCP IAM and BigQuery internals, and you'll automate away recurring toil.
Who You Are
- You have solid experience building and operating backend services in production with Java or Scala.
- You're comfortable on GCP and Kubernetes. You can read kubectl output, reason about IAM and quotas, and work with CRDs and operators. You don't need deep expertise in all of these on day one.
- You've worked with Apache Iceberg or JVM-based processing frameworks like Spark, Beam or Flink, or with a cloud data warehouse such as BigQuery, Snowflake or Databricks.
- You think in systems and can connect the dots when a problem crosses several layers of the stack.
- You like owning things in production, and you're motivated by seeing your systems run reliably for thousands of users.
- You communicate clearly in writing and work well asynchronously across US and EU time zones.
- You own the quality of what you ship, including code written with AI assistance.
Where You'll Be
- This role is based in London or Stockholm. We offer the flexibility to work where you work best! There will be some in person meeetings, but still allows for flexibility to work from home.Â
Backend Engineer - Data Platform · Spotify