Build and support large-scale data infrastructure and datasets for Spotify’s user behavior data. Design data architectures, develop collection and processing systems, and improve reliability, scalability, and usability. The role supports personalization, recommendation systems, and internal data teams while requiring collaboration across product and engineering disciplines.
The Platform team creates the technology that enables Spotify to learn quickly and scale easily, enabling rapid growth in our users and our business around the globe. Spanning many disciplines, we work to make the business work; creating the infrastructure, tooling, frameworks, and capabilities needed to welcome a billion customers.
The Data Platform team develops and maintains the infrastructure that supports Spotify’s data ecosystem. Within Data Platform, the Data Collection product area builds infrastructure that makes it easier to collect and consume data at scale. You’ll join a team focused on making user behaviour data seamless and reliable for data teams across Spotify.
What You'll Do
- Collaborate with product and engineering teams to build and support critical data infrastructure and datasets.
- Build systems that collect and process user behaviour data at scale.
- Help power Spotify’s personalization and recommendation systems, as well as the tools teams use to build and test new features.
- Design and evolve data architectures that support high data volumes and a broad range of internal users.
- Improve the reliability, scalability, and usability of our data infrastructure.
- Work across disciplines to solve complex data challenges and make an impact across Spotify.
Who You Are
- You are an experienced Data Engineer, and proficient in backend development using Java.
- You are comfortable working with large-scale datasets using SQL and platforms such as BigQuery.
- You have experience with JVM-based data processing frameworks such as Flink, Beam, Dataflow, or Spark.
- You are familiar with DevOps practices, cloud infrastructure, containerized applications, and Kubernetes fundamentals.
- You have experience with data modeling and schema design.
- You care about high-quality code and engineering practices such as continuous delivery and automated testing.
- You value experimentation, iterative development, and collaborative software development practices.
- You are comfortable navigating ambiguity and open-ended problems, using data and sound judgment to make decisions.
Where You'll Be
- This role is based in London
- We offer you the flexibility to work where you work best! There will be some in person meetings, but still allows for flexibility to work from home.
Spotify is an equal opportunity employer. You are welcome at Spotify for who you are, no matter where you come from, what you look like, or what’s playing in your headphones. Our platform is for everyone, and so is our workplace. The more voices we have represented and amplified in our business, the more we will all thrive, contribute, and be forward-thinking! So bring us your personal experience, your perspectives, and your background. It’s in our differences that we will find the power to keep revolutionizing the way the world listens.
At Spotify, we are passionate about inclusivity and making sure our entire recruitment process is accessible to everyone. We have ways to request reasonable accommodations during the interview process and help assist in what you need. If you need accommodations at any stage of the application or interview process, please let us know - we’re here to support you in any way we can.
Similar Jobs
AdTech • eCommerce • Information Technology • Software • Travel • Generative AI
Design, build, operate, and evolve scalable, reliable, observable data lake and catalog platform services. Lead technical design, develop well-tested software, support SLOs and incident learning, collaborate with stakeholders, create implementation plans, mentor engineers, and contribute to platform architecture and open-source initiatives.
Top Skills:
Amazon EmrAmazon S3Apache IcebergAWSAws GlueHadoopHiveJavaJvmKotlinSpark
Artificial Intelligence • Big Data • Enterprise Web • Fintech • Software • Financial Services
Design, build, and operate a high-throughput, low-latency real-time data platform using Flink and Kafka. Develop Java and .NET processing and APIs, manage cloud infrastructure (AWS, Terraform), implement CI/CD (Harness), and ensure observability, scalability, performance, and cost-efficiency across production systems.
Top Skills:
.NetApache FlinkAWSHarnessJavaKafkaTerraform
eCommerce • Retail
Design, build, and maintain scalable data pipelines and platform components using Spark/PySpark on Azure. Develop reusable templates, enforce data engineering best practices (CI/CD, testing, observability), support product teams, and optimise pipeline performance, reliability, security, and cost-efficiency.
Top Skills:
AdlsAzureAzure Data FactoryCi/CdDatabricksDbtDelta LakePysparkPythonScalaSparkTerraform
What you need to know about the Manchester Tech Scene
Home to a £5 billion digital ecosystem, including MediaCity, which consists of major players like the BBC, ITV and Ericsson, Manchester is one of the U.K.'s top digital tech hubs, at the forefront of advancements in film, television and emerging sectors like as e-sports, while also fostering a community of professionals dedicated to pushing creative and technological boundaries.

.png)

