RefolkCandidates
Open nowData and MLData and Insights

Senior Data Engineer

Strava · San Francisco, California

Location
San Francisco, California, United States
Workplace
Hybrid
Employment
Full time
Level
Senior
Posted
2 weeks ago

About this role

About Strava

Strava is the app for active people. With over 200 million athletes in more than 185 countries, it’s more than tracking workouts - it’s where people make progress together, from new habits to new personal bests. No matter your sport or how you track it, Strava’s got you covered. Find your crew, crush your goals, and make every effort count. Start your journey with Strava today.

Our mission is simple: to motivate people to live their best active lives. We believe in the power of movement to connect and drive people forward.

About This Role

We are looking for a Senior Data Engineer to join our Data Team and help build reliable, scalable data systems that support analytics, data science, and critical business use cases across Strava.

In this role, you will design and operate data pipelines, build high-quality domain data models, improve our dbt and data transformation workflows, and help ensure data is accurate, well-governed, and easy to use. You will also contribute to areas such as data ingestion, data quality, privacy and GDPR workflows, and the ongoing evolution of our data warehouse and data lake.

We follow a flexible hybrid model that translates to more than half of your time on-site in our San Francisco office - three days per week.

What You’ll Do:

  • Design, build, and operate foundational data systems and shared data assets that serve a broad range of analytical, operational, and business use cases across Strava.

  • Build and evolve scalable data ingestion and transformation frameworks that move, process, clean, standardize, and organize data across our data lake and data warehouse.

  • Develop reusable data engineering tools and abstractions that improve how engineers build and operate data pipelines, including frameworks and capabilities around technologies such as dbt.

  • Design high-quality, durable domain data models - such as user, subscription, activity, or other core business domains - that provide consistent definitions and reusable foundations for teams across the company.

  • Build systems and workflows that support data governance, privacy, and regulatory requirements, including GDPR-related deletion, retention, access, and data lifecycle management.

  • Improve the reliability and observability of our data platform through automated testing, data quality checks, lineage, monitoring, alerting, and operational tooling.

  • Optimize large-scale data processing and storage for performance, maintainability, scalability, and cost across both warehouse and data lake environments.

  • Partner with data engineers, analytics engineers, software engineers, data scientists, security, privacy, and infrastructure teams to establish scalable data architecture and engineering standards.

You will be successful here by:

  • Thinking beyond individual pipelines and designing reusable systems, abstractions, and data models that solve common problems across multiple teams and use cases.

  • Building well-defined domain data assets with clear semantics, ownership, lineage, and interfaces so downstream consumers can confidently build on top of them.

  • Applying strong data modeling principles to represent complex business entities and relationships in ways that are extensible, understandable, and efficient.

  • Maintaining a high bar for data correctness, reliability, privacy, and operational excellence across critical production data systems.

  • Making thoughtful engineering tradeoffs across data freshness, scalability, storage, compute cost, complexity, and developer productivity.

  • Proactively identifying recurring pain points in the data development lifecycle and creating tooling or platform capabilities that eliminate manual work and improve engineering velocity.

  • Designing data systems with governance and regulatory requirements in mind, rather than treating privacy and compliance as downstream concerns.

  • Bringing software engineering discipline to data infrastructure through testing, modular design, version control, CI/CD, observability, documentation, and code review.

What You’ll Bring to the Team:

  • You have 3 - 5+ years of professional experience in Data Engineering, Data Infrastructure, Software Engineering, or a related field, with experience owning production data systems.

  • You have strong expertise in SQL and data modeling, including experience designing dimensional, normalized, or domain-oriented data models for large-scale analytical systems.

  • You have experience building and operating ETL/ELT and data processing systems using technologies such as dbt, Airflow, Spark, or similar frameworks.

  • You have experience developing reusable tooling, frameworks, or abstractions that improve how data pipelines and transformations are built, tested, deployed, or operated.

  • You are proficient in at least one general-purpose programming language such as Python, Scala, Java, or Go and are comfortable applying software engineering principles to data systems.

  • You understand modern data warehouse and data lake architectures and have worked with technologies such as Snowflake, Databricks, BigQuery, Redshift, Iceberg, Delta Lake, or similar systems.

  • You have experience processing and transforming large datasets, including handling schema evolution, data normalization, deduplication, backfills, incremental processing, and data quality.

  • You understand data governance and data lifecycle concepts such as lineage, retention, deletion, access control, PII handling, and GDPR/privacy requirements.

  • You have experience implementing production-grade data quality, monitoring, alerting, testing, and observability for data pipelines and datasets.

  • You can independently reason about data architecture and make sound technical decisions around modeling, ingestion, transformation, storage, reliability, scalability, and maintainability.

  • You are comfortable working with cloud infrastructure such as AWS, GCP, or Azure and understand the infrastructure that supports large-scale data processing systems.

Experience with Kafka, Flink or other streaming systems; Kubernetes; Iceberg or other open table formats; data catalogs and lineage systems; schema management; CDC; or internal developer platforms for data engineering is a plus.

For information on benefits, please click here.

Why Join Us?

Movement brings us together. At Strava, we’re building the world’s largest community of active people, helping them stay motivated and achieve their goals.

Our global team is passionate about making movement fun, meaningful, and accessible to everyone. Whether you’re shaping the technology, growing our community, or driving innovation, your work at Strava makes an impact.

When you join Strava, you’re not just joining a company - you’re joining a movement. If you’re ready to bring your energy, ideas, and drive, let’s build something incredible together.

Strava builds software that makes the best part of our athletes’ days even better. Just as we’re deeply committed to unlocking their potential, we’re dedicated to providing a world-class, inclusive workplace where our employees can grow and thrive, too. We’re backed by Sequoia Capital, TCV, Madrone Partners and Jackson Square Ventures, and we’re expanding in order to exceed the needs of our growing community of global athletes. Our culture reflects our community. We are continuously striving to hire and engage teammates from all backgrounds, experiences and perspectives because we know we are a stronger team together.

Strava is an equal opportunity employer. In keeping with the values of Strava, we make all employment decisions including hiring, evaluation, termination, promotional and training opportunities, without regard to race, religion, color, sex, age, national origin, ancestry, sexual orientation, physical handicap, mental disability, medical condition, disability, gender or identity or expression, pregnancy or pregnancy-related condition, marital status, height and/or weight.

We will ensure that individuals with disabilities are provided reasonable accommodation to participate in the job application or interview process, to perform essential job functions, and to receive other benefits and privileges of employment. Please contact us to request accommodation.

California Consumer Protection Act Applicant Notice

As published by Strava. Applications are handled on their site.

Skills this posting mentions

ELTData EngineeringApache Kafka

About Strava

Strava is Swedish for “strive,” which epitomizes who we are and what we do. We’re a passionate and committed team, unified by our mission to connect athletes to what motivates them and help them find their personal best. And with billions of activity uploads from all over the world, we have a humbling and audacious vision: to be the record of the world’s athletic activities and the technology that makes every effort count. Strava builds software that makes the best part of our athletes’ days even better. And just as we’re deeply committed to unlocking their potential, we’re dedicated to providing a world-class, inclusive workplace where our employees can grow and thrive, too. We’re backed by Sequoia Capital, Madrone Partners and Jackson Square Ventures, and we’re expanding in order to exceed the needs of our growing community of global athletes. Our culture reflects our community - we are continuously striving to hire and engage diverse teammates from all backgrounds, experiences and perspectives because we know we are a stronger team together.When you’re ready for a challenge and a team that will support you along the way, join us!

All 30 openings at Strava

One click, then it is written

Apply to Strava with a resume written for this role.

Queue Senior Data Engineer and I read the posting, rewrite your resume against it, draft the cover letter, and score the fit. Then you press send, or press one button and I fill in Strava’s form for you.

  1. 01Drop your resume

    A PDF or a LinkedIn URL. About a minute, once.

  2. 02I rank the openings

    Every weekday morning, the live catalog scored against your history. Up to 20 worth your time, not two hundred links.

  3. 03Each one is written up

    Resume rewritten for the posting, a cover letter, a fit score. Press send, or let me fill in the form.

  • 25 sent a week, free
  • No card
  • Nothing sent until you say so

More roles at Strava

See all

Similar roles elsewhere

See more

Put this to work

Paste your career in once. Every application after that is written for you.

Drop a resume or a LinkedIn URL. I rank the live openings against it, rewrite the resume and write a cover letter for the best of them, and fill in the employer's form when you press the button. You read, you decide what goes out.

  1. 01Drop your resume

    A PDF or a LinkedIn URL. About a minute, once.

  2. 02I rank the openings

    Every weekday morning, the live catalog scored against your history. Up to 20 worth your time, not two hundred links.

  3. 03Each one is written up

    Resume rewritten for the posting, a cover letter, a fit score. Press send, or let me fill in the form.

  • New matches ranked and written before you are up.
  • Every bullet stays inside what your history supports. Nothing invented.
  • Queued, submitted, interviewing, offer: one screen, not a spreadsheet.

500 free credits on sign-up. No card. Nothing is sent until you say so.

Listed from the job board Strava publishes. Refolk is not the employer and does not handle their hiring. Applications go to Strava directly.