The opportunity
As a Data Engineer on Analytics Data Engineering, you will build and operate the pipelines and data models the rest of Dropbox relies on to understand its products and its business. You will own well-scoped pipelines end to end — design, build, test, ship, monitor — with senior…
What you'll do
Build and maintain Spark and SparkSQL jobs that populate company data models
Own well-scoped pipelines end to end, from requirements through deployment, monitoring, and iteration
Contribute to data quality frameworks, testing, and data lineage instrumentation
Partner with data scientists, analysts, product managers, and engineers to turn data needs into durable models
Extend datamarts and data models supporting recurring reporting and analysis across products
Improve the reliability and cost efficiency of existing pipelines, dashboards, and frameworks
What they're looking for
- + years of development experience in Spark, Python, Java, C++, or Scala
- + years of SQL experience, including query performance tuning
- + years of experience with schema design and dimensional data modeling
- Experience building and maintaining production data pipelines that others depend on
- Working exposure to a cloud data lake or lakehouse platform, Databricks preferred
- Clear written and verbal communication with non-engineering partners, and a: track record of asking for help and feedback early
- BS in Computer Science or a related technical field involving coding (e.g.: physics or mathematics), or equivalent technical experience
- + years of SQL experience