Senior Data Engineer, Python, Spark
Roku
As a Senior Data Engineer, you will design scalable data models and build robust pipelines to measure device and product metrics across Roku devices and apps. You’ll work with cross-functional partners to surface insights that inform feature decisions and improve the user experience. You help scale a fault-tolerant, petabyte-scale data platform that enables data-driven growth and experimentation. This role offers a blend of architecture influence, hands-on data work, and collaboration in a fast-growing, mission-driven company.
Pay / Benefits- global mental health and financial wellness support
- healthcare (medical, dental, vision)
- retirement options
- time off policies
- local benefits where applicable
- accommodations available
- Build scalable batch and streaming data processing systems handling tens of terabytes daily
- Design robust data solutions and simplify complex datasets into self-service models
- Develop pipelines ensuring data quality and resilience to imperfect source data
- Define and maintain data mappings, transformations and quality standards
- Debug, measure performance, and optimise large production clusters
- Contribute to architecture discussions and drive new initiatives from concept to delivery
- Maintain and evolve platforms, introducing modern technologies where applicable
- Strong SQL skills
- Proficiency in Python
- Proficiency in at least one object-oriented language
- Experience with big data technologies (HDFS, YARN, MapReduce, Hive, Kafka, Spark, Airflow, Presto)
- Experience with cloud platforms (AWS, GCP) and Looker is advantageous
- Solid background in data modelling for scalable architectures
- BS in Computer Science (MS preferred)
- collaboration
- problem-solving
- ownership
- SQL
- Python
- object-oriented programming
Reference: WJ-747_30144074