Senior Data Engineer, Python, Spark
Roku
In this role you will design and build scalable data pipelines and models to surface metrics across Roku devices and platforms, enabling data-driven decisions. You’ll contribute to a world-class data platform that supports internal and external partners, helping understand feature resonance and improve user experiences. You will influence the product roadmap through architecture discussions and own initiatives end-to-end. Join a fast-growing company shaping how the world watches TV and collaborate with cross-functional teams.
Pay / Benefits- global mental health and financial wellness resources
- healthcare (medical, dental, vision)
- retirement options
- flexible Fridays for remote work
- local leave policies and personal needs support
- accommodations available
- Build scalable, fault-tolerant batch and streaming data processing systems handling tens of terabytes daily
- Design robust data solutions that translate complex datasets into self-service models
- Develop pipelines ensuring high data quality and resilience to imperfect source data
- Define and maintain data mappings, business logic, transformations and data quality standards
- Debug low-level systems, measure performance and optimize large production clusters
- Participate in architecture discussions and own initiatives from concept to delivery
- Maintain and evolve existing platforms, introducing modern technologies and architectures
- Strong SQL skills
- Proficiency in Python
- Proficiency in at least one object-oriented language
- Experience with big data technologies (HDFS, YARN, MapReduce, Hive, Kafka, Spark, Airflow, Presto)
- Experience with AWS or GCP; Looker is advantageous
- Solid background in data modelling for scalable architectures
- BS in Computer Science (MS preferred)
- Collaborative mindset
- Ownership and problem-solving attitude
- Ability to influence product direction
- SQL
- Python
- Object-oriented programming
Reference: WJ-747_30145843