This role will help shape the data foundation that supports product, engineering, and business decision-making across the organization. You will design and scale pipelines and infrastructure that power batch, streaming, and real-time workloads with an emphasis on reliability, observability, and performance. The position also involves strengthening data quality, governance, and event standards while partnering closely with cross-functional teams and mentoring other engineers.
Location: Remote - US based candidates only, no visa sponsorship available
Compensation: $186,500 – $255,000 annually
Responsibilities
- Design data solutions for batch, streaming, and real-time workloads
- Build reliable data pipelines using Spark, Kafka, and other tools
- Contribute to the evolution of the data lake and data platform
- Scale cloud data infrastructure with a focus on reliability and observability
- Implement and enhance data quality and governance across systems
- Define standards for event instrumentation and governance
- Mentor engineers and collaborate across teams on technical initiatives
- BA/BS degree or equivalent experience
- 5+ years of experience in data engineering with production data pipelines
- Strong expertise in Spark and distributed data processing
- Experience with Kafka, Spark Structured Streaming, and CDC
- Proficiency in cloud data infrastructure and CI/CD
- Demonstrated leadership on complex data engineering initiatives
- Strong programming and SQL skills
- Ownership through equity in the company
- Comprehensive health coverage for employees and dependents
- 12 weeks of paid parental leave and additional support for family planning
- Flexible vacation and sabbatical programs to recharge
- Access to mental health resources and wellness support
- 401(k) with 100% employer match up to $6,000/year
- Monthly stipends for work and wellness-related expenses
- Eligibility for annual performance-based bonus program





