Data Engineer at Superhuman
on-site · full_time · Visa sponsorship
Apply for this role at Superhuman
Responsibilities:
- Architect and lead design of large-scale data pipelines and data lakes handling tens of billions of daily events.
- Build robust real-time and batch ETL/ELT systems to make data available for analytics, ML, and product features.
- Ensure data reliability, security, and scalability through monitoring, testing, and observability practices.
- Collaborate with backend, analytics, data science, and ML teams to deliver data products and enable business outcomes.
- Mentor engineers, influence platform strategy, and contribute to architecture and operational best practices.
- Optimize cost, performance, and throughput of data infrastructure and cloud resources.
Requirements:
- Significant experience with distributed data systems and high-volume pipelines (Spark, Kafka or equivalent).
- Strong SQL skills and experience with ETL/ELT orchestration tools (Airflow, Dagster, dbt, etc.).
- Proficiency in one or more languages used for data engineering (Python, Scala, Java).
- Knowledge of cloud data platforms, data lakes, and storage formats; experience with cost and performance optimization.
- Strong data modeling, governance, and observability experience; familiarity with privacy/regulatory concerns a plus.
- Excellent collaboration, communication, and mentoring skills; proven track record building production-grade systems.