Lead Data Engineer – Data & AI, Supply Chain
Russell Tobin · San Francisco, CA · 1 mo ago
On-siteOTHR$65–$70/hrFull-time
Key Responsibilities
- Design, develop, and implement scalable data pipelines and data products on Google Cloud Platform (GCP).
- Build and optimize enterprise data solutions using BigQuery, Dataproc, SQL, and dbt.
- Design scalable and efficient data models to support analytics and reporting.
- Develop and maintain ETL/ELT pipelines to ingest, transform, and publish data from enterprise systems.
- Collaborate with Product Managers, Business Analysts, Architects, and business stakeholders to translate business requirements into technical solutions.
- Lead technical design discussions, perform code reviews, and promote engineering best practices.
- Optimize data platforms for performance, scalability, reliability, and cost efficiency.
- Implement monitoring, testing, and operational best practices for production workloads.
- Contribute to reusable frameworks, engineering standards, and technical documentation.
- Support production issue resolution and continuous platform improvements.
- Participate in Agile ceremonies including sprint planning, estimation, and backlog refinement.
- Mentor and guide junior engineers while fostering engineering excellence.
Requirements
- Bachelor's degree in Computer Science, Engineering, Information Systems, or a related field (or equivalent experience).
- 8+ years of experience in Data Engineering, including technical leadership on enterprise-scale projects.
- Strong hands-on experience with Google Cloud Platform (GCP).
- Expert-level proficiency in:
- BigQuery
- Dataproc
- SQL
- dbt (Data Build Tool)
- Strong experience designing and implementing modern ETL/ELT solutions.
- Solid understanding of data modeling, including dimensional modeling, normalized data models, and data warehouse design.
- Experience building scalable, cloud-native data pipelines.
- Experience with Git, CI/CD pipelines, and software engineering best practices.
- Strong analytical, troubleshooting, and problem-solving skills.
- Excellent communication and collaboration skills.
Skills
- Experience with Apache Airflow for workflow orchestration (nice to have).
- Experience with Apache Kafka or other streaming technologies (nice to have).
- Experience with PySpark for distributed data processing (nice to have).
- Strong Python programming skills for data engineering and automation (nice to have).
- Knowledge of data quality, metadata management, and data governance practices (nice to have).
Nice to Have Skills
- Experience in one or more of the following domains is highly desirable:
- Retail (Apparel)
- Supply Chain and Logistics
- Transportation Management
- Warehouse Management Systems (WMS)
- Distribution Center Operations
Benefits
- Comprehensive healthcare coverage (medical, dental, and vision plans).
- Supplemental coverage (accident insurance, critical illness insurance and hospital indemnity).
- 401(k)-retirement savings.
- Life & disability insurance.
- An employee assistance program.
- Legal support.
- Auto, home insurance, pet insurance.
- Employee discounts with preferred vendors.
Pay
$65-70 an hour on W2 (DOE)
Schedule
Hybrid