Data Engineer (AI Pipelines)
DeWinter Group · Campbell, CA · 2 wk ago
Information Technology$50–$175/hrContract
Our client, a leader in AI testing, is looking for a skilled Data Engineer (AI Pipelines) to join their team for a 12-month remote contract engagement. This project involves building scalable ETL/ELT pipelines to ingest, clean, and transform massive datasets for AI training, inference, and low-latency real-time applications. This is a high-impact role that requires a self-motivated professional who can hit the ground running and deliver results quickly.
Responsibilities
- Build scalable ETL/ELT pipelines to ingest, clean, and transform massive datasets for AI training and inference.
- Implement data quality checks and automated validation to prevent "garbage in, garbage out" in AI systems.
- Manage the storage and versioning of large datasets using tools like DVC or Snowflake.
- Optimize data retrieval patterns for low-latency RAG systems and real-time model serving.
- Collaborate with ML engineers to ensure data features are consistent across training and production.
Requirements
- 4+ years of experience in Data Engineering.
- Deep expertise in SQL, Spark, Python, and modern data stack tools (Airflow, dbt).
- Demonstrated ability to work autonomously and manage your own time effectively to meet project goals.
- Experience with cloud data warehouses (Snowflake, BigQuery) and Git.
- Strong communication skills to provide clear and concise status updates to the project team.
Pay
$50/hr – $175/hr
Schedule
12-month contract, starting immediately. Remote work.