Senior Data Engineer
Tata Consultancy Services · Raleigh, NC · 3 days ago
Information Technology$100k–$120k/yrFull-time
Responsibilities
- Design and implement scalable, resilient data pipelines using Snowflake features including Snowpipe, Tasks, Streams, Dynamic Tables, and advanced SQL.
- Build and maintain DBT models with strong testing, documentation, and lineage.
- Develop Python ingestion frameworks for files and APIs, including schema validation, retries, and metadata capture.
- Engineer ingestion for CSV, fixed width multi record layouts, JSON, XML, Excel, and semi structured formats.
- Design Mainframe VSAM data ingestion pattern for complex EBCDIC data formats.
- Detect, analyze, and manage schema drift across file, API, and replicated database sources.
- Implement metadata driven schema evolution strategies to ensure downstream stability.
- Carefully coordinate schema changes through controlled CI/CD workflows.
- Configure and manage Qlik Replicate tasks for CDC and full load replication from Oracle, SQL Server, and DB2.
- Ensure idempotent, auditable, and recoverable replication pipelines with strong monitoring and reconciliation.
- Implement Snowflake Data Masking policies, including dynamic masking, conditional masking, and role based masking rules.
- Apply Protegrity tokenization for sensitive data fields across ingestion and transformation layers.
- Enforce RBAC, data access controls, and governance standards across Snowflake and supporting systems.
- Build and schedule workflows using Astronomer Airflow, ensuring dependency management, retries, SLAs, and observability.
- Integrate pipelines with enterprise DevOps processes using GitLab and Azure DevOps for CI/CD automation.
- Manage code repositories using GitLab, including branching strategies, merge requests, code reviews, and approvals.
- Implement monitoring and alerting for ingestion pipelines, schema drift, replication, and transformation workloads.
- Optimize Snowflake compute, storage, and query performance; scale ingestion pipelines to meet evolving data volume and latency requirements.
Requirements
- 8+ years of hands on data engineering experience.
- Deep expertise with Snowflake, including data masking policies, RBAC, performance tuning, and advanced SQL.
- Strong experience with Qlik Replicate for CDC and database replication.
- Excellent proficiency in Python and Pyspark for ingestion frameworks and automation.
- Hands on experience with DBT Cloud and Astronomer Airflow.
- Experience with schema drift detection and schema evolution patterns.
- Experience with GitLab and CI/CD pipelines.
- Familiarity with Protegrity or similar data protection platforms.
Pay
$100,000 - $120,000 per year