ETL Developer / Data Engineer
Staffing Spot, Inc. · Charlotte, NC · 2 wk ago
On-siteBusiness DevelopmentContract
Location: Charlotte, NC
Work arrangement: 5 days/week onsite
Employment type: Contract
About the role
We are seeking a Mid-Level Software Engineer / Data Engineer with strong hands-on experience in Ab Initio, Python, UNIX Shell Scripting, SQL/PLSQL, and enterprise ETL development.
Responsibilities
- Design, develop, maintain, and optimize scalable batch and near-real-time data pipelines.
- Develop complex ETL solutions using Ab Initio, Python, PySpark, SQL, and PL/SQL.
- Build and maintain complex Ab Initio graphs, Psets, and reusable components.
- Perform data integration, cleansing, normalization, transformation, and validation across multiple source systems.
- Translate technical designs and solution blueprints into high-quality, reusable production code.
- Develop complex SQL/PLSQL queries and perform query and pipeline performance tuning.
- Work with Oracle, Teradata, BigQuery, and other enterprise data platforms.
- Support migration and modernization of legacy data-processing solutions to GCP and cloud-native technologies.
- Build, schedule, orchestrate, and deploy data solutions through enterprise CI/CD pipelines.
- Work with AutoSys, Airflow, and/or Cloud Composer for job scheduling and orchestration.
- Participate in code reviews, automated testing, deployment, and Git-based development workflows.
- Monitor production data pipelines, troubleshoot issues, perform root-cause analysis, and support SLA-driven operations.
- Collaborate with Product Owners, Architects, Engineers, and business stakeholders.
- Contribute to data quality, governance, metadata, lineage, and operational excellence initiatives.
- Evaluate opportunities to modernize Ab Initio workloads using Python, PySpark, Spark, and SQL patterns.
- Use AI-assisted development tools appropriately to improve engineering productivity while maintaining security and code-quality standards.
Requirements
- 5+ years of Data Engineering / ETL development experience or equivalent combination of education, training, and professional experience.
- 4+ years of hands-on Ab Initio development, including complex graphs, Psets, debugging, and performance tuning.
- Strong UNIX/Linux Shell Scripting experience.
- 3+ years of Python development experience.
- Hands-on experience with PySpark and distributed data processing.
- 4+ years of SQL and PL/SQL development experience.
- Strong experience with Oracle and/or Teradata.
- Experience with complex query development, debugging, and performance tuning.
- 3+ years of ETL development, ETL design, data warehousing, and data modeling experience.
- Experience developing and supporting production-grade enterprise data pipelines.
- Strong understanding of data integration, transformation, cleansing, and reconciliation processes.
Preferred Qualifications
- Financial services or mortgage industry experience.
- Experience supporting production applications, including monitoring, SLAs, incident management, root-cause analysis, and performance optimization.
- Experience working in hybrid on-premises and cloud environments.
- Experience with GCP, BigQuery, and Dataplex.
- Experience migrating legacy ETL/data platforms to cloud-native solutions.
- Experience with AutoSys, Airflow, or Google Cloud Composer.
- Experience with Jenkins, Harness, and/or uDeploy.
- Strong Git-based development, code review, and automated testing experience.
- Experience migrating or converting Ab Initio graphs to Spark/PySpark/SQL-based solutions.
- Experience with Informatica Data Quality, including profiling, rules, scorecards, metrics, and exception workflows.
- Understanding of data governance concepts such as metadata, classification, stewardship, and lineage.
- Experience designing near-real-time or micro-batch data processing patterns.
- Understanding of late-arriving and out-of-order data handling.
- Familiarity with GCP operational concepts including IAM/service accounts, job monitoring, quotas, and cost controls.
- Experience using AI-assisted coding tools in day-to-day development while following enterprise security and quality standards.
Technical Environment
- ETL & Data Engineering: Ab Initio, ETL, Data Warehousing, Data Modeling
- Programming: Python, PySpark, SQL, PL/SQL, UNIX Shell
- Databases: Oracle, Teradata, BigQuery
- Cloud: Google Cloud Platform (GCP), BigQuery, Dataplex
- Orchestration: AutoSys, Airflow, Cloud Composer
- CI/CD: Jenkins, Harness, uDeploy, Git
- Data Quality & Governance: Informatica Data Quality, Data Profiling, Metadata, Classification, Lineage
What We Value
We are looking beyond job titles and keyword matches. Strong candidates should be able to clearly demonstrate:
- What they have actually built and delivered.
- The business problems their solutions addressed.
- Their level of hands-on technical ownership.
- Experience working with complex enterprise data environments.
- Strong ETL, data integration, and data transformation expertise.
- Experience contributing to modernization and cloud migration initiatives.
- The ability to learn, adapt, and work effectively across evolving technologies.