Lead Data Engineer
Biogensys · Tallahassee, FL · 4 wk ago
Information TechnologyFull-time
About the role
We are hiring a Lead Data Engineer for one of our clients in Tallahassee, FL.
Responsibilities
- Design, develop, and maintain enterprise data warehouses, cloud data lakes, and lakehouse architectures.
- Build scalable ETL/ELT pipelines using Informatica IDMC.
- Develop and optimize Snowflake data warehouse solutions.
- Design conceptual, logical, physical, and dimensional data models.
- Implement cloud-native data lake solutions utilizing AWS S3, Parquet, Delta Lake, Hudi, and Iceberg.
- Develop robust SQL and Python solutions for data transformation and automation.
- Build and support scalable data integration pipelines, including batch, CDC, and streaming architectures.
- Implement data quality, profiling, observability, and governance best practices.
- Support metadata management, data cataloging, and master/reference data management initiatives.
- Collaborate with business users, analysts, and data scientists to deliver analytics solutions.
- Support BI platforms including Tableau and Qlik.
- Implement DataOps and CI/CD best practices for data engineering.
- Support Dataiku machine learning environments and model lifecycle management.
- Ensure high availability, performance, security, and scalability across enterprise data platforms.
Requirements
- 7+ years of data modeling experience, including Conceptual, Logical, Physical, and Entity Relationship (ER) modeling.
- 5+ years of experience collaborating with business stakeholders and data science teams to translate business requirements into data and analytics solutions.
- 5+ years of experience designing, implementing, and supporting enterprise Data Warehouses, including at least 2 years of Snowflake.
- 5+ years of ETL/ELT development and data integration experience using Informatica.
- 5+ years of SQL programming experience.
- 5+ years of experience with relational and NoSQL databases.
- 5+ years of data quality, testing, and quality assurance experience.
- 4+ years supporting Analytics & Business Intelligence platforms such as Qlik and Tableau.
- 3+ years of cloud data lake development using AWS S3 and Apache Parquet.
- 3+ years of Python or similar object-oriented programming experience.
- 3+ years working in AWS Cloud.
- 2+ years of Databricks Lakehouse implementation using Delta Lake, Apache Hudi, or Apache Iceberg.
- 2+ years implementing DevOps/DataOps practices.
- 2+ years supporting Metadata Management and Data Catalog solutions.
- 2+ years implementing Master Data Management (MDM) and Reference Data Management (RDM) using Informatica Customer 360 and Reference 360.
- 2+ years supporting Dataiku or similar Machine Learning platforms.
Primary Technologies
- Snowflake
- AWS
- Databricks
- Informatica IDMC
- SQL
- Python
- Tableau
- Qlik
- Dataiku
- Erwin Data Modeler
About Us
We are specialized in recruiting and delivering the best professional talent in the industry. We are committed to providing the best experience for both our clients and job seekers. With over two decades of experience in the recruitment industry, we are dedicated to finding the right talent with a high success rate of talent delivery, which helps us continue to be the best in the industry. By responding to this job posting, you are consenting to receive text/SMS messages from us. Thank you.