Senior Data Engineer | Emerging Products
Jobgether · United States · 1 mo ago
RemoteRemoteInformation Technology$110k–$200k/yrFull-time
Accountabilities
- Design and develop a modern data platform that supports large-scale analytics and product innovation.
- Build reliable data infrastructure, improve platform performance, and enable teams across the organization to make data-driven decisions.
- Design and implement scalable lakehouse architectures using structured data layers, including Bronze, Silver, and Gold patterns.
- Build and maintain high-performance streaming and batch data pipelines using technologies such as Kafka, Spark, and Airflow.
- Develop and manage open table format solutions, including Iceberg, Delta Lake, and Hudi, selecting the appropriate approach based on workload requirements.
- Optimize distributed query environments and analytics platforms to improve performance, scalability, and reliability.
- Manage data platform reliability by monitoring pipeline health, troubleshooting data quality issues, and implementing continuous improvements.
- Architect and maintain solutions supporting hybrid streaming and batch processing environments.
- Partner with data scientists, analysts, and product teams to understand requirements and deliver effective data platform capabilities.
- Create and maintain Gold-layer datasets optimized for business intelligence and self-service analytics.
- Contribute to platform standards, engineering best practices, and the evolution of enterprise data infrastructure.
Requirements
- Experienced data engineering professional with deep expertise in distributed systems, modern data architectures, and scalable platform development.
- Significant experience building enterprise-scale data platforms.
- Proven experience designing and implementing lakehouse architectures and Medallion data models.
- Strong hands-on expertise with Apache Kafka, Apache Spark, and Apache Airflow for streaming and batch workloads.
- Advanced knowledge of open table formats such as Apache Iceberg, Delta Lake, or Hudi.
- Experience working with Databricks for large-scale data processing and analytics.
- Strong programming skills in Python and SQL.
- Experience with distributed query engines such as Trino or Starburst is a plus.
- Knowledge of event-driven architectures and hybrid streaming/batch systems.
- Strong understanding of data modeling, performance optimization, and scalable pipeline design.
- Excellent communication skills with the ability to collaborate effectively across technical and business teams.
- Must be a U.S. citizen or lawful permanent resident due to security requirements.
Benefits
- Competitive compensation package with a base salary range of $110,000-$200,000 USD annually, depending on location, experience, skills, and market factors.
- Fully remote work flexibility within eligible U.S. locations, with hybrid options available in select office locations.
- Comprehensive medical, dental, and vision insurance coverage.
- 401(k) retirement plan to support long-term financial planning.
- Life insurance coverage.
- Unlimited paid time off to support work-life balance.
- Opportunities for career growth, advancement, and professional development.
- A collaborative and supportive work environment focused on innovation and continuous learning.