Sr. Data Engineer
Linde · Tonawanda, NY · 3 wk ago
Information Technology$108k–$159k/yrFull-time
req31963
About the role
Linde is a leading global industrial gases and engineering company with 2025 sales of $34 billion. We live our mission of making our world more productive every day by providing high-quality solutions, technologies and services which help our customers succeed and contribute to sustainability, decarbonization, and protecting our planet. Our industrial gases and technologies enable applications in space exploration, semiconductor manufacturing, healthcare, clean hydrogen production, and carbon capture.
Responsibilities
- Architect, build, and maintain scalable, distributed data pipelines using Apache Spark and Microsoft Fabric to process large structured and unstructured datasets
- Integrate data from diverse internal and external systems, ensuring reliability, lineage, and consistency across the enterprise
- Lead optimization of ETL/ELT workloads for large-scale analytics to improve cost efficiency, throughput, and reliability
- Define and implement standards for data quality, metadata management, cataloging, lineage, and governance compliance
- Collaborate with data scientists, analysts, architects, and IT teams to define requirements, deliver insights, and integrate analytical models
- Develop and maintain comprehensive documentation for pipeline architectures, workflows, schemas, and operational processes
- Evaluate emerging technologies and introduce modern data engineering practices such as Lakehouse patterns, delta formats, automation, and real-time processing
- Lead resolution of complex data pipeline failures, ensuring platform stability, enterprise-grade reliability, and enforce data security, privacy, and access governance policies
Requirements
- Bachelor’s or master’s degree in computer science, Engineering, Information Systems, or related field
- 7+ years of professional experience in data engineering, data architecture, or large-scale data platform development
- Expertise in Apache Spark for large-scale batch and streaming workloads
- Deep hands-on experience with Microsoft Fabric, including Data Engineering, Data Factory, Data Pipelines, and Lakehouse implementations
- Advanced proficiency in SQL, Python, and/or Scala
- Hands-on expertise with Microsoft Azure (preferred), AWS, or GCP cloud data services
- Strong understanding of distributed systems, lakehouse patterns, and data modeling
- Proven experience designing and optimizing complex ETL/ELT workflows
- Strong communication skills, leadership capability, and technical mentorship experience
Benefits
- Competitive compensation and an outstanding benefits package
- Health, dental, disability, and life insurance
- Paid holidays and vacation
- Employee discount program
- Opportunities for educational and professional growth
Pay
Salary Range: $108,225 - $158,730