Senior Data Engineer
Oracle · Nashville, TN · 2 days ago
Information Technology$102k–$210k/yrInternship
Responsibilities
- Data Processing & Pipelining – Data Requirements, Collection, and Infrastructure: Identifies data requirements and business objectives of a project or initiative in collaboration with cross-functional teams. Designs and builds proper data pipelines required for optimal data processing from a variety of data sources. Independently analyzes, designs, and troubleshoots data flows based on business needs. Analyzes business requirements and translates that into a technical specification. Adjusts data collection processes that involve indexing and query optimizations, for optimal performance. Builds Extract, Transform, and Load (ETL) pipelines to support efficient data collection and extraction, independently. Analyzes data sources for profiling the data to ensure pipeline build success. Defines success and failure thresholds for data collection pipelines.
- Data Processing & Pipelining – Data Governance: Implements data governance policies and procedures for data handling (e.g., data retention) to manage data consistency, integrity, accuracy, and reliability throughout the data lifecycle. Redacts Personally Identifiable Information (PII) and Protected Health Information (PHI) data to ensure compliance with data privacy and security standards. Follows data security measures to protect data from unauthorized access, use, disclosure, alteration, or destruction. Ensures data compliance with relevant laws, regulation, and industry standards.
- Data Processing & Pipelining – Data Validation & Quality Assurance: Implements rigorous data validation and integrity checks, identifying and addressing any data quality issues that could impact data pipeline and model performance. Defines data annotation and labeling processes to ensure data quality, independently. Designs and implements automation of data validation and governance.
- Data Pipeline and Solutions Engineering – Pipeline Design: Independently designs, develops, and optimizes automated and scalable data pipeline architectures using ETL processes to build reusable data products. Implements appropriate data storage solutions to store the processed data to be used in a scalable and optimized way for access and analysis. Manages the flow of data pipeline and storage day-to-day operations.
- Data Pipeline and Solutions Engineering – Data Solutions Engineering: Works independently and collaboratively in an agile development environment with other engineers to develop, maintain, and debug data solutions that are scalable, efficient, cost effective, and reliable. Writes runnable code and performs testing and debugging of data solutions, independently.
Qualifications
- Disclaimers: Certain U.S. based or U.S. customer or client-facing roles may be required to comply with applicable requirements, such as immunization/occupational health mandates, and/or drug testing requirements.
- Hiring Range in USD: $102,300 - $209,500 per year. May be eligible for bonus and equity.
- Benefits: Oracle offers a comprehensive benefits package which includes medical, dental, and vision insurance, short term disability and long term disability, life insurance and AD&D, supplemental life insurance, health care and dependent care, flexible spending accounts, pre-tax commuter and parking benefits, 401(k) savings and investment plan with company match, paid time off, paid sick leave, adoption assistance, employee stock purchase plan, financial planning and group legal, and voluntary benefits including auto, homeowner and pet insurance.