Intern - Data Engineer
Job Summary
The Data Solutions Team is a core part of Optimum Media. Our team incorporates data engineering, data management, data onboarding, and data analysis into a singular group. We are the data specialists that ensure the highest operational efficiency of reporting and analytics within the organization and are seeking a Data Engineer Intern to support the development and maintenance of data pipelines, data infrastructure, and analytical data solutions.
Responsibilities
- Collaborate with data engineers and cross-functional teams to support the development and maintenance of data pipelines, data structures, and ETL/ELT processes.
- Work with large relational databases and data warehouse platforms (e.g. Google BigQuery, Snowflake, Databricks) to support data processing and reporting needs.
- Write SQL queries & Python code for data structuring, data validation, process automation, and data engineering workflows.
- Agregate, onboard, and investigate data from multiple sources; identify and troubleshoot data quality issues.
- Document data processes, architecture, and development activities following established engineering best practices.
Qualifications
- Bachelor's/Master’s Degree in Data Science/Statistics/Computer Science/Computer Engineering/Information Science/Mathematics or equivalent job experience
- Graduate degree in related field, preferred
- 1-2 Years minimum relevant working experience, preferred
- Intermediate SQL programming skills, including working with relational databases, data manipulation, and data transformation
- Comfortable reading, writing, troubleshooting, and understanding Python code
- Exposure to cloud data platforms and technologies such as Google BigQuery, AWS Redshift or Databricks
- Familiarity with ETL/ELT concepts and data processing workflows
- Strong problem-solving skills and willingness to learn new technologies and tools
- Team-oriented with strong communication skills and the ability to collaborate with both technical and non-technical stakeholders
- Exposure to Snowflake or other cloud-based data warehouse technologies, preferred
- Familiarity with PySpark, Spark or other distributed data processing technologies through coursework, projects, or internships, preferred
- Experience scripting and deploying ETL jobs, cron jobs and code changes, preferred
- Experience with Linux or Bash scripting is a plus
- Experience with Git, GitHub, or other code versioning tools is a plus
- Basic statistical analysis, data science, or machine learning experience is a plus
- Experience in Media, Television, Advertising, or Digital Marketing is a plus
Benefits
We are an Equal Opportunity Employer committed to recruiting, hiring and promoting qualified people of all backgrounds regardless of gender, race, color, creed, national origin, religion, age, marital status, pregnancy, physical or mental disability, sexual orientation, gender identity, military or veteran status, or any other basis protected by federal, state, or local law. The Company collects personal information about its applicants for employment that may include personal identifiers, professional or employment related information, photos, education information and/or protected classifications under federal and state law. This information is collected for employment purposes, including identification, work authorization, FCRA-compliant background screening, human resource administration and compliance with federal, state and local law. Applicants for employment with The Company will never be asked to provide money (even if reimbursable) as part of the job application or hiring process.
Pay
The starting pay rate/range at time of hire for this position in the posted location is $35.00 / hour. The rate/range provided herein is the anticipated pay at the time of hire, and does not reflect future job opportunity.