Senior Data Modeler (Databricks)
Clarkston Consulting · Austin, TX · 1 mo ago
HybridConsultingFull-time
About the role
Clarkston Consulting is seeking motivated, self-driven leaders to join a firm that values its culture and people as its biggest strengths. As a Senior Data Modeler, you will deliver creative business solutions to market-leading clients in the life sciences, consumer products, and retail industries.
Responsibilities
- Serve our solution architects, software developers, data analysts and data scientists on data and analytics initiatives
- Serve in a lead capacity on client initiatives
- Ensure optimal data delivery architecture is consistent throughout ongoing projects
- Be self-directed and comfortable supporting the data needs of multiple teams, systems, and products
- Be excited by the prospect of building our company's data engineering capabilities to support our next generation of products and data initiatives
- Lead, grow, and mentor a data engineering team in developing, maintaining, and monitoring data solutions for clients
- Assist in defining client solutions and support managing their delivery
How You'll Grow
- Receive the support and mentorship of your Clarkston colleagues and leaders
- Expand your existing skillset with internal and external professional development opportunities
Requirements
- An ideal Senior Data Modeler candidate would have experience:
- Re-engineering one or more modules and associated reports within Databricks environment
- Documenting existing data models and designing the future state data models and architecture, following the set guidelines and standards
- Analyzing existing data models, reports and analysis
- Owning delivery of data products specific to a module or functional area
- Designing and developing data models and transformations using notebooks and pipelines within Azure Databricks
- Serving in role of a hands-on Data Engineer
- Applying expertise in Databricks, Azure Data Factory, SQL, PySpark, and Python
- Developing data management solutions using data lake platform tools and technologies
- Assembling large, complex data sets that meet functional / non-functional business requirements
- Identifying, designing, and implementing internal process improvements: automating manual processes, optimizing data delivery, re-designing infrastructure for greater scalability, etc.
- Leveraging the appropriate infrastructure required for optimal extraction, transformation, and loading of data from a wide variety of data sources using SQL and ‘big data' technologies
- Building analytics tools that utilize the data pipeline to provide actionable insights into customer acquisition, operational efficiency and other key business performance metrics
- Working with stakeholders including the Product, Data, and Design teams to assist with data-related technical issues and support their data infrastructure needs
- Working with stakeholders including the Product, Data, and Design teams to assist with data-related technical issues and support their data infrastructure needs
Additional Qualifications
- 3+ years of experience in a Data Modeler role
- Advanced working SQL knowledge and experience working with relational databases, query authoring (SQL) as well as working familiarity with a variety of databases
- Experience building and optimizing ‘big data' data pipelines, architectures and data sets
- Experience performing root cause analysis on internal and external data and processes to answer specific business questions and identify opportunities for improvement
- Strong analytic skills related to working with unstructured datasets
- Build processes supporting data transformation, data structures, metadata, dependency and workload management
- A successful history of manipulating, processing and extracting value from large disconnected datasets
- Experience supporting and working with cross-functional teams in a dynamic environment
- Experience with Azure or AWS cloud services
- Experience with object-oriented/object function scripting languages such as Scala, Python, Java, and / or C++
- Graduate degree in Computer Science, Statistics, Informatics, Information Systems or an equivalent field