Junior Biostatistician
GDIT delivers technology solutions and mission services to U.S. government, defense, and intelligence agencies. Our work supports complex projects that ensure today’s safety and tomorrow’s advancements.
About the role
The Junior Biostatistician supports data engineering, statistical analysis, and analytics functions that inform readiness, performance, and decision-support activities. This role processes, curates, and analyzes structured and unstructured data to create actionable insights, dashboards, and foundational data products. The Junior Biostatistician assists senior analysts and enterprise data teams by preparing datasets, developing pipelines, and supporting reporting tools used across programs and operational environments.
Responsibilities
- Acquire, integrate, and prepare data from multiple sources to enable operational reporting, readiness analytics, and decision support.
- Design and implement ETL/ELT pipelines using SQL Server and Databricks leveraging Apache Spark (PySpark) for scalable batch processing and data transformation.
- Develop and maintain curated, analysis-ready datasets (e.g., dimensional models, governed reporting tables) with attention to data quality, lineage, and repeatability.
- Optimize data processing performance (partitioning strategies, efficient joins, job tuning) and contribute to reliable scheduling/orchestration and monitoring of recurring workloads.
- Apply statistical methods, data mining, and machine learning to generate actionable insights, forecasts, and predictive analytics relevant to military missions and readiness.
- Create dashboards and analytical products using Power BI, including semantic modeling and measures to support consistent metrics and reporting.
- Support low-code solutions by integrating data products into Power Apps (Canvas and Model-Driven) and automations using Power Automate.
- Contribute to text and language-focused use cases by preparing and analyzing unstructured data (basic NLP preprocessing, labeling/curation, feature creation).
- Demonstrate familiarity with LLMs and agent concepts (prompting patterns, retrieval-style workflows, evaluation considerations) and assist senior team members in implementing these capabilities where appropriate.
- Document data sources, transformations, assumptions, and analytical methodologies to ensure transparency and operational continuity.
Requirements
- Bachelor’s degree with 1–2 years of experience in quantitative science, social science, or a related discipline.
- Proficient with Microsoft Office programs (Word, Excel, Access).
- Working knowledge of SQL Server (SQL/T-SQL), including querying, joins, indexing basics, and relational schemas.
- Hands-on experience or strong exposure to Databricks and Apache Spark (preferably PySpark) for scalable data preparation.
- Foundational programming skills in Python or R for analysis, data manipulation, and ML workflows.
- Understanding of data engineering concepts: ETL/ELT, data modeling basics, data quality checks, reproducible pipelines, and version control (Git).
- Experience building visual analytics with Power BI (reports, datasets/semantic models, DAX fundamentals preferred).
- Familiarity with the Power Platform: Power Apps (Canvas and Model-Driven) and Power Automate for workflow integration.
- Basic familiarity with NLP and Large Language Models, including typical use cases and data requirements (text preparation, evaluation datasets, governance considerations).
- Ability to communicate clearly with stakeholders and translate mission needs into data requirements and analytical outputs.
- Experience with web application development technologies: HTML, CSS, JavaScript, React, plus backend exposure to Django or ASP.NET Core.
- Experience with advanced Spark/Databricks patterns (Delta tables, incremental loads, job/workflow scheduling, performance tuning).
- Familiarity with deploying analytics/ML into applications (APIs, dashboards, or integrated workflows) and basic MLOps concepts.
- U.S. citizenship required.
- Secret or Top-Secret Clearance, or the ability to obtain clearance.
Skills
- Analytics
- Datasource
- Data Transformation
- Structured Query Language (SQL)
Pay
The likely salary range for this position is $63,312 – $85,658. Salary will be set based on experience, geographic location, and possibly contractual requirements.
Schedule
- Scheduled Weekly Hours: 40
- Travel Required: Less than 10%
- Telecommuting Options: Onsite
Benefits
- 401K with company match
- Comprehensive health and wellness packages (medical, dental, vision)
- Paid time off: vacation, sick, personal time, holidays, parental leave, military leave, bereavement, jury duty
- Short- and long-term disability benefits
- Life, accidental death and dismemberment, personal accident, critical illness, and business travel and accident insurance
- Internal mobility team dedicated to career growth
- Professional growth opportunities including paid education and certifications
Locations: Various OCONUS & CONUS sites, including USA NC Fort Bragg, USA CA San Diego, USA FL Hurlburt Field, USA KY Fort Campbell, USA WA Joint Base Lewis-McChord.