Baseball Tech Manager - Data Systems
Role Overview
The Baseball Tech Manager - Data Systems is responsible for the production readiness, deployment, monitoring, support and continuous improvement of Baseball’s post-game data systems, while maintaining close alignment with live game operations across both cloud and on-prem environments.
Responsibilities
Own production deployments for Baseball post-game data systems and related cloud/on-prem components. Test, validate and approve releases before deployment into production environments. Maintain staging and production environments, including configuration, release documentation and deployment processes. Act as a production-readiness gatekeeper, escalating or blocking releases that are not supportable, scalable or safe to deploy.
Serve as a Tech subject matter expert for Baseball data products across backend, frontend, cloud and operational workflows. Own data reprocessing workflows, including technical execution, quality validation and delivery of reprocessed feeds to customers. Maintain data feeds, feed specifications, validation processes and sample data generation for internal and external stakeholders. Build scripts, tools and automation to improve reprocessing reliability, reduce manual effort and scale delivery.
Stay closely aligned with Live Operations to understand how live capture, on-prem systems, cloud delivery and operational workflows impact post-game data quality. Participate in live game incident escalation where system health, data delivery or post-game outputs may be affected. Investigate data-quality issues across live and post-game workflows, identifying root causes with Operations, Solution Owners and Engineering. Translate recurring operational issues into improvements in tooling, monitoring, automation, documentation or deployment processes.
Own monitoring dashboards for Baseball post-game systems, live delivery touchpoints and operational data-quality workflows. Build and maintain Datadog, AWS and other observability dashboards for Live Operations, Tech leadership and executive reporting. Monitor data pipelines, product health, reprocessing systems, queues, cloud services and production environments. Create alerts and reporting views that provide visibility of system health, data quality, customer delivery risk and incident trends.
Provide advanced troubleshooting and third-line support for Baseball data products, rollouts, live delivery touchpoints and production incidents. Manage and troubleshoot components such as RabbitMQ queues, Keycloak/client credentials, APIs, databases, cloud services and deployment environments. Investigate production issues across cloud and on-prem systems, escalating to Engineering or DevOps where required. Contribute to root-cause analysis and post-incident reviews, ensuring learnings are converted into improved tooling, monitoring, documentation or process changes.
Build scripts, tools and dashboards to analyse product performance, data quality and operational reliability. Automate repeatable operational processes across reprocessing, validation, deployments, monitoring and reporting. Identify manual workflows that limit Baseball’s ability to scale and work with Tech leadership to improve them. Contribute technical input to Tech forecasting and budget planning, particularly around cloud usage, tooling, infrastructure and support requirements.
Own operational documentation and training materials for Baseball post-game systems and related Tech workflows. Maintain runbooks, deployment guides, rollback procedures, troubleshooting guides, monitoring guides and reprocessing documentation. Train Operations and Tech users on new features, workflows, dashboards and production support processes. Ensure new releases are handed over with clear operational guidance, known risks, escalation routes and support expectations.
Qualifications
3+ years’ experience in technical operations, product engineering, DevOps, platform support, data operations or a similar role.
Experience supporting production systems in a cloud, data-platform or software-hardware environment.
Strong understanding of AWS services, data pipelines, APIs, databases and monitoring tools.
Experience with observability and monitoring platforms such as Datadog, AWS CloudWatch or similar.
Experience with SQL, scripting and automation for operational tooling, data analysis or workflow improvement.
Familiarity with message queues, authentication/credential management and production environment support.
Ability to troubleshoot complex technical issues across backend systems, frontend workflows, cloud infrastructure and operational processes.
Strong communication skills, with the ability to work across Engineering, Operations, Product, Solution Owners and customer-facing teams.
Ability to translate operational pain points into scalable technical improvements.
Comfortable working in high-pressure live service environments where production incidents can affect customer delivery.