Network Modeling / Automation Engineer
Volta · Palo Alto, CA · 4 days ago
HybridInformation TechnologyFull-time
About The Role
The Network Modeling / Automation Engineer turns Volta's network fabric — spine-leaf, InfiniBand/RoCE, and the underlay/overlay stack — into models and automation that scale with the fleet. Reporting to the Network Platform Team Lead, you will own the tooling that provisions, verifies, and monitors network configuration across every site, replacing manual, one-off work with reproducible, testable code. You will work closely with the network engineering team to capture design intent as data and code, build the pipelines that push validated changes to production safely, and give the fleet a single source of truth for what the network should look like versus what it actually is.
What You Will Be Doing
- Build and maintain network modeling tools that represent Volta's fabric topology, configuration, and intent as structured, version-controlled data
- Design and own automation pipelines for network provisioning, configuration deployment, and change validation across the fleet
- Develop pre-deployment verification and simulation tooling to catch design and configuration errors before they hit production
- Build drift detection between intended network state and live configuration, and automate remediation where safe to do so
- Partner with the Network Platform Team Lead and network engineers to translate fabric designs into automated, repeatable build processes
- Integrate network automation into broader platform CI/CD, working alongside the platform engineering team's Kubernetes-native tooling
- Instrument the network stack for observability, feeding meaningful telemetry into fleet-wide monitoring and health tooling
- Maintain documentation and tooling that let new sites come online with minimal manual network configuration
- Support incident response with tooling that speeds up root-cause identification for network issues
What You Bring
- Experience building network automation for large-scale or multi-site network environments
- Strong software engineering fundamentals — Python essential; comfortable writing production-grade automation, not just scripts
- Working knowledge of network modeling, simulation, or verification tools (e.g. Batfish, or equivalent) and infrastructure-as-code approaches (e.g. Ansible, Nornir, NAPALM)
- Solid understanding of data centre networking fundamentals — spine-leaf design, BGP/EVPN, and RDMA-based fabrics (InfiniBand or RoCE)
- Experience integrating network automation into CI/CD pipelines
- Comfortable working from source-of-truth systems (e.g. NetBox or equivalent) to drive automated configuration
- Strong debugging instincts and a bias toward building tools that prevent repeat incidents
Nice to Have
- Experience with network digital twin or simulation platforms for pre-deployment testing
- Familiarity with high-performance/HPC networking and GPU cluster fabric requirements
- Background contributing to or operating network source-of-truth systems at scale
- Experience with observability tooling (Prometheus, Grafana, OpenTelemetry) applied to network telemetry
- Exposure to Kubernetes-native infrastructure and how network automation interfaces with platform operators