Senior Software Engineer, Platform
Overview
This is a hybrid role requiring 3 days a week. You must be work authorized in the United States without the need for employer sponsorship.
- Ad Tech / MarTech industry experience, specifically in e-commerce, travel, and finance.
Responsibilities
The Systems You Work On:
- Observability & monitoring — Prometheus, Alertmanager, Grafana, and OpenTelemetry across every production region.
- Kubernetes & cloud infrastructure — multi-region EKS clusters: upgrades, node group and Karpenter management, controller lifecycle, and add-on / configuration management.
- AWS networking — VPC design, subnet allocation and CIDR management, VPC peering, Transit Gateway, security groups, Route53, and NAT gateway topology across multiple accounts and regions, plus 24/7 networking alarms for all prod networking between clusters and squad resources.
- CI/CD & artifact management — GitLab administration (runner fleet, AMI updates, cache, access - not just pipeline authoring), GitOps delivery through ArgoCD, and the Nexus artifact repository including its storage lifecycle as it grows.
- Access & identity — Vault secrets management, IAM roles and service accounts for apps in clusters, cluster permission management for audit compliance, and AI model access management (including cost alerts and reporting).
- Cost observability — OpenCost, EBS orphan cleanup, cost anomaly investigation, and rightsizing attribution across teams, so waste is attributable rather than shared overhead.
- Internal tools & delivery — the container image build pipeline and base image standards, code audit tooling, HedgeDoc, and the UI CDN (S3 + CloudFront), plus adopted applications with no other owner.
Qualifications
- 6+ years of experience in software and/or infrastructure engineering.
- Bachelor's degree or equivalent experience.
- Deep, hands-on production experience operating Kubernetes and AWS at scale, across multiple accounts and regions.
- Real operational depth in at least one system we own beyond the cluster - most importantly the observability stack (Prometheus/Alertmanager at scale), but AWS networking, Vault, or artifact/CI infrastructure also count.
- We are filtering for people who have run these systems, not just used them.
- Strong AWS networking judgment (VPC, peering, Transit Gateway, subnet/CIDR design).
- A track record as a critical reviewer - spotting subtle infrastructure issues and long-term risks before they ship.
- Experience leading technical work and mentoring engineers; can manage, clarify, and plan around uncertainty.
- Effective communication and the ability to influence design in a product-focused way.
Nice To Have:
- Prometheus long-term storage / sharding (Thanos, Cortex, Mimir, or equivalent) run in production.
- Experience owning a container image / base image pipeline.
- Policy-as-code (Kyverno / OPA) and admission webhook design.
- Building claim-based self-service platform capabilities.
Additional Information
This is a hybrid role requiring 3 days a week in office. CJ is the leader in Performance Marketing. We take pride in our innovative technology, comprehensive data solutions and our people. We equip our teams with advanced tools, training and career development opportunities all to provide modern solutions, strategies and support to deliver high quality results for our clients. We work in an enthusiastic, collaborative team setting that values outstanding performance. We're a community of creative and passionate problem solvers who go the distance to tackle the tough questions, think creatively, and drive resourceful growth, for our clients—and ourselves. We foster and embody an inclusive and collaborative culture where diverse perspectives are sought, relationships are valued, and people feel accepted with a sense of belonging in expressing themselves authentically. We pride ourselves in having a workplace environment that values both work and play.
Compensation
Compensation Range: USD $110,580.00 - USD $166,430.00/Annually.