Jobs · Quality Assurance · Texas

Sr. SysDev Eng, OTS - Data ANCHOR Team

Amazon · Austin, TX · 2 wk ago
Quality AssuranceFull-time

About the role

Are you a builder who thrives in ambiguity? Do you get energized turning proof-of-concepts into production-grade AI systems, fast and with a high bar? If you want to design agentic solutions that solve real operational problems at Amazon's global scale, this role is for you.

Responsibilities

  • Design and build production-grade agentic AI solutions — from agent orchestration to API integrations with ServiceNow (3P) and Amazon internal systems (1P)
  • Develop MCPs, agentic frameworks, and automation systems that enable scalable, reusable architectures across the Decision Intelligence portfolio
  • Partner with Data Scientists and Data Engineers to productionize ML models, build AI-ready data pipelines, and develop agent capabilities (LLM orchestration, prompt engineering, RAG, tool-use patterns)
  • Turn POCs into real, scalable production systems with a high quality bar — rapidly prototype, validate, and harden solutions for long-term deployment
  • Architect end-to-end systems that connect AI agents to enterprise workflows, enabling automated ticket processing, device management, and operational decision-making at scale
  • Build integrations across OTS and RME ecosystems, connecting agents to multiple upstream/downstream services, data sources, and streaming platforms
  • Design user-facing platforms and dashboards where AI integrations surface prescriptive recommendations to field operations partners
  • Design guardrails, evaluation frameworks, and safety mechanisms to ensure production AI systems operate reliably at scale
  • Raise the engineering bar — establish coding standards, drive code reviews, define architectural patterns, and mentor team members on best practices
  • Own the full development lifecycle — write design documents, implement solutions, build CI/CD pipelines, deploy to production, and monitor operational health
  • Participate in on-call rotations and drive operational excellence to maintain SLAs for production agents
  • Ensure we deploy long-term scalable agents and system integrations — not short-lived prototypes — through rigorous testing, observability, and production-readiness standards

Qualifications

  • Experience leading the design, automation, deployment, and support of large-scale infrastructure
  • Experience programming with at least one modern language such as Python, Ruby, Golang, Java, C++, C#, Rust 3+ years of designing or architecting (design patterns, reliability and scaling) of new and existing systems experience
  • Bachelor's degree in Computer Science, Engineering, a related field, or equivalent experience
  • Experience with designing and building applications using AWS services such as Lambda, AWS Elastic Beanstalk, Kubernetes
  • Experience in Kafka, or experience in software development and experience in any Bigdata architecture
  • Experience integrating with third-party APIs and enterprise platforms (e.g., ITSM tools, ServiceNow, or similar)
  • Experience building AI/ML-powered systems — LLM orchestration, prompt engineering, RAG architectures, agentic frameworks, or MCP integrations
  • Understanding of IT operations, reliability engineering, or maintenance workflows in large-scale operational environments
  • Track record of raising engineering standards — design docs, code review culture, reusable frameworks, and mentoring engineers
  • Comfort working in ambiguous, early-stage environments across geographically distributed teams where you must learn business processes to define technical solutions

Skills

  • Experience in designing and building application using AWS services such as Lambda, AWS Elastic Beanstalk, Kubernetes
  • Experience in Kafka, or experience in software development and experience in any Bigdata architecture
  • Experience integrating with third-party APIs and enterprise platforms (e.g., ITSM tools, ServiceNow, or similar)
  • Experience building AI/ML-powered systems — LLM orchestration, prompt engineering, RAG architectures, agentic frameworks, or MCP integrations
  • Understanding of IT operations, reliability engineering, or maintenance workflows in large-scale operational environments
  • Track record of raising engineering standards — design docs, code review culture, reusable frameworks, and mentoring engineers

Benefits

Amazon offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at https://amazon.jobs/en/benefits.

Pay

The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location.

Schedule

Final compensation will be determined based on factors including experience, qualifications, and location.

Similar jobs

Data Anchor

Ford Motor CompanyLouisville, KY· 1 mo ago
Marketing$85k–$193k/yrapply on careers.ford.com

Team Member 11586

Dhanani Group IncWestmont, IL· 6 days ago
Management$50k/yrapply on myjobs.adp.com

Team Member 25832

Dhanani Group IncSalisbury, MA· 6 days ago
Management$50k/yrapply on myjobs.adp.com

Team Member (4054)

Checkers & Rally’s Drive-In RestaurantsMarion, OH· 2 days ago
Manufacturingapply on apply.checkers.com

Team Member - 10202

Pollo TropicalFruitville, FL· 5 mo ago
Managementapply on frgi.wd1.myworkdayjobs.com