Jobs · Engineering · New York

Data Architect / Engineer

Appnovation · New York, United States · 2 days ago
EngineeringFull-time

Appnovation is a global, full-service digital partner that combines Strategy, Experience & Design, Engineering and Managed Services. We build digital solutions that deliver real impact today and serve as foundations for future growth. The technology department focuses on delivering software solutions that enable rich consumer experiences, from mobile and web applications to advanced analytics and machine learning, to content and engagement management service enablement platforms.

About the role

As a Data Architect / Engineer, you will join a highly motivated and experienced team and own both the data structures behind our graph-RAG platform and the pipelines that get content into it — modeling the knowledge and pumping it in at scale. You will define the schema, taxonomy and graph model for our knowledge domains, build resilient ingestion pipelines from enterprise content sources such as SharePoint and the wider Microsoft stack, and keep the knowledge graph accurate as source documents evolve. We are looking for people who can bring a strong, solution-focused mindset and contribute to architectural decisions, best practices, and get things done.

Responsibilities

  • Define the schema, taxonomy, and graph modeling for the knowledge domains.
  • Design how documents, versions, and their machine-readable representations are structured, and how facts evolve over time (including temporal replay).
  • Build ingestion/ETL pipelines from enterprise content sources (e.g., SharePoint and the wider Microsoft stack), handling multimodal content (presentations, documents, transcripts).
  • Keep the knowledge graph in sync as source documents change frequently.
  • Advise on whether the current backend approach (Supabase/Postgres) is the right long-term architecture, and validate ingested content quality with QA.
  • Establish data quality, lineage, and governance standards across pipelines.
  • Build monitoring, alerting, and error-handling for ingestion jobs.
  • Document data models, schemas, and pipeline architecture.
  • Optimize pipeline cost, throughput, and reliability over time.

Requirements

  • Bachelor’s Degree in Computer Science or a related field.
  • 5+ years of experience in data engineering / architecture roles.
  • Deep data modeling across relational and graph paradigms; graph databases (Neo4j) and Postgres/Supabase at scale.
  • Strong ETL / data pipeline engineering in Python, including large-scale, incremental, and change-data ingestion.
  • Experience with enterprise content ecosystems: SharePoint, ideally Power Platform / Power Automate.
  • Knowledge representation, taxonomies/ontologies, and versioned/temporal data.
  • Understanding of how data structures feed RAG and LLM retrieval quality.
  • SQL proficiency; experience with orchestration tools (e.g., Airflow) and cloud data services.
  • Familiarity with data governance, security, and privacy practices.
  • Strong documentation and cross-team communication skills.

Skills

  • Think about how to scale, automate, and operate, not just how to build a solution to an immediate problem.
  • Understand lean thinking.
  • Set high standards for code quality, performance/scalability, and security, and seek continuous improvement.
  • Solid analytical, problem-solving, and decision-making skills.
  • Customer-first mindset and a devotion to customer service.
  • Engage and build positive internal and external client relationships while managing multiple initiatives, often with competing priorities.
  • Strong self-initiative, passion, interpersonal, oral and written communication, and collaboration skills with the ability to work, influence, and make an impact in a cross-functional environment.
  • Responsive and thrive in a fast-paced, diverse, high-performance environment with rapidly changing business needs.
  • Actively seek out things outside your comfort zone with the ability to rapidly learn and take advantage of new concepts, business models, and technologies.
  • Prior experience in consulting.
  • Prior experience and connections in the Life Sciences industry is preferred.

Similar jobs

Data Engineer

Wolfe, LLCPittsburgh, PA· Yesterday
Information Technologyapply on apply.wolfe.com

Data Engineer

Ascot GroupDallas, TX· Yesterday
Information Technology$110k–$125k/yrapply on fa-emkq-saasfaprod1.fa.ocs.oraclecloud.com

Data Engineer

ProlaioChicago, IL· 2 days ago
Analyst$134k/yrapply on job-boards.greenhouse.io

Data Engineer

Evlo AINew York, NY· 2 days ago
RemoteEngineeringapply on evlo.ai

Data Engineer

Chenega MIOS SBUHuntsville, AL· 1 mo ago
apply on chenega.jibeapply.com

Data Engineer

HaystackWashington, DC· 1 mo ago
RemoteInformation Technology$62k–$141k/yrapply on haystack.cv