Senior Software Engineer - Semantic Data Lake

wex

IndiaremotePosted 25 days ago
wex logo

Skill Required

Software-EngineerData-EngineeringSemantic-Data-PlatformAI-Native-EngineeringPlatform-EngineeringSenior-Data-EngineeringSenior-Big-Data-EngineerSenior-AI-ML-Data-EngineerSenior-Cloud-Data-EngineerSenior-Lead-Data-EngineeringData-Platform-EngineerData EngineerSoftware EngineerSoftware Developersoftware engineeringdata engineeringObservabilityEngineeringKubernetesautomationData StructuresTerraformdesigningsimilar)analyticssecuritybuildingEmbedded CAirflowPrometheusPythondesignDevOpsGitScalaCI/CDDesign PatternsRustRAGFulltime

Key highlights

  • Level Required: 4–8 years of experience in data engineering or software engineering

Role overview

WEX is reimagining its enterprise data platform to transform raw data into semantically meaningful, reusable, and trusted business assets. As a Senior Software Engineer on the Semantic Data Team, you'll contribute to designing, building, and maintaining the control plane for the Semantic layer and AI Native data platforms supporting core 360 data objects.

Responsibilities

  • Engineer the Operational Backbone: You will design and implement the core infrastructure of the Control Plane, serving as the central execution engine that orchestrates the entire data lifecycle—from ingestion and transformation to final distribution.
  • Shift from Reactive to Proactive: You will transform platform reliability by building sophisticated observability frameworks and automated quality gates, ensuring that data trust is engineered into the system rather than inspected after the fact.
  • Drive Self-Service Autonomy: You will eliminate operational bottlenecks by developing intuitive portals and "golden templates," empowering domain teams to autonomously build, manage, and consume trusted data products.
  • Champion Platform Governance: You will automate critical compliance and security policies, including end-to-end data lineage, role-based access control (RBAC), and PII protection, making security an inherent feature of our data objects.
  • Prepare for AI-Native Scale: You will build the foundational data architecture required for our AI future, contributing to the development of the context and RAG-based systems that will power the next generation of data-driven intelligence.
  • Leverage AI coding assistants (Claude, Copilot, Cursor, and similar) to accelerate development—drafting transformation logic, generating tests, refactoring pipelines, exploring datasets, and producing semantic documentation—while critically reviewing AI output for correctness and alignment with business rules.
  • Share patterns, prompts, and workflows that help the team get more leverage out of AI tooling, contributing to AI-native engineering practices across the Semantic Data Team.
  • Work closely with domain experts, data scientists, and product stakeholders to translate business concepts into framework that supports development
  • Implement logic for classifications, KPIs, scoring algorithms, and business rules, ensuring traceability and data lineage.
  • Follow and contribute to standards for data modeling, documentation, and governance within the semantic layer—including responsible, auditable use of AI-generated code and artifacts.
  • Collaborate across teams to integrate with ingestion, MDM, and data product layers.

Requirements

  • 4–8 years of experience in data engineering or software engineering with a focus on data transformation, modeling, or analytics platforms.
  • Strong proficiency in SQL and at least one general-purpose language such as Python or Scala.
  • Demonstrated experience as an AI-native engineer—using tools like Claude, GitHub Copilot, Cursor, or similar as a regular part of your development workflow.
  • AI-Native Architecture: You are passionate about the future of data intelligence and have hands-on experience with LLM-driven architectures. You are comfortable designing RAG systems, building agentic workflows (e.g., LangGraph, CrewAI), and working with vector/graph databases to manage context and ontology.
  • Familiarity with modern AI engineering practices such as prompt design, Spec-Driven Development (SDD), and AI-assisted code review.
  • Advanced Data Orchestration & Pipeline Reliability: You possess deep expertise in building and managing complex orchestration systems (specifically Airflow) and are obsessed with reliability. You know how to engineer "quality-first" pipelines using frameworks like Great Expectations to ensure data trust at scale.
  • Infrastructure as Code (IaC) & Platform Engineering: You are fluent in the modern DevOps stack—including Terraform, Kubernetes, and CI/CD automation—allowing you to build self-healing, scalable infrastructure that minimizes operational toil.
  • Enterprise Governance & Metadata Management: You have a proven track record of implementing automated governance, specifically utilizing DataHub or similar platforms to manage end-to-end lineage, cataloging, and RBAC policies that make security inherent to the data object.
  • Observability & Incident Management: You understand that "the platform is the product." You bring experience building comprehensive observability frameworks (e.g., Grafana, xMatters integration) that move teams from reactive firefighting to proactive, automated alerting.
  • Self-Service Mindset: You advocate for "developer autonomy" over manual intervention. You have experience building intuitive portals and developer tools that enable domain teams to build, manage, and consume trusted data products independently.
  • Solid understanding of data quality practices—including validation, enrichment, schema enforcement, and business rule encoding.
  • Comfort operating in a collaborative, cross-functional environment, balancing business logic with platform scalability.
  • A proven track record for traceability, reproducibility, and semantic clarity—you build data models others can trust and reuse.

Additional details

  • Senior Software Engineer, Semantic Data Team
  • Originally posted on Himalayas
Apply now