About the role
We're hiring a Senior Data Engineer to own the governed data foundation that powers our AI and analytics work. You'll design the pipelines, semantic layer, and vector indexes that feed LLM systems and reporting alike — and partner directly with the AI practice to make data trustworthy, fast, and well-documented.
What you'll do
- Design and operate batch and streaming pipelines that move enterprise data reliably at scale.
- Build and maintain a governed semantic layer and the vector indexes that feed retrieval systems.
- Establish data quality, lineage, and observability standards across the platform.
- Partner with AI/ML engineers to shape how data is modeled and served to production models.
- Mentor engineers and set technical direction for the data team.
What you'll bring
- 5+ years of data engineering experience, including production pipeline ownership.
- Strong SQL and Python; deep experience with a modern warehouse (Snowflake, BigQuery, Redshift) and orchestration (Airflow, dbt).
- Hands-on work with data modeling, governance, and pipeline reliability.
- Clear communication with both engineers and non-technical stakeholders.
Nice to have
- Experience building vector / embedding pipelines for AI use cases.
- Familiarity with streaming tools (Kafka, Spark Structured Streaming).
- Background working with regulated or sensitive data.