×
Register Here to Apply for Jobs or Post Jobs. X

Data Engineer

Job in Northern, Floyd County, Kentucky, USA
Listing for: Harbor Compliance
Full Time position
Listed on 2026-09-12
Job specializations:
  • Software Development
    Data Engineering
Salary/Wage Range or Industry Benchmark: 172000 - 215000 USD Yearly USD 172000.00 215000.00 YEAR
Job Description & How to Apply Below

Harbor Compliance is building its first end-to-end data platform — unifying fragmented data from Hub Spot, financial systems, and HRIS into a single, AI-augmented source of truth for executive decision‑making. As Staff Data Engineer, you'll be a foundational technical hire, working closely with the Sr. Director of Data Platform & Analytics to design and build the real‑time, AI‑ready data infrastructure that powers this platform.

This is a hands‑on role for someone who wants to build from the ground up, not maintain what already exists.

Key Responsibilities
  • Design, build, and own near real‑time data pipelines (CDC, streaming ingestion, event‑driven architectures) as the backbone of the platform's data flow
  • Evaluate, implement, and maintain vector database infrastructure and embedding pipelines to support AI-augmented use cases (semantic search, retrieval‑augmented generation, AI agents acting on company data).
  • Design and build ELT/ETL pipelines ingesting data from our platform, financial platforms, CRM (Hub Spot) and HRIS, feeding both real‑time and batch use cases.
  • Partner with the Sr. Director to architect the underlying warehouse/lakehouse as a supporting system of record — the storage layer beneath the streaming and AI infrastructure.
  • Build lightweight transformation layers (e.g., dbt) as needed to enable our Analytics Engineering team translate raw data into business‑ready datasets aligned to core metrics like ARR, CAC, and churn.
  • Own pipeline reliability and observability — monitoring, automated failure alerting, and lineage tracking across both streaming and batch pipelines.
  • Build the technical foundation for self‑service and AI‑powered reporting, partnering with BI, Product & Engineering stakeholders on recurring executive and departmental reports.
  • Implement data governance practices, including documentation standards and access controls.
  • Partner cross‑functionally with Finance, Marketing, Customer Success, and Operations to translate data needs into reliable, low‑latency data products.
  • Leverage AI‑augmented development workflows (e.g., Claude Code) to accelerate pipeline development and documentation.
Requirements
  • 7+ years of hands‑on data engineering experience, with meaningful depth in streaming/event‑driven systems, not just batch pipelines.
  • Proven experience designing and building near real‑time pipelines from scratch (e.g., Kafka, Kinesis, Flink, Debezium/CDC) in a production environment.
  • Hands‑on production experience with vector databases and embeddings (e.g., Zilliz, Pinecone, Weaviate, pgvector, Milvus) — ideally having built this infrastructure from the ground up rather than just consuming a managed AI feature.
  • Advanced proficiency in SQL and Python.
  • Working knowledge of a cloud warehouse/lakehouse platform (Snowflake, Big Query, or Databricks) and dbt — you'll use these, but they're the storage/transform layer supporting the streaming and AI work, not the main focus.
  • Proven experience building or materially contributing to an end‑to‑end production data environment, ideally as an early or founding data hire.
  • Familiarity with B2B SaaS and recurring revenue data models (customer lifecycle, pipeline/conversion data).
  • Working knowledge of BI/reporting tools (e.g., Looker, Tableau, Power BI) as a downstream consumer of your data models.
  • Ability to work independently and drive multi‑stakeholder projects in a lean, scrappy, fast‑moving environment — comfortable with ambiguity and building without a lot of existing infrastructure or process.
Skills and Knowledge
  • Strong command of event streaming and CDC tooling.
  • Hands‑on experience with vector databases (Pinecone, Weaviate, pgvector, Milvus, or similar), including embedding strategies and chunking approaches for retrieval use cases.
  • Working…
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary