Habeo is building the IT asset management and CMDB platform designed specifically for higher education, replacing ServiceNow ITAM with a system that integrates natively with Jamf, Intune, Google Admin, and Workday. We're looking for a Data Engineer to help build and scale the data infrastructure that powers our asset ledger, CMDB relationship graph, and continuous device discovery pipelines from multiple source systems.
What you'll do
- Design, build, and maintain data pipelines that ingest continuous device and identity data from sources like Jamf, Intune, Google Admin, and Workday
- Develop and optimize the data models underlying the asset ledger and CMDB relationship graph
- Ensure data quality, consistency, and auditability across ingestion, transformation, and storage layers
- Build and maintain ETL/ELT processes that support real-time and batch data needs
- Collaborate with product and engineering teams to expose reliable, well-structured data to application and reporting layers
- Monitor pipeline performance and reliability, and troubleshoot data issues as they arise
What we're looking for
- Professional experience building and operating production data pipelines
- Strong SQL skills and experience with relational and/or graph data models
- Experience with a modern data stack (e.g., orchestration tools, warehousing, streaming or batch processing frameworks)
- Proficiency in a programming language commonly used for data engineering, such as Python, Java, or Scala
- Experience integrating with third-party APIs and handling data from multiple heterogeneous sources
- Strong understanding of data quality, testing, and observability practices
Nice to have
- Experience working with immutable or append-only audit log architectures
- Familiarity with identity and directory data (SSO, SCIM, or similar)
- Experience in a startup or fast-growing engineering team
- Exposure to compliance-driven data handling requirements (e.g., FERPA, SOC 2)