Skip to content
← Back to job listings

Staff Data Engineer

Butterfly Network · New York, United States

External listingfull-time11 days ago

About The Role

Join Butterfly Network as a Staff Data Engineer, where you will design and build the data layers that power our dashboards, data products, and AI agents. You will work on a small, high-ownership team and take ownership of ingestion pipelines, transformation layers, and the migration from GCP to AWS/Databricks. You will also translate business needs into data requirements and have the opportunity to make a significant impact in the healthcare industry.

  • Concevoir et construire des solutions de données robustes, en prenant en charge l'ensemble du pipeline d'ingestion et de transformation.
  • Diriger la migration de GCP vers AWS/Databricks, en prenant en charge les coupures de pipeline actives et en validant la parité.
  • Collaborer avec les équipes interfonctionnelles pour comprendre leurs flux de travail et leurs systèmes sources, et traduire leurs besoins en exigences de données.
  • Demonstrated experience with both streaming and batch ingestion patterns, including CDC pipelines, event-driven architectures, and scheduled bulk loads from operational sources
  • Hands-on experience integrating with operational data sources: CRM (Salesforce), ERP (NetSuite), payments (Stripe), or similar
  • Solid Python for data engineering: PySpark, pipeline development, utilities, and custom tooling
  • 5+ years of data engineering experience, with clear evidence of seniority: owned complex domains, designed solutions from scratch, and made architectural decisions — not just executed tickets
  • Experience with data quality, testing, and pipeline observability: dbt tests, Great Expectations, alerting, SLA tracking
  • Ability to design before building: write the doc, define the interface, identify the failure modes, then implement
  • Experience with Databricks (Delta Lake, Delta Live Tables) or a comparable lakehouse platform
  • Strong hands-on dbt experience: models, tests, sources, macros, documentation, CI integration, and refactoring existing work
  • Exposure to AI/ML platform patterns: vector search, RAG pipelines, or model serving data flows
  • Familiarity with reverse ETL or integration platforms (Mulesoft, Boomi, or equivalent)
  • Experience with Unity Catalog or an equivalent governance layer for column masking, row-level security, lineage, and sensitivity tags
  • Experience in healthcare, life sciences, or another regulated environment (HIPAA, PII/PHI classification and handling)
  • IaC experience at the data workload level (Terraform, Spacelift, or equivalents)

This is an external listing. JobSpring does not represent or verify the employer. Report this listing