← Back to job listings
BN
Staff Data Engineer
Butterfly Network · New York, United States
About The Role
Join Butterfly Network as a Staff Data Engineer, where you will design and build the data layers that power our dashboards, data products, and AI agents. You will work on a small, high-ownership team and take ownership of ingestion pipelines, transformation layers, and the migration from GCP to AWS/Databricks. You will also translate business needs into data requirements and have the opportunity to make a significant impact in the healthcare industry.
- Concevoir et construire des solutions de données robustes, en prenant en charge l'ensemble du pipeline d'ingestion et de transformation.
- Diriger la migration de GCP vers AWS/Databricks, en prenant en charge les coupures de pipeline actives et en validant la parité.
- Collaborer avec les équipes interfonctionnelles pour comprendre leurs flux de travail et leurs systèmes sources, et traduire leurs besoins en exigences de données.
- Demonstrated experience with both streaming and batch ingestion patterns, including CDC pipelines, event-driven architectures, and scheduled bulk loads from operational sources
- Hands-on experience integrating with operational data sources: CRM (Salesforce), ERP (NetSuite), payments (Stripe), or similar
- Solid Python for data engineering: PySpark, pipeline development, utilities, and custom tooling
- 5+ years of data engineering experience, with clear evidence of seniority: owned complex domains, designed solutions from scratch, and made architectural decisions — not just executed tickets
- Experience with data quality, testing, and pipeline observability: dbt tests, Great Expectations, alerting, SLA tracking
- Ability to design before building: write the doc, define the interface, identify the failure modes, then implement
- Experience with Databricks (Delta Lake, Delta Live Tables) or a comparable lakehouse platform
- Strong hands-on dbt experience: models, tests, sources, macros, documentation, CI integration, and refactoring existing work
- Exposure to AI/ML platform patterns: vector search, RAG pipelines, or model serving data flows
- Familiarity with reverse ETL or integration platforms (Mulesoft, Boomi, or equivalent)
- Experience with Unity Catalog or an equivalent governance layer for column masking, row-level security, lineage, and sensitivity tags
- Experience in healthcare, life sciences, or another regulated environment (HIPAA, PII/PHI classification and handling)
- IaC experience at the data workload level (Terraform, Spacelift, or equivalents)
This is an external listing. JobSpring does not represent or verify the employer. Report this listing
JobSpring