Skip to content
← Back to job listings

Staff Data Engineer (Real World Evidence)

H1 · New York, United States

External listingfull-time3 months ago

About The Role

Join H1 as a Staff Data Engineer on the Real World Evidence (RWE) team. In this role, you will be a key technical leader, driving high-visibility data initiatives and reducing bottlenecks across teams. You will own the end-to-end architecture for critical data assets, design and optimize large-scale data pipelines, and partner with various teams to deliver high-value outcomes. You will also mentor other engineers and represent engineering in cross-functional forums.

  • Conduire des initiatives de données visibles et réduire les goulets d'étranglement entre les équipes.
  • Diriger des projets RWE de haute visibilité, en commençant par les données de réclamation, et maintenir plusieurs initiatives en mouvement.
  • Posséder l'architecture de bout en bout pour des actifs de données critiques, en veillant à ce que les solutions soient évolutives, fiables et alignées sur la vision à long terme de l'entreprise.
  • You’re a hands-on Staff IC and technical leader who thrives in complex data environments.
  • You bring clarity to ambiguity, turn messy problems into reliable systems, and operate with a strong sense of ownership and impact.
  • You collaborate effectively across functions, help others move faster, and are comfortable working across the full data and infrastructure stack
  • You invest in others through mentorship, pairing, and constructive feedback
  • You bring deep experience building and evolving large-scale data architectures, pipelines, or distributed systems
  • You operate well in high-ambiguity environments, make pragmatic trade-offs, and keep execution moving
  • You communicate clearly with both technical and non-technical partners and influence direction without needing formal authority
  • You raise the bar for engineering excellence through thoughtful design, high-quality code, and strong documentation
  • You have a proven track record of leading large, complex technical projects from concept to production
  • Strong programming experience in Python (or a modern language with the ability to quickly ramp up in Python)
  • Experience with large-scale data processing (e.g., Spark/PySpark on EMR or similar) or scalable distributed backend systems, with the ability to quickly deepen expertise in our data stack (PySpark, EMR, Hudi/Delta)
  • Hands-on experience with modern engineering workflows and tooling such as Git, JIRA, and CI/CD systems (e.g., CircleCI)
  • Experience with workflow orchestration or job scheduling tools (e.g., Airflow, Argo)
  • 8+ years as a software, data, or backend engineer building and operating scalable, production-grade systems
  • Comfort deploying and troubleshooting distributed workloads in cloud environments such as AWS EMR or Kubernetes
  • Experience using AI-assisted coding tools (e.g., GitHub Copilot, Claude Code) to accelerate development while maintaining quality is encouraged
  • Experience designing systems or large-scale datasets/pipelines with attention to performance, reliability, and maintainability
  • Strong proficiency in SQL, including writing and optimizing complex queries over large datasets
  • Experience with streaming/messaging technologies (e.g., Kafka, Kinesis) nice to have
  • Background in RWE, healthcare data, or other complex/regulated data domains is preferred
  • Demonstrated ability to independently drive complex, cross-team technical initiatives and influence stakeholders without formal authority

This is an external listing. JobSpring does not represent or verify the employer. Report this listing