Skip to content
← Back to job listings

Software Engineer (Core Services)

Snorkel AI · San Francisco, United States

External listingfull-time2 months ago

About The Role

Join Snorkel AI, a leading company in AI-native development. As a Software Engineer in the Core Services team, you will play a crucial role in shaping the data platform that powers the entire organization. You will be involved in a foundational architecture shift, designing and implementing event-driven data flows, building systems for data tracking and governance, and contributing to infrastructure cost visibility and optimization. This is an exciting opportunity to make a significant impact on the company's data flow and development workflow.

  • Concevoir et mettre en œuvre des flux de données pilotés par des événements en utilisant des courtiers d'événements, des connecteurs CDC, un registre de schémas, un routage d'événements.
  • Construire et maintenir la bibliothèque d'accès aux données partagées et les SDK utilisés par les équipes de la plateforme, de l'emballage et de l'API des ensembles de données.
  • Contribuer à l'optimisation de la visibilité et des coûts de l'infrastructure - estimation des coûts des requêtes, dimensionnement des charges de travail.
  • Hands-on experience with SQL and at least two of: Snowflake, Redshift, Postgres. You understand the performance characteristics of each and can write queries that don't bring down production
  • Experience with Kubernetes. Our workloads run on EKS and you will deploy, debug, and scale services on K8s
  • Experience with AWS — S3, RDS, EKS, EventBridge, IAM. Comfortable working in a Terraform-managed environment
  • Familiarity with data orchestration tools (Prefect, Airflow, or Dagster) and transformation frameworks (dbt)
  • Fluency with AI-assisted development tools (Claude Code, Cursor, or similar). This is a hard requirement — the team uses these tools daily and we expect engineers to leverage them for code generation, debugging, and investigation
  • 4+ years building platform infrastructure, data infrastructure, data platforms, or backend systems with significant data components. You have built and operated pipelines, data access layers, or ETL/ELT systems in production
  • Understanding of data governance concepts — RBAC, PII handling, audit logging, data lineage
  • Strong proficiency in Python. Our stack is Python-heavy across Prefect, FastAPI, dbt, and the SDK layer
  • Experience with Ray for distributed compute workloads
  • Prior work in regulated environments (SOC 2, FedRAMP, HIPAA) where compliance requirements shaped system design
  • Experience with OpenTelemetry, ClickHouse, or similar observability infrastructure
  • Experience building shared libraries or SDKs consumed by multiple teams — versioning, backwards compatibility, migration support
  • Experience with event-driven architectures — CDC, event buses, schema registries, at-least-once delivery semantics

This is an external listing. JobSpring does not represent or verify the employer. Report this listing