← Back to job listings
SA
Software Engineer (Core Services)
Snorkel AI · San Francisco, United States
About The Role
Join Snorkel AI, a leading company in AI-native development. As a Software Engineer in the Core Services team, you will play a crucial role in shaping the data platform that powers the entire organization. You will be involved in a foundational architecture shift, designing and implementing event-driven data flows, building systems for data tracking and governance, and contributing to infrastructure cost visibility and optimization. This is an exciting opportunity to make a significant impact on the company's data flow and development workflow.
- Concevoir et mettre en œuvre des flux de données pilotés par des événements en utilisant des courtiers d'événements, des connecteurs CDC, un registre de schémas, un routage d'événements.
- Construire et maintenir la bibliothèque d'accès aux données partagées et les SDK utilisés par les équipes de la plateforme, de l'emballage et de l'API des ensembles de données.
- Contribuer à l'optimisation de la visibilité et des coûts de l'infrastructure - estimation des coûts des requêtes, dimensionnement des charges de travail.
- Hands-on experience with SQL and at least two of: Snowflake, Redshift, Postgres. You understand the performance characteristics of each and can write queries that don't bring down production
- Experience with Kubernetes. Our workloads run on EKS and you will deploy, debug, and scale services on K8s
- Experience with AWS — S3, RDS, EKS, EventBridge, IAM. Comfortable working in a Terraform-managed environment
- Familiarity with data orchestration tools (Prefect, Airflow, or Dagster) and transformation frameworks (dbt)
- Fluency with AI-assisted development tools (Claude Code, Cursor, or similar). This is a hard requirement — the team uses these tools daily and we expect engineers to leverage them for code generation, debugging, and investigation
- 4+ years building platform infrastructure, data infrastructure, data platforms, or backend systems with significant data components. You have built and operated pipelines, data access layers, or ETL/ELT systems in production
- Understanding of data governance concepts — RBAC, PII handling, audit logging, data lineage
- Strong proficiency in Python. Our stack is Python-heavy across Prefect, FastAPI, dbt, and the SDK layer
- Experience with Ray for distributed compute workloads
- Prior work in regulated environments (SOC 2, FedRAMP, HIPAA) where compliance requirements shaped system design
- Experience with OpenTelemetry, ClickHouse, or similar observability infrastructure
- Experience building shared libraries or SDKs consumed by multiple teams — versioning, backwards compatibility, migration support
- Experience with event-driven architectures — CDC, event buses, schema registries, at-least-once delivery semantics
This is an external listing. JobSpring does not represent or verify the employer. Report this listing
JobSpring