Skip to content
← Back to job listings

Senior Software Engineer (Cloud Infrastructure)

Altruist · San Francisco, United States

External listingfull-time2 months ago

About The Role

Join Altruist as a Senior Cloud Infrastructure Engineer. In this high-impact role, you will architect, build, and operate the AWS-based infrastructure that powers our broker-dealer and clearing platform. You will own critical infrastructure domains end-to-end, drive technical decisions, and shape infrastructure strategy. This is not a ticket-driven infrastructure role; you will be the technical authority for critical infrastructure areas and lead complex cross-functional initiatives.

  • Architect, build, and operate the AWS-based infrastructure that powers the broker-dealer and clearing platform.
  • Own critical infrastructure domains end-to-end and drive technical decisions that affect the reliability, security, and scalability of systems.
  • Define and enforce security architecture standards across AWS environments, and partner with Security, Compliance, and Audit teams to ensure infrastructure meets regulatory requirements.
  • AWS certifications at the Professional or Specialty level: Solutions Architect Professional, DevOps Engineer Professional, Security Specialty, or Machine Learning Specialty
  • Experience authoring RFCs, ADRs, or technical strategy documents that influenced engineering-wide decisions
  • 5+ years of hands-on experience in cloud infrastructure engineering, with deep, production-proven expertise in AWS
  • Extensive production experience operating Kubernetes (EKS strongly preferred) at scale — cluster lifecycle management, multi-tenancy patterns, Helm governance, and GitOps workflows
  • Proven experience designing and operating observability platforms (Datadog, Prometheus/Grafana, CloudWatch, OpenSearch/ELK) at organizational scale
  • Demonstrated experience with AI/ML infrastructure: provisioning GPU compute, SageMaker/Bedrock integration, vector databases, MLOps pipelines, or AIOps automation in production
  • Experience defining and executing organizational rollout strategies for AI developer tools, including governance frameworks, usage analytics, and cost management
  • Track record of owning and driving infrastructure initiatives end-to-end — from design and architecture through implementation, rollout, and operational excellence
  • Excellent technical communication skills — ability to write clear ADRs, present to leadership, and translate infrastructure complexity for non-technical stakeholders
  • Demonstrated ability to lead disaster recovery planning, execute DR simulations, and design HA architecture patterns for mission-critical systems
  • Hands-on experience with API gateway management and platform design (Kong, AWS API Gateway)
  • Experience with event streaming platforms at scale (Amazon MSK / Apache Kafka) including cluster operations, partition strategy, and consumer group management
  • 7+ years of infrastructure or platform engineering experience, including 3+ years operating at a senior or staff level
  • Strong Linux systems engineering skills and advanced scripting proficiency (Python, Bash, or Go)
  • Studies have shown that women and people of color are less likely to apply to jobs unless they meet every single qualification
  • Strong experience with database infrastructure and data layer architecture (Aurora PostgreSQL, RDS, ElastiCache/Redis, OpenSearch, DynamoDB)
  • Experience in financial services, fintech, broker-dealer, or other heavily regulated industries with FINRA/SEC compliance requirements
  • FinOps certification or demonstrated experience leading cloud cost optimization programs at scale
  • You may be just the right candidate for this or other roles
  • Contributions to open-source projects, conference talks, or published technical writing
  • At Altruist we are dedicated to building a diverse, inclusive, and authentic workplace, so if you’re excited about this role, but your past experience doesn’t align perfectly with every qualification in the job description, we encourage you to apply anyways
  • Expert-level understanding of cloud networking (VPC architecture, Transit Gateway, peering, DNS, load balancing) and security (IAM, KMS, WAF, GuardDuty, Secrets Manager)
  • Expert-level proficiency with Terraform, including module design, state management strategies, and establishing IaC standards for engineering teams
  • Deep expertise in CI/CD platforms (GitHub Actions, ArgoCD, Jenkins) with experience designing deployment strategies for multi-service architectures
  • Proven track record of mentoring engineers and elevating team capabilities through knowledge sharing, design reviews, and tooling improvements
  • Proficiency with policy-as-code frameworks (OPA/Rego, Sentinel, Kyverno) for infrastructure governance and compliance automation
  • Don’t meet every single requirement?

This is an external listing. JobSpring does not represent or verify the employer. Report this listing