Skip to content
← Back to job listings

Software Engineer (Data Infrastructure)

Scale AI · Washington, United States

External listingfull-timeabout 1 month ago

About The Role

Join Scale AI, a leading company in the AI industry, as a Software Engineer specializing in Data Infrastructure. In this role, you will be a key member of the Public Sector Engineering team, responsible for designing and implementing systems that integrate with existing customer workflows. You will architect and build the foundational data infrastructure for complex simulations, ensuring minimal latency and optimal performance. This position offers a unique opportunity to work on mission-driven technology that supports national security and defense modernization.

  • Architect the Data Ensemble: Design and implement the architecture to ensemble various sources of injected context into a unified, highly queryable format optimized for LLM consumption.
  • Build highly scalable, resilient data architectures from scratch, optimizing for moving, transforming, and processing massive quantities of simulation output data.
  • Design sophisticated, highly relational data models that accurately represent massive, state-based simulation environments, making them easily interpretable by machine learning models.
  • Information Retrieval & Context Surfacing: You don't need a background in AI agents, but you must be an expert in surfacing the right needle from an ocean of hay to feed decision-making engines. We highly value engineers with backgrounds in:
  • High-Throughput / Low-Latency Data: Proven track record of processing massive datasets. You understand how to optimize massive batch jobs and parallel processing across distributed simulation nodes without sacrificing speed
  • Mission-Driven: A strong desire to build robust, foundational technology that supports national security and defense modernization
  • Engineering Excellence: Deep, expert-level proficiency in systems languages (e.g., Rust, Go, C++, or highly optimized Python/Java, Spark, PySpark) and a fundamental understanding of memory management, compute limits, and distributed systems architecture
  • High-Frequency Trading (HFT): Processing disparate, massive streams of data for algorithmic decision-making
  • Experience: 5+ years of backend or data infrastructure experience, operating at a Senior, Staff, or Principal level
  • Gaming / MMOs: Managing complex state, data relationships, and telemetry for massive, highly populated simulations
  • Search & RecSys: Building complex information retrieval systems or recommendation engines
  • Experience with LLM context optimization, vector embeddings, or agentic AI frameworks (e.g., advanced RAG architectures)
  • Security Clearance: An active Secret or TS/SCI clearance is a nice to have for this role. If you do not have an active clearance, you must be eligible and willing to obtain one
  • Previous experience in a high-growth, 0-to-1 startup environment
  • Deep domain experience working with wargaming data, complex systems modeling, or distributed simulation protocols

This is an external listing. JobSpring does not represent or verify the employer. Report this listing