jobloom

JobLoom finds jobs directly from company career sites before many job boards, then routes you into detailed role pages like this one.

infrastructure

Posted 2 weeks ago

Software Engineer, Data Infrastructure

at scaleai

New York, United StatesOn-site

Responsibilities

  • Architect the Data Ensemble: Design and implement the architecture to ensemble various sources of injected context (deeply structural simulation data, historical game states, and dynamic user inputs) into a unified, highly queryable format optimized for LLM consumption.

Requirements

  • Scale AI is seeking a highly skilled and motivated Software Engineer to join our dynamic Public Sector Engineering team.
  • Your systems will manage enormous batch throughput jobs with strict, minimal latency requirements, ensuring that downstream AI systems and language models have the exact context they need to actionably reason over complex, multi-dimensional scenarios. Key Responsibilities
  • Engineering Excellence: Deep, expert-level proficiency in systems languages (e.g., Rust, Go, C++, or highly optimized Python/Java, Spark, PySpark) and a fundamental understanding of memory management, compute limits, and distributed systems architecture.
  • You understand how to optimize massive batch jobs and parallel processing across distributed simulation nodes without sacrificing speed.
  • Information Retrieval & Context Surfacing: You don't need a background in AI agents, but you must be an expert in surfacing the right needle from an ocean of hay to feed decision-making engines.
  • Experience with LLM context optimization, vector embeddings, or agentic AI frameworks (e.g., advanced RAG architectures). Deep domain
  • experience working with wargaming data, complex systems modeling, or distributed simulation protocols. Previous
  • experience in a high-growth, 0-to-1 startup environment.
  • At Scale, our mission is to develop reliable AI systems for the world's most important decisions.
  • Our products provide the high-quality data and full-stack technologies that power the world's leading models, and help enterprises and governments build, deploy, and oversee AI applications that deliver real impact.
  • We are expanding our team to accelerate the development of AI applications.

Experience

  • Experience: 5+ years of backend or data infrastructure experience, operating at a Senior, Staff, or Principal level.

Benefits

  • Compensation packages at Scale for eligible roles include base salary, equity, and benefits.
  • The range displayed on each job posting reflects the minimum and maximum target for new hire salaries for the position and may be inclusive of several career levels at Scale; it will be determined during the interview process based on work location and additional factors, including job-related skills, experience, qualifications, interview performance, and relevant education or training.
  • Scale employees in eligible roles are also granted equity based compensation, subject to Board of Director approval.
  • Your recruiter can share more about the specific salary range for your preferred location during the hiring process, and confirm whether the hired role will be eligible for equity grant. You'll also receive
  • benefits including, but not limited to: comprehensive health, dental and vision coverage, retirement benefits, a learning and development stipend, and generous PTO.
  • benefits such as a commuter stipend.
  • For pay transparency purposes, the base salary range for this full-time position in the locations of San Francisco, New York, Seattle is: $186,400 — $233,000 USD
  • We comply with the United States Department of Labor's Pay Transparency provision .

Contact

  • If you need assistance and/or a reasonable accommodation in the application or recruiting process due to a disability, please contact us at accommodations@scale.com.

Additional details

  • As a part of this team, you will play a critical role in supporting Scale’s government customers by scoping and developing onsite solutions.
  • Our scalable, high-performance platform is the foundation for these customer solutions, and your expertise will be instrumental in designing and implementing systems that can handle interactions with existing customer systems to help our products integrate into existing customer workflows. The Role
  • We are not looking for someone to stitch together off-the-shelf data frameworks.
  • You will be responsible for designing highly novel data models and processing pipelines capable of handling massive quantities of output data from complex simulations.
  • At the core of this role is the challenge of building a foundational data ensemble —a unified architecture that seamlessly aggregates, structures, and stages diverse sources of simulation outputs and user inputs.
  • Massive Batch Infrastructure: Build highly scalable, resilient data architectures from scratch.
  • You will optimize for moving, transforming, and processing massive quantities of simulation output data via enormous batch jobs, maintaining the minimal latency required for rapid wargame iterations.
  • First-Principles Problem Solving: Navigate highly ambiguous product
  • requirements to design custom, ground-up systems where existing open-source or enterprise tools simply cannot handle the structural complexity or scale.
  • Technical Leadership: Set the technical standard for the data infrastructure team, driving rigorous code quality, system performance, and architectural clarity. What We’re Looking For •

Find more real-time jobs on JobLoom.