Skip to main content
← Back to Jobs

Software Engineer, Data Infrastructure

Scale AINew York, NY; Washington, DC
Full-timeMid Level
👁️ 1 views📝 0 applicationsPosted 6/23/2026Expires 9/29/2026
Tailor Resume for This JobCheck ATS Score

Get alerts for roles like this

More Software Engineer, Data Infrastructure roles in New York, NY; Washington, DC — straight to your inbox. No account needed.

Applying to this role? Tailor your résumé to this job description in one click, then download it clean — no watermark, no subscription.

Job Description

Scale AI is seeking a highly skilled and motivated Mission Software Engineer to join our dynamic Federal Engineering team. As a part of this team, you will play a critical role in supporting Scale’s government customers by scoping and developing onsite solutions. Our scalable, high-performance platform is the foundation for these customer solutions, and your expertise will be instrumental in designing and implementing systems that can handle interactions with existing customer systems to help our products integrate into existing customer workflows. The Role We are looking for an exceptional Senior Software Engineer to architect and build the foundational data infrastructure that will serve as the brain of a project ecosystem. We are not looking for someone to stitch together off-the-shelf data frameworks. You will be responsible for designing highly novel data models and processing pipelines capable of handling massive quantities of output data from complex simulations. At the core of this role is the challenge of building a foundational data ensemble —a unified architecture that seamlessly aggregates, structures, and stages diverse sources of simulation outputs and user inputs. Your systems will manage enormous batch throughput jobs with strict, minimal latency requirements, ensuring that downstream AI systems and language models have the exact context they need to actionably reason over complex, multi-dimensional scenarios. Key Responsibilities Architect the Data Ensemble: Design and implement the architecture to ensemble various sources of injected context (deeply structural simulation data, historical game states, and dynamic user inputs) into a unified, highly queryable format optimized for LLM consumption. Massive Batch Infrastructure: Build highly scalable, resil

Required Skills

C++JavaLeadershipLearning & DevelopmentMachine LearningOil & Gas OperationsProblem SolvingPythonRecruitmentRustSpark

The best-paying roles in your field. Every week. Free.

Join 10,000+ professionals getting job alerts and salary insights in their inbox

We respect your privacy. Unsubscribe anytime with one click.