Software Engineer - Hosted Model Infrastructure
Palo Alto, CA Hybrid
$145,000–$200,000 a yearJobFig found this opening at its original source and checks that it remains available.
About the role
We are a software engineering team with expertise in enabling ML models in production. We deploy AI models to run in variety of environments: air-gapped government networks, forward-deployed defense environments, edge nodes, and enterprises with strict data sovereignty requirements. Our customers rely on us for frontier AI capabilities running on hardware they control, often with constrained GPU resources and limited direct access. Rising to that challenge and meeting those expectations is what Palantir's excels at. We treat models like any other software: continuously tested, continually delivered, packaged for reproducible deployment, and built for long-term maintainability. You will own services end-to-end, and work across the full stack, from inference engines, GPU scheduling to deployment pipelines, observability, and integration with Palantir's platform. The goal is to deliver new models and capabilities quickly and continuously. Join us if you want to solve problems at the intersection of infrastructure and machine learning that directly enable critical customers.
What you'll bring
- 4+ years of professional software engineering experience building and operating production systems
- Engineering background in Computer Science, Mathematics, Software Engineering, Physics, or similar field
- Strong coding skills with demonstrated proficiency in programming languages, such as Java, C++, Python, Rust, or similar languages.
- Familiarity with the Python ML ecosystem is valuable.
- Experience with containers, Kubernetes, and deploying backend services in production environments
- Strong written and verbal communication skills and ability to iterate quickly with teammates, incorporating feedback and holding a high bar for quality