Palantir
Software Engineer - Hosted Model Infrastructure
About this role
Palantir seeks a Software Engineer to build and maintain infrastructure for deploying ML models across diverse, resource-constrained environments including air-gapped networks and edge devices. You'll own end-to-end services spanning inference engines, GPU scheduling, deployment pipelines, and platform integration to enable frontier AI capabilities for critical customers.
What you'll do
- Design and maintain ML model deployment infrastructure for air-gapped and edge environments
- Manage GPU scheduling, resource allocation, and inference engine optimization
- Build and support continuous deployment pipelines for models and capabilities
- Develop observability and monitoring solutions for production ML systems
- Integrate hosted model infrastructure with Palantir's broader platform
- Collaborate with customers to solve deployment challenges in constrained environments
What they're looking for
- Software engineering and full-stack development
- ML/AI infrastructure and deployment
- GPU scheduling and resource optimization
- Inference engine design or experience
- Continuous integration/continuous deployment (CI/CD)
- Observability and monitoring systems
- Cloud or on-premise infrastructure management
- Systems programming and optimization
Opens the application — the Jobs AI extension fills it for you. Set up autofill
Opens the official application on the employer’s site. No login required.
Palantir
Palantir builds data platforms and software solutions that help government and enterprise customers tackle complex operational challenges, with a focus on responsible AI governance and privacy. The company is hiring software engineers and interns for forward-deployed customer roles, infrastructure and platform teams, and specialized privacy and civil liberties engineering positions.
- Website
- palantir.com
Likely interview questions
- Describe your experience deploying ML models to production. How have you handled constraints like limited GPU resources or air-gapped environments?
- Walk us through how you've designed or maintained a service end-to-end, from inference to deployment and monitoring. What was your ownership model?