The opportunity
A World-Changing Company Palantir builds the world’s leading software for data-driven decisions and operations. By bringing the right data to the people who need it, our platforms empower our partners to develop lifesaving drugs, forecast supply chain disruptions, locate missing children, and more.
What you'll do
Different backend languages, including Java, Rust, Python and Go
Model serving engines for GPU-accelerated inference
Docker and Kubernetes for containerization and orchestration
Industry-standard build tooling, including Gradle and GitHub
Building high-performance model serving infrastructure that integrates with: security models, hardware constraints, and different inference engines
Designing intelligent request handling including authentication, rate: limiting, concurrency control, and audit logging for multi-tenant model access
What they're looking for
- Building and maintaining packaging and deployment pipelines enabling fast,: secure, and reliable model rollouts across on-premises and air-gapped environments
- Developing observability for production AI systems to enable easy service: monitoring and fast incident triage and resolution
- Debugging complex issues and performance problems throughout the stack,: including open source inference engines, container runtimes, and GPU drivers, in environments you cannot always access directly
- Designing and running testing and benchmarking infrastructure that validates: model deployments across varying GPU hardware before they reach production