The opportunity
At Anyscale , we're on a mission to democratize distributed computing and make it accessible to software developers of all skill levels. We’re commercializing Ray , a popular open-source project that's creating an ecosystem of libraries for scalable machine learning.
What you'll do
Improve the performance of Ray Data and multi-modal batch inference use cases.
Ensure efficient scaling across different stages of the Data pipeline in a heterogeneous environment.
Building data loading solutions for production training workloads.
Focus on stability and fault tolerance at high scale
Working with customers and new age AI native companies in scaling their AI workloads.
At least 3-4 years of relevant work experience
What they're looking for
- Solid background in building scalable and fault-tolerant distributed systems
- Experience with data processing, database internals.
- Passionate about large scale systems and performance for AI.