About Ray Data Team:Ray Data is Python-native data processing engine that is a one stop shop for all AI data processing needs. Ray Data provides performant, first-class integration with cutting edge AI frameworks using both multi-modal and structured data.
The Ray Data team currently develops and maintains Ray Data. We are a team of engineers passionate about building a Data processing engine which is a one-stop shop for all of your ML/AI needs. We are looking for exceptional engineers to build, optimize, and scale Ray for modern and increasingly complex AI workloads.
As part of this role, you will:- Improve the performance of Ray Data and multi-modal batch inference use cases.
- Ensure efficient scaling across different stages of the Data pipeline in a heterogeneous environment.
- Building data loading solutions for production training workloads.
- Focus on stability and fault tolerance at high scale
- Working with customers and new age AI native companies in scaling their AI workloads.
We'd love to hear from you if have:- At least 3-4 years of relevant work experience
- Solid background in building scalable and fault-tolerant distributed systems
- Experience with data processing, database internals.
- Passionate about large scale systems and performance for AI.