Anyscale

Software Engineer, Ray Data

Anyscale
Apply
1 year ago
Bengaluru, IndiaSenior

Responsibilities

  • Develop high-quality open-source software for distributed programming with Ray.
  • Identify, implement, and evaluate architectural improvements to Ray Core and Ray Datasets.
  • Optimize Ray Datasets performance at large scale using Arrow primitives and Ray object-manager improvements.
  • Integrate Ray with machine-learning training systems and data sources.
  • Improve Ray’s testing process and stability infrastructure to support smooth releases.
  • Lead future streaming-workload integrations such as Beam on Ray.
  • Differentiate data operations in Anyscale’s hosted Ray service.
  • Communicate technical work through talks, tutorials, and blog posts.

Requirements

  • At least 5 years of relevant work experience.
  • Strong background in algorithms, data structures, and system design.
  • Experience building scalable and fault-tolerant distributed systems.
  • Experience with data processing and database internals, including Spark or Dask; streaming experience is a plus.

Tech Stack

Apache BeamApache SparkC++PythonXGBoost

Categories

BackendData Engineering
Anyscale

About Anyscale

501-1,000 employees

Anyscale builds a cloud platform and tools to run Ray, the open-source framework for distributed Python and AI/ML workloads, enabling teams to scale data prep, training, and inference. It monetizes through a managed service, enterprise features, and support for Ray deployments. Founded in 2019 and headquartered in San Francisco, this privately held company is the commercial steward of Ray, widely used to power production AI systems.

Contact me