Skip to main content

Anyscale

Software Engineer (Ray Core)

San Francisco (Remote)$215k–$265kfulltimemidAdded 1 month ago

About this role

Anyscale is hiring a Software Engineer to contribute to Ray Core, the C++ backend of an open-source distributed computing framework used by companies like OpenAI and Spotify. You'll work on performance optimization, fault tolerance, and architectural improvements while mentoring junior engineers and contributing to high-quality open-source software.

What you'll do

  • Develop and maintain Ray's C++ backend including distributed scheduler and runtime components
  • Optimize performance of large-scale distributed workloads and improve stability through stress testing
  • Lead cross-team projects on architectural improvements and fault tolerance enhancements
  • Improve testing infrastructure and release processes for Ray
  • Mentor junior team members on systems software best practices
  • Communicate work through talks, tutorials, and technical documentation

What they're looking for

  • C/C++ programming with low-level systems experience
  • Distributed systems design and fault-tolerant architecture
  • Algorithms and data structures
  • System design and performance optimization
  • GPU programming (preferred)
  • Distributed model training/inference (preferred)
  • Open-source software development
  • Testing infrastructure and quality assurance
Apply with Autofill

Opens the application — the Jobs AI extension fills it for you. Set up autofill

Opens the official application on the employer’s site. No login required.

Anyscale

Anyscale builds Ray, an open-source distributed computing framework and enterprise platform for scaling AI workloads across Kubernetes and cloud providers. The company is hiring forward-deployed engineers to work embedded with customers, software engineers to develop Ray Core, LLM inference specialists, and customer support engineers who combine technical expertise with post-sale success.

View all jobs at Anyscale

Likely interview questions

  • Can you walk us through a distributed system you've built or significantly contributed to? What were the key architectural decisions around scheduling, fault tolerance, or performance?
  • Describe your experience with C/C++ in systems programming. Have you worked on low-level OS components, memory management, or performance-critical code?