HomeJobsAnyscale › Engineering

Distributed LLM Inference Engineer

Anyscale San Francisco FullTime Engineering Posted 3w+ ago
Apply Now →

Culture signals in this specific role

Inferred from the employer’s wording in this job description. These role-level signals are separate from the company-wide employee-review scores below.

OwnershipThe description emphasizes ownership or end-to-end responsibility.
Deep technical workThe role highlights substantial technical, research, or systems work.
Role contextThis is a engineering role; confirm team-specific working norms during the interview.

What it’s like to work at Anyscale

The company behind Ray · San Francisco

4.2
Employee Rating
3.8
Work-Life Balance
19
Open Roles
eng-drivenopen-sourceflatship-fast

What employees love

  • Brilliant colleagues from top companies
  • Direct impact on AI/ML ecosystem

What could be better

  • Context switching between projects
  • Documentation can lag behind development
View full Anyscale culture profile →

About the Role

About Anyscale

At Anyscale, we're on a mission to democratize distributed computing and make it accessible to software developers of all skill levels. We’re commercializing Ray, a popular open-source project that's creating an ecosystem of libraries for scalable machine learning. Companies like OpenAIUberSpotifyInstacartCruise, and many more, have Ray in their tech stacks to accelerate the progress of AI applications out into the real world.


With Anyscale, we’re building the best place to run Ray, so that any developer or data scientist can scale an ML application from their laptop to the cluster without needing to be a distributed systems expert.


Proud to be backed by Andreessen Horowitz, NEA, and Addition with $250+ million raised to date.


About the role

As a Distributed LLM Inference Engineer, you will help systems and optimizations that push the boundaries of performance for inference at large scale. This is an incredibly critical role to Anyscale as it allows us to achieve a market leading position for  AI infrastructure.

As part of this role, you will

We'd love to hear from you if you have

Bonus points!

Compensation

At Anyscale, we take a market-based approach to compensation. We are data-driven, transparent, and consistent.  As the market data changes over time, the target salary for this role may be adjusted.

This role is also eligible to participate in Anyscale's Equity and Benefits offerings, including the following:

Anyscale Inc. is an Equal Opportunity Employer. Candidates are evaluated without regard to age, race, color, religion, sex, disability, national origin, sexual orientation, veteran status, or any other characteristic protected by federal or state law. 

Anyscale Inc. is an E-Verify company and you may review the Notice of E-Verify Participation and the Right to Work posters in English and Spanish

Similar Roles

More at Anyscale
Software Engineer, Ray Core
Bengaluru, Karnataka
Software Engineer (Ray Data)
San Francisco
Software Engineer, Infrastructure
San Francisco
Software Engineer (Ray Core)
San Francisco
Software Engineer, Platform Infrastructure (Foundations)
Bengaluru, Karnataka

Frequently Asked Questions

What is the work-life balance like at Anyscale?
Anyscale has a work-life balance score of 3.8/5 based on employee reviews. This is about average for the AI/tech industry.
What is Anyscale’s culture like?
Anyscale is characterized by these culture values: eng-driven, open-source, flat, ship-fast. Based on employee reviews, the company has an overall rating of 4.2/5. Brilliant colleagues from top companies
How many open roles does Anyscale have?
Anyscale currently has 19 open roles across departments including engineering, product, sales, and more. Roles are refreshed daily from their careers page.
Is this role remote-friendly?
This role is located in San Francisco. Check the job description above for specific location and remote work details.
Apply for this role at Anyscale →