HomeJobsTogether AI › Engineering

Staff Software Engineer, Inference / Compute Infrastructure Engineering

Together AI San Francisco Full-time Engineering Posted Jul 22, 2026
Apply Now →

Culture signals in this specific role

Inferred from the employer’s wording in this job description. These role-level signals are separate from the company-wide employee-review scores below.

OwnershipThe description emphasizes ownership or end-to-end responsibility.
Cross-functionalThe role explicitly involves collaboration across teams or functions.
Deep technical workThe role highlights substantial technical, research, or systems work.
Operational responsibilityThe description includes production, reliability, or operational responsibility.

What it’s like to work at Together AI

AI Infrastructure · San Francisco

4.1
Employee Rating
3.8
Work-Life Balance
59
Open Roles
open-sourceeng-drivenlearningflatmany-hats

What employees love

  • Open-source AI infrastructure — real mission alignment, not just marketing
  • Small team with outsized research output; your work has direct impact

What could be better

  • Early-stage means wearing many hats constantly — not for specialists
  • Limited career ladder; growth paths are still being defined
View full Together AI culture profile →

About the Role

About the Role

We're looking for a Software Engineer to build the systems that treat infrastructure as software. This role owns the software state machines that provision hardware, bring it into service, and manage its full lifecycle — turning racks of GPUs into running inference clusters without a human touching a runbook. The Research and Inference team is your customer: today they file tickets and wait; the target state is that they issue a single API call to stand up, scale, or tear down a cluster, and the system takes care of the rest. The platform is manifest-driven such that teams declare the desired state of a cluster or host — shape, topology, software stack — and the system is responsible for reconciling reality to that manifest, continuously, through every stage of its lifecycle. You will design the engines that manifest the schema, the engines that execute against it, and the workflows that carry a piece of hardware or a cluster from one state to the next—taking it from bare metal to a fully functioning AI cluster for training or inference.

You'll write production code which is typed, tested, versioned, and deployed through CI/CD that models infrastructure state and reconciles it, the same way a Kubernetes controller reconciles a cluster's desired state. Success looks like eliminating manual provisioning work, not documenting it better.

A product mindset - you've built internal platforms or APIs consumed by other engineering teams and care about the developer experience of what you ship.You build it, you own it. You are not only responsible for delivering the software but also for operating and supporting it in production.

Responsibilities

Requirements

Core requirements (all levels):

Nice to have:

About Together AI

Together AI is a research-driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancement such as FlashAttention, Hyena, FlexGen, and RedPajama. We invite you to join a passionate group of researchers and engineers in our journey in building the next generation AI infrastructure.

Compensation

We offer competitive compensation, startup equity, health insurance and other competitive benefits. The US base salary range for this full-time position is: $240,000 - $280,000 + equity + benefits. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge.

Equal Opportunity

Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.

Similar Roles

More at Together AI
AI Infrastructure Engineer
San Francisco
AI infrastructure Engineer (SRE) Amsterdam
Amsterdam
AI infrastructure Engineer (SRE) Bangalore
Bangalore, India
Lead/Manager Site Reliability Engineering Team (Amsterdam)
Amsterdam
Lead/Manager Together Cloud Infrastructure
Amsterdam
Similar roles at other companies
Member of Technical Staff (Software Engineer, Monetization)
Perplexity AI · San Francisco
Principal Software Engineer, Developer Acceleration (Release Engineering)
HubSpot · Remote - USA
Member of Technical Staff, Backend Engineer, API
Runway · Remote
Staff Software Engineer, Product
Replit · Foster City, CA
Member of Technical Staff - Systems
Modal · New York

Frequently Asked Questions

What is the work-life balance like at Together AI?
Together AI has a work-life balance score of 3.8/5 based on employee reviews. This is about average for the AI/tech industry.
What is Together AI’s culture like?
Together AI is characterized by these culture values: open-source, eng-driven, learning, flat, many-hats. Based on employee reviews, the company has an overall rating of 4.1/5. Open-source AI infrastructure — real mission alignment, not just marketing
How many open roles does Together AI have?
Together AI currently has 59 open roles across departments including engineering, product, sales, and more. Roles are refreshed daily from their careers page.
Is this role remote-friendly?
This role is located in San Francisco. Check the job description above for specific location and remote work details.
Apply for this role at Together AI →