Site Reliability Engineer

latent· Engineering
Apply Now ↗
📍 San FranciscoFullTime💰 USD 200K–275K/yr

About this role

SRE

Location: San Francisco, CA (5 Days In-Office)

You are the infrastructure expert who enables our rapid product development and guarantees 99.9%+ stability and performance of our clinical AI platform for major health systems. Your focus on operational excellence is directly tied to a patient's access to life-saving treatment.

What We Look for in a Great Engineer

You have the intensity and technical mastery to own mission-critical infrastructure. You hold yourself and others to high standards and thrive in a high-energy, in-office culture where everyone is in it to win it.

  • Tool Proficiency: You are highly proficient with your tools—you speak command line fluently and have mastered keyboard shortcuts.

  • Ownership: You thrive on owning complex systems and have a proven track record of scaling mission-critical deployments.

  • Automation Drive: You love automating things, always finding new ways to increase your own leverage, and defining standards for operational excellence.

  • Problem Solver: You won't wait for someone else to solve a problem that you're in a position to solve; you are willing to jump into whatever needs to get done.

What You'll Work On (Responsibilities)

As our SRE, you will own the entire production environment and improve the development experience:

  • Infrastructure Ownership: Design, implement, and maintain the production environment, having previously handled 500+ machine deployments.

  • Kubernetes Mastery: Own our containerized infrastructure, leveraging deep expertise in Kubernetes and Helm to manage deployment, scaling, and operational health.

  • CI/CD & Deployment Optimization: Optimize and streamline both the TypeScript and Python/ML deployment pipelines to support high-velocity feature release while maintaining the highest reliability.

  • DevX Support: Support Developer Experience (DevX) work to streamline developer workflows, enhance tool proficiency, and improve CI/CD systems.

  • Infrastructure as Code (IaC): Manage and maintain infrastructure definitions using Terraform.

Technical Qualifications & Environment

  • IaC & Orchestration: Deep, demonstrable experience with Kubernetes, Helm, and Terraform.

  • Scaling Systems: Proven ability to architect and maintain complex, distributed systems with high-availability requirements.

  • Deployment Experience: Hands-on experience optimizing deployment pipelines for both application code (TypeScript) and machine learning models (Python/ML). Also PostgreSQL, Redis, Kakfa.

  • Core Team Member: Excitement about working five days per week in our San Francisco office.

Frequently Asked Questions

What is the salary for the Site Reliability Engineer role at latent?
The listed salary for this Site Reliability Engineer position at latent is USD 200K–275K/yr. This is an FullTime role.
Where is the Site Reliability Engineer position at latent located?
This Site Reliability Engineer role at latent is based in San Francisco. The position is listed as on-site or hybrid. Check the full job description or apply directly to confirm the work arrangement.
Is the Site Reliability Engineer role at latent full-time or part-time?
This is listed as a FullTime position. It is posted as a Site Reliability Engineer role in the Engineering department at latent.
Which team or department does the Site Reliability Engineer at latent belong to?
This Site Reliability Engineer position is part of the Engineering department at latent. See the full job description for more information about the team structure and responsibilities.
How do I apply for the Site Reliability Engineer position at latent?
Click the "Apply Now" button on this page. You will be redirected to latent's official application portal hosted on ashby where you can submit your application directly.
When was the Site Reliability Engineer job at latent posted?
This Site Reliability Engineer position at latent was posted on Dec 5, 2025. Apply as soon as possible — early applications are often reviewed first.
Site Reliability Engineer
latent · 💰 USD 200K–275K/yr
Apply for this role ↗

You'll be redirected to latent's official application page on Ashby ATS.