Turn on analytics and we load Google Analytics: Google gets the pages you open and what you do here — searches, jobs you view, jobs you apply to — and sets its own cookies. Leave it off and the only cookies we set are your login, your theme, and this answer. Accept All also records a yes to advertising, which nothing uses yet. Privacy Policy.
142 open listings · Page 3 of 6
Roles posted by organizations here, alongside roles we found on employers’ own careers sites. Crawled roles say so on the card and send you to the employer to apply. Only show organizations on Career App
Mastercard
Full-time · O Fallon, MO · $122,000 – $207,000
Lead a team of Site Reliability Engineers ensuring Mastercard's Java and Spring-based systems remain stable, scalable, and performant across global operations. This role suits experienced SREs who want to mentor others while driving automation, operational excellence, and a culture of reliability.
Listed on Mastercard’s careers site · Apply there ↗
iHeartMedia
Full-time · San Antonio, TX
iHeartMedia is seeking a senior-level reliability engineer to lead a team managing cloud infrastructure and system performance across their massive audio platform. This role suits experienced engineers who thrive in leadership, building high-performing teams, and ensuring stability at scale.
Listed on iHeartMedia’s careers site · Apply there ↗
Blackstone Group
Full-time · New York, NY · $140,000 – $225,000
Blackstone seeks a Site Reliability Engineer to lead observability and resilience practices across their infrastructure and development teams. This role combines hands-on incident response with strategic initiatives to improve system reliability, operational efficiency, and on-call experience through automation and SRE methodologies.
Listed on Blackstone Group’s careers site · Apply there ↗
Axiom Space
Houston, TX
This role develops reliability strategies and conducts failure analysis for commercial space station systems, working across engineering teams to predict and prevent component failures throughout orbital operations. It suits experienced aerospace engineers with deep spacecraft reliability expertise who want to contribute to humanity's expansion into space.
Listed on Axiom Space’s careers site · Apply there ↗
Comcast
2 locations · $109,178 – $163,768
This role manages the production infrastructure powering FreeWheel's high-volume advertising platform, requiring expertise in cloud infrastructure, Kubernetes, and automation to support business-critical systems serving major media clients. It suits engineers experienced with AWS, infrastructure-as-code, and on-call operational support who thrive solving complex distributed systems problems at scale.
Listed on Comcast’s careers site · Apply there ↗
Rogue Fitness
Columbus, OH
Rogue Fitness seeks a Site Reliability Engineer to design and maintain highly available infrastructure supporting their retail and manufacturing systems. This role suits someone with deep cloud and container experience who can automate deployments, manage incidents, and collaborate with development teams to keep systems running reliably.
Listed on Rogue Fitness’s careers site · Apply there ↗
CrowdStrike
$140,000 – $215,000
A senior systems engineer role focused on building and maintaining the reliability infrastructure that powers CrowdStrike's massive-scale cloud security platform. This suits experienced backend engineers who want to shape architecture across product teams and work on distributed systems challenges without operational ticket management.
Listed on CrowdStrike’s careers site · Apply there ↗
Cognition AI
Full-time · New York, NY · $260,000 – $300,000
This role owns production reliability and platform engineering for Devin and Windsurf, AI developer tools used by hundreds of thousands daily. You'll define SLOs, lead incident response, build CI/CD pipelines, and partner with product teams to engineer reliability from the start—combining on-call ownership with infrastructure-as-code and observability work.
Listed on Cognition AI’s careers site · Apply there ↗
Relativity Space
Long Beach, CA · $128,000 – $192,000
This role leads a build reliability team at a rocket manufacturer, responsible for embedding quality and risk management into vehicle design and production. It suits experienced manufacturing engineers with people leadership skills who want to scale reliability processes from development into high-rate production.
Listed on Relativity Space’s careers site · Apply there ↗
MyFitnessPal
Full-time · Remote - US · $120,000 – $165,000
This role leads site reliability and CI/CD infrastructure for a health and fitness platform at scale, owning production reliability, incident response, and security in the delivery pipeline. It suits senior engineers with hands-on experience in Kubernetes, cloud platforms, and automation who want to influence system architecture and cross-team operational practices.
Listed on MyFitnessPal’s careers site · Apply there ↗
Cohere Health
Full-time · United States · $100,000 – $110,000
A healthcare infrastructure role focused on keeping cloud-based systems running reliably, split between responding to live incidents and building automation to reduce repetitive work. This suits someone with strong AWS and Python skills who has managed production data pipelines and enjoys the intersection of operations and infrastructure engineering.
Listed on Cohere Health’s careers site · Apply there ↗
Obsidian Securty
Full-time · Palo Alto, CA · $165,000 – $190,000
This role maintains and scales Obsidian's SaaS security platform infrastructure across AWS and GCP, working with Kubernetes, CI/CD systems, and observability tools. It suits someone with several years of cloud operations experience who wants to embed deeply with engineering teams and own reliability challenges.
Listed on Obsidian Securty’s careers site · Apply there ↗
Supabase
Full-time · Remote, Global
A distributed SRE role focused on establishing reliability practices and frameworks across Supabase's engineering organization rather than owning infrastructure directly. This suits someone with deep SRE experience who wants to drive systemic improvements and influence multiple teams through expertise and collaboration.
Listed on Supabase’s careers site · Apply there ↗
InstaLily
Full-time · New York, NY · $150,000 – $190,000
InstaLILY is hiring a Site Reliability Engineer to build and operate their Internal Developer Platform and multi-cloud Kubernetes infrastructure that supports their AI agent products. You'll work in a small team with direct impact, partnering with engineers to create self-service tooling and golden paths while managing production systems that power major enterprise customers.
Listed on InstaLily’s careers site · Apply there ↗
Eight Sleep
Full-time · San Francisco, CA · $150,000 – $180,000
Lead hardware reliability efforts at a sleep technology company, driving testing strategies and design validation for electromechanical systems that must perform reliably in real-world conditions. This role suits engineers with 5+ years of hands-on reliability experience who want to own the full lifecycle of hardware reliability from design through production.
Listed on Eight Sleep’s careers site · Apply there ↗
The New York Times Company
Full-time · New York, NY · $120,000 – $142,000
This role leads product strategy for site reliability programs across The New York Times' infrastructure, working with engineering teams to build platforms and practices that help services operate reliably at scale. It suits someone with technical product experience in platform or infrastructure domains who understands SRE practices and wants to influence how organizations approach operational readiness and observability.
Listed on The New York Times Company’s careers site · Apply there ↗
Bend Studio
Full-time · Aliso Viejo, CA · $145,700 – $218,500
A Site Reliability Engineer focused on cloud gaming infrastructure, managing API gateways, service meshes, and distributed systems at scale for PlayStation's gaming platform. This role suits engineers with strong Linux and production systems experience who want to own reliability and operational excellence across a high-traffic gaming service.
Listed on Bend Studio’s careers site · Apply there ↗
Palantir
Full-time · 2 locations · $96,000 – $140,000
This role focuses on maintaining the reliability and performance of Palantir's production systems, combining on-call incident response with longer-term infrastructure and stability improvements. It suits engineers who enjoy troubleshooting complex problems, take ownership of end-to-end service health, and want to balance reactive firefighting with proactive resilience work.
Listed on Palantir’s careers site · Apply there ↗
Cohere
Full-time · New York, NY · $160,000 – $260,000 · Remote
This role develops and operates the infrastructure that serves Cohere's language models at scale, focusing on Kubernetes automation, GPU clusters, and high-availability systems. It suits engineers with production infrastructure experience who want to work on the platform layer enabling enterprise AI deployment.
Listed on Cohere’s careers site · Apply there ↗
Astranis
San Francisco, CA · $135,000 – $235,000
Astranis seeks a senior reliability engineer to ensure their spacecraft electronics meet demanding operational standards over multi-year missions in geosynchronous orbit. This role suits someone with deep experience in physics of failure and design for reliability who can lead analysis, testing strategies, and failure investigations for complex space systems.
Listed on Astranis’s careers site · Apply there ↗
Alcoa
AU PTL Portland
Listed on Alcoa’s careers site · Apply there ↗
NVIDIA
Santa Clara, CA
Listed on NVIDIA’s careers site · Apply there ↗
The Home Depot
GEORGIA - VIRTUAL - GA01 · Remote
Listed on The Home Depot’s careers site · Apply there ↗
Samsung Semiconductor
Taylor, TX
Listed on Samsung Semiconductor’s careers site · Apply there ↗
Boeing
USA - Berkeley, MO
Listed on Boeing’s careers site · Apply there ↗