Turn on analytics and we load Google Analytics: Google gets the pages you open and what you do here — searches, jobs you view, jobs you apply to — and sets its own cookies. Leave it off and the only cookies we set are your login, your theme, and this answer. Accept All also records a yes to advertising, which nothing uses yet. Privacy Policy.
108 open listings · Page 2 of 5
Roles posted by organizations here, alongside roles we found on employers’ own careers sites. Crawled roles say so on the card and send you to the employer to apply. Only show organizations on Career App
Together AI
Full-time · Remote · $160,000 – $230,000
A technical support engineer who handles customer issues with AI inference and fine-tuning platforms, working weekend shifts after an initial ramp-up period. The role bridges customer needs and internal engineering teams, requiring deep infrastructure expertise and hands-on troubleshooting of GPU clusters and Kubernetes deployments.
Listed on Together AI’s careers site · Apply there ↗
Intuit
2 locations · $202,500 – $274,000
This is a senior site reliability engineering role focused on maintaining highly available fintech infrastructure serving millions of small business users. The position combines hands-on systems design, automation, and incident leadership within a distributed cloud environment at scale.
Listed on Intuit’s careers site · Apply there ↗
Obsidian Securty
Philadelphia, PA · $223,000 – $239,000
A Staff Software Engineer role leading cross-team platform and product work at a SaaS security company protecting enterprise applications. This suits experienced engineers who thrive in ambiguity, design systems at scale, and can influence technical direction across multiple teams.
Listed on Obsidian Securty’s careers site · Apply there ↗
Crusoe
Full-time · New York, NY · $155,000 – $200,000
This role involves helping enterprise customers successfully deploy AI and machine learning workloads on Crusoe's GPU infrastructure, moving from technical discovery through proof-of-concept to production launch. It's suited to cloud infrastructure engineers who want customer-facing technical responsibility and hands-on ownership of AI infrastructure challenges.
Listed on Crusoe’s careers site · Apply there ↗
Supabase
Full-time · Remote, Global
This role owns the reliability of Supabase's deployment pipelines and control plane, treating them as production systems with SLOs and on-call ownership. It suits someone with deep SRE experience who has operated systems at scale and wants to eliminate deployment toil through automation, observability, and incident-driven improvement.
Listed on Supabase’s careers site · Apply there ↗
Kong Company
Washington, United States · $153,000 – $218,000 · Remote
Lead a newly-formed team of Site Reliability Engineers managing Kong's cloud-hosted API gateway service, responsible for maintaining enterprise-grade uptime and performance while staying hands-on with critical implementations. The role suits experienced engineering managers who thrive in high-ownership environments and can balance strategic team-building with direct technical contribution.
Relativity Space
Long Beach, CA · $140,000 – $214,000
Relativity Space seeks a senior infrastructure engineer to design and operate cloud and on-premises platforms supporting their rocket manufacturing and launch operations. This role combines platform reliability, automation, and cross-functional collaboration at a company building commercial spacecraft.
Listed on Relativity Space’s careers site · Apply there ↗
Menlo Ventures Portfolio
Full-time · New York, NY · $320,000 – $485,000
This role owns the data pipelines, observability systems, and operational tooling that track and optimize Anthropic's massive multi-cloud GPU and accelerator fleet. It suits engineers who enjoy moving between data engineering, systems operations, and observability work, and who want to directly influence infrastructure decisions affecting billions of dollars in compute spending.
Listed on Menlo Ventures Portfolio’s careers site · Apply there ↗
Geotab
Full-time · Atlanta, GA
This role leads the evaluation and implementation of AI-driven solutions for Geotab's enterprise IT applications, requiring someone to architect strategies, mentor team members, and manage the full lifecycle from roadmap to deployment. It suits experienced technologists who are passionate about AI automation, comfortable with 24/7 on-call responsibilities, and want to shape the future of how enterprise systems leverage intelligent workflows and autonomous operations.
Listed on Geotab’s careers site · Apply there ↗
BlackSky
Full-time · Herndon, VA · $135,000 – $150,000 · Remote
BlackSky seeks an engineer to design and operate cloud infrastructure, Kubernetes clusters, and GitOps pipelines across public, private, and air-gapped networks for intelligence platforms. This role suits someone with deep Kubernetes expertise and hands-on experience managing enterprise-scale deployments in restrictive network environments.
Listed on BlackSky’s careers site · Apply there ↗
Life360
Full-time · Remote, USA; Remote, Canada · $118,500 – $216,500
This role builds the experimentation and machine learning infrastructure that powers personalization across Life360's platform serving nearly 100 million users. You'll design high-throughput data pipelines and low-latency serving systems while pioneering AI-native engineering practices—using AI coding assistants as genuine development partners rather than just tools.
Listed on Life360’s careers site · Apply there ↗
InstaLily
Full-time · New York, NY · $150,000 – $190,000
InstaLILY is hiring a Site Reliability Engineer to build and operate their Internal Developer Platform and multi-cloud Kubernetes infrastructure that supports their AI agent products. You'll work in a small team with direct impact, partnering with engineers to create self-service tooling and golden paths while managing production systems that power major enterprise customers.
Listed on InstaLily’s careers site · Apply there ↗
Temporal
Full-time · United States - Remote Opportunity · $140,000 – $180,000
A technical support engineer who helps developers deploy and operate Temporal in production, focusing on troubleshooting infrastructure, optimizing performance, and building automation solutions in cloud-native environments. This role suits experienced developers comfortable with distributed systems, cloud platforms, and direct customer interaction.
Listed on Temporal’s careers site · Apply there ↗
Tatari
Full-time · 3 locations · $190,000 – $240,000
This is a systems and infrastructure role focused on keeping Tatari's data platform reliable and stable across all environments, rather than building data pipelines. It suits experienced SRE or platform engineers who have learned to respect production through hard experience and bring operational discipline to infrastructure scaling and deployment.
Listed on Tatari’s careers site · Apply there ↗
Replit
Full-time · Foster City, CA · $220,000 – $325,000 · Remote
This role leads infrastructure reliability and automation efforts across Replit's platform, designing systems and mentoring engineers to handle millions of developers. It suits experienced engineers who enjoy solving complex distributed systems problems and want to establish best practices that make software development more accessible.
Listed on Replit’s careers site · Apply there ↗
San Francisco, CA
Reddit is seeking a Staff-level Site Reliability Engineer to lead reliability efforts across their platform's user-facing systems at massive scale. This role combines technical depth in distributed systems with leadership responsibilities, focusing on improving availability, performance, and operational excellence across web, mobile, and real-time infrastructure.
Listed on Reddit’s careers site · Apply there ↗
Bend Studio
Full-time · San Mateo, CA · $140,500 – $210,700
Test quality assurance engineer for machine learning systems at PlayStation, focused on validating probabilistic models and building automated test frameworks for ML inference services. Suits someone with backend testing expertise who wants to specialize in ML validation and work across distributed, high-scale systems.
Listed on Bend Studio’s careers site · Apply there ↗
AST SpaceMobile
Lanham, MD
This role involves building and maintaining automation workflows within a network operations platform, with equal focus on developing orchestration scripts and validating them through comprehensive testing. It suits experienced automation engineers who want to own end-to-end reliability of operational systems and collaborate closely with service teams in a telecom environment.
Listed on AST SpaceMobile’s careers site · Apply there ↗
Axios
Remote · $130,000 – $165,000
Axios seeks a backend engineer to build and maintain the APIs and services powering their news platform. You'll work with Golang and Python in a collaborative environment focused on reliability, performance, and leveraging AI tools to accelerate development.
Listed on Axios’s careers site · Apply there ↗
Deepgram
USA | Remote · $160,000 – $220,000
This role bridges research and production at an AI voice platform, building the infrastructure that takes experimental models through CI/CD pipelines to serving millions of API requests. You'll own deployment automation, monitoring, and testing systems for machine learning models at scale.
Listed on Deepgram’s careers site · Apply there ↗
PlanetScale
San Francisco, CA · $120,000 – $290,000 · Remote
PlanetScale is building database observability tools that help developers understand performance and query patterns across MySQL, PostgreSQL, and Vitess. This role suits engineers with solid backend experience who have shipped monitoring or observability products before and enjoy working across the full stack.
Listed on PlanetScale’s careers site · Apply there ↗
SentiLink
Full-time · United States · $180,000 – $250,000 · Remote
SentiLink seeks a Staff Infrastructure Engineer to design and maintain the cloud infrastructure, tooling, and systems that support their identity verification platform and engineering teams. This role suits experienced engineers who excel at building scalable, secure infrastructure and improving operational reliability across distributed systems.
Listed on SentiLink’s careers site · Apply there ↗
ExtraHop Networks
Seattle, WA · Remote
This is a general talent community signup for ExtraHop's engineering organization, accepting resumes from software engineers to be contacted when relevant positions open. It suits engineers interested in network security, performance optimization, and distributed systems who want to explore opportunities at a well-established NDR platform company.
Listed on ExtraHop Networks’s careers site · Apply there ↗
Runway
Remote · $280,000 – $340,000
Lead the API and platform infrastructure team at an AI company building world simulation technology. This role suits experienced engineering managers who combine technical depth in cloud systems with product thinking and the ability to scale teams in a fast-moving startup.
Listed on Runway’s careers site · Apply there ↗
Privateer Space
Washington, DC · $180,000 – $200,000
A senior-level security engineer who embeds defensive practices into automated deployment pipelines, managing cloud infrastructure and container systems while ensuring compliance and incident response. This role suits experienced DevOps practitioners ready to lead security integration across development lifecycles.
Listed on Privateer Space’s careers site · Apply there ↗