$150,000 – $190,000
Listed on InstaLily’s own careers site. You apply with them directly — we never stand between you and the employer.
What this role is
InstaLILY is hiring a Site Reliability Engineer to build and operate their Internal Developer Platform and multi-cloud Kubernetes infrastructure that supports their AI agent products. You'll work in a small team with direct impact, partnering with engineers to create self-service tooling and golden paths while managing production systems that power major enterprise customers.
Our summary, not InstaLily’s wording. The full posting is on their site.
Skills this role names
- Amazon Web Services (AWS)
- ArgoCD
- Datadog
- GitHub Actions
- Jenkins
- Kubernetes
- Microsoft Azure
- OpenTelemetry
- Prometheus
- Terraform
Log in to see which of these are already on your profile.
What they ask for
Required
- Bachelor's or Master's degree in Computer Science, Engineering, or related field, or equivalent practical experience
- 3 to 5 years of platform, cloud, infrastructure, or DevOps engineering experience
- Production experience operating Kubernetes clusters
- Hands-on experience with at least one major cloud platform (AWS, GCP, or Azure)
- Infrastructure as Code proficiency with OpenTofu or Terraform
- GitOps workflow experience (ArgoCD, Flux)
- CI/CD tools experience (GitHub Actions, Jenkins, or GitLab CI)
- Cloud networking fundamentals (VPCs, load balancers, DNS, CDNs, service meshes)
- Cloud security and IAM knowledge
- Product-minded approach to internal tooling
- Problem-solving and collaborative work in fast-paced environments
- Effective communication with technical and non-technical teams
Nice to have
- Greenfield Kubernetes build-out experience
- Multi-cloud experience
- Internal Developer Platform contribution or build experience
- GPU workload, model serving, or vector database infrastructure experience