Listed on Moderna’s own careers site. You apply with them directly — we never stand between you and the employer.
What this role is
This role involves building and leading an enterprise-wide observability platform that monitors applications, infrastructure, and AI systems across Moderna's global operations. The position suits experienced engineers who want to shape platform strategy, work with modern technologies like OpenTelemetry and Grafana, and drive operational excellence in a regulated healthcare environment.
Our summary, not Moderna’s wording. The full posting is on their site.
Skills this role names
- Amazon Web Services (AWS)
- Ansible
- Bash (Scripting)
- Datadog
- Dynatrace
- Elasticsearch
- Grafana
- Jira
- Kubernetes
- Microsoft Azure
- OpenTelemetry
- PagerDuty
- Prometheus
- Python
- ServiceNow
- Terraform
Log in to see which of these are already on your profile.
What they ask for
Required
- 7+ years in site reliability engineering, observability engineering, or platform engineering
- Hands-on experience designing and operating modern observability platforms
- Strong understanding of metrics, logs, traces, telemetry pipelines, and SLO/SLI frameworks
- Experience with observability technologies such as OpenTelemetry, Prometheus, Grafana, VictoriaMetrics, Elastic, Datadog, or Dynatrace
- Experience supporting applications, infrastructure, containers, cloud services, and distributed systems
- Experience integrating observability platforms with incident management workflows
- Hands-on experience with automation and infrastructure-as-code (Python, Terraform, Ansible, Bash)
- Experience in cloud-native and hybrid environments (AWS and/or Azure)
- Strong analytical, troubleshooting, and problem-solving skills
- Strong communication and stakeholder management skills
Nice to have
- Experience in biotech, pharmaceutical, healthcare, or regulated environments (GxP, HIPAA)
- Experience with AI observability tools (Langfuse, Arize, Phoenix, LangSmith)
- Experience monitoring AI agents, LLM applications, or agentic workflows
- Experience with enterprise logging platforms and large-scale log management
- Experience integrating observability with PagerDuty, ServiceNow, Jira
- Relevant certifications in AWS, Azure, Kubernetes, or observability