Senior Site Reliability Engineer
Latitude.sh
- United States, United States
- Remote
- Posted Nov 2, 2025
Job description
About the role
The Senior Site Reliability Engineer will build reliable, observable, and self‑healing systems at scale, enhancing the resilience of the infrastructure that powers Latitude.sh’s global bare‑metal cloud.
About the company
Latitude.sh is a global computing platform that enables businesses to programmatically deploy single‑tenant bare‑metal instances.
Requirements
- Strong verbal and written English communication skills
- Advanced knowledge of Linux/Unix systems in production environments
- Experience with Kubernetes and container orchestration
- Proficiency with infrastructure automation tools (e.g., Terraform, Ansible)
- Experience with observability stacks (e.g., Prometheus, Grafana, Loki, ELK)
- Familiarity with scripting and programming languages such as Bash, Python, Go, or Ruby
- Working knowledge of Git and CI/CD pipelines
- Solid understanding of incident management and root cause analysis processes
- Knowledge of cloud-native reliability and security best practices