Senior Site Reliability Engineer

Latitude.sh

  • United States, United States
  • Remote
  • Posted Nov 2, 2025
Sign up — let your agent apply Sign in

BestApply tailors your resume and applies for you.

GoRubyGrafanaGitCloud-Native SecurityAnsibleCloud-native reliabilityRoot Cause AnalysisTerraformLinuxPythonCI/CD

Job description

About the role

The Senior Site Reliability Engineer will build reliable, observable, and self‑healing systems at scale, enhancing the resilience of the infrastructure that powers Latitude.sh’s global bare‑metal cloud.

About the company

Latitude.sh is a global computing platform that enables businesses to programmatically deploy single‑tenant bare‑metal instances.

Requirements

  • Strong verbal and written English communication skills
  • Advanced knowledge of Linux/Unix systems in production environments
  • Experience with Kubernetes and container orchestration
  • Proficiency with infrastructure automation tools (e.g., Terraform, Ansible)
  • Experience with observability stacks (e.g., Prometheus, Grafana, Loki, ELK)
  • Familiarity with scripting and programming languages such as Bash, Python, Go, or Ruby
  • Working knowledge of Git and CI/CD pipelines
  • Solid understanding of incident management and root cause analysis processes
  • Knowledge of cloud-native reliability and security best practices