Remote Senior Site Reliability Engineer

Posted 1 hour ago

Share:

Please let Clera know you found this job on RemoteYeah. This helps us get more companies to post jobs here for you.

Description:

  • Join a well-funded AI/ML company focused on geospatial intelligence and climate technology.
  • Own and advance GCP infrastructure and DevOps practices, improving incident management and observability.

Requirements:

  • 3+ years of Site Reliability Engineering or production SRE experience.
  • Proficiency with Google Cloud Platform (GCP) and cost optimization.
  • Hands-on experience with Kubernetes for cluster management.
  • Infrastructure as Code experience with Terraform or Deployment Manager.
  • Scripting skills in Python, Bash, or Go.
  • Strong experience with observability tools: Prometheus, Grafana, OpenTelemetry.
  • Ability to define and implement SLOs, SLIs, and error budgets.
  • Experience with incident management and on-call rotation.
  • Nice to have: Experience designing cloud infrastructure at scale and familiarity with DORA metrics.

Benefits:

  • Fully remote role available for candidates in the EU, UK, or North America.
  • Compensation details to be discussed during the interview process.

Report this job

Job expired or something else is wrong with this job?

Report job
SerpApi

SerpApi

Scrape Google and other search engines from our fast, easy, and complete API.

RemoteYeah Ads