Remote Jobs RockRemote Jobs Rock

Senior Site Reliability Engineer

🕒 3 days ago
KubernetesSRE PrinciplesAWSAzure

📜 Description

  • Participate in an on-call rotation to diagnose and resolve production issues using runbooks and playbooks.
  • Design, build, and maintain observability dashboards and alerting based on SLIs and SLOs.
  • Operate and enhance our Kubernetes-based compute platform for infrastructure management.
  • Investigate and resolve production incidents, conducting root-cause analysis and remediation.
  • Collaborate with product engineering teams to review architecture and infrastructure decisions.

🛠️ Requirements

  • 8-10+ years of relevant hands-on experience.
  • Strong, hands-on understanding of Kubernetes.
  • Practical experience with SLIs, SLOs, and error budgets.
  • Solid grounding in AWS or Azure, including networking fundamentals.
  • Deep experience with modern observability stacks like OpenTelemetry or Prometheus.
  • Strong understanding of CI/CD systems, preferably GitHub Actions.
  • Strong programming skills in .NET, ASP.NET, Python, or Java.

Benefits

  • Company-paid medical
  • Dental
  • Vision with 100% employer paid options
  • 401k match and telehealth options including memberships to One Medical.
  • Parental leave and support, up to $20k in fertility services.
Bitwarden

Senior Site Reliability Engineer - FedRAMP

Bitwarden👥 10,000+ employees🏢 Software Development🤝 B2B
🕒 9 days ago

As a Senior Site Reliability Engineer - FedRAMP, you will manage and enhance the Bitwarden Gov cloud infrastructure, ensuring security, reliability, and compliance in a multi-cloud environment.

FedrampCloud InfrastructureSite Reliability EngineeringIncident Response

Trusted by Remote Workers