All roles

Distinguished Site Reliability Engineer – Cloud

Remote · USA Full-time New today

Job Description:

  • Lead, design, implement and support operational and reliability aspects of large scale Kubernetes clusters with focus on performance at scale, real time monitoring, logging and alerting
  • Engage in and improve the whole lifecycle of services—from inception and design through deployment, operation and refinement
  • Support services before they go live through activities such as system design consulting, developing software tools, platforms and frameworks, capacity management and launch reviews
  • Maintain services once they are live by measuring and monitoring availability, latency and overall system health
  • Scale systems sustainably through mechanisms like automation, and evolve systems by pushing for changes that improve reliability and velocity
  • Practice sustainable incident response and blameless postmortems
  • Be part of an on call rotation to support production systems

Requirements:

  • BS degree in Computer Science or a related technical field involving coding (e.g., physics or mathematics), or equivalent experience
  • 16+ years of experience with Infrastructure automation, distributed systems design, experience with design, develop tools for running large scale private or public cloud system in Production
  • Experience in one or more of the following: Python, Go, Perl or Ruby
  • In depth knowledge on Linux, Networking and Containers

Benefits:

  • equity
  • benefits

Apply tot his job Apply To this Job

Related roles

Site Reliability Engineer, IDaaS Data Platform

Remote · USA Full-time

Site Reliability/Platform Engineer (Linux/ Kubernetes / Python) - 180-190K

Remote · USA Full-time

Site Reliability Engineering Manager

Remote · USA Full-time

Site Reliability Engineer – SkillBridge Intern

Remote · USA Full-time

DevOps Engineer - Kubernetes, AWS & Docker Skills Required (Fully Remote )

Remote · USA Full-time

FSO Audit LABS - Kubernetes DevOps Engineer - Senior - Bay Area

Remote · USA Full-time

Team Lead, Site Reliability Engineering - Storage Layer Service

Remote · USA Full-time

Site Reliability Engineer-SkillBridge Intern

Remote · USA Full-time

SRE Architect + Strong Dynatrace exp

Remote · USA Full-time

Software Engineer – Java, Spring Boot, Kubernetes, AWS

Remote · USA Full-time

Remote arenaflex Customer Care Specialist - Work from Home | Premier Airline Passenger Support Representative

Remote · USA Full-time

Experienced Part-Time Remote Data Entry Specialist – Financial Services Industry

Remote · USA Full-time

Remote Live Chat Agent Work From Home at Game Day USA Naperville, IL

Remote · USA Full-time

Senior Network Engineer

Remote · USA Full-time

Experienced Customer Service/Sales Representative – Driving Sales Growth and Exceptional Customer Experience at arenaflex

Remote · USA Full-time

Remote Data Entry Specialist – Entry Level Part-Time Position | arenaflex Logistics & Supply Chain Excellence

Remote · USA Full-time

Manager Immobilien- & Wohnraumbeschaffung (all genders) - auf den kanarischen Inseln

Remote · USA Full-time

Senior Analyst, Programs and Research, Government Renewal Project

Remote · USA Full-time

Supervisor, TeleNeuro

Remote · USA Full-time

Experienced Remote Customer Service Representative – Work From Home Opportunity with arenaflex

Remote · USA Full-time