Loading...

Senior Site Reliability Engineer

  • Company: Get A Job.ai
  • Salary: Pay not listed
  • Work type: Remote

Website Get A Job.ai

Represented by Get A Job.ai

About This Opportunity

We are representing a confidential enterprise software organization seeking a Senior Site Reliability Engineer to join their fully remote team. Our client is a rapidly growing platform company operating across multiple global markets, and they need an experienced infrastructure professional who thrives on building reliable, scalable systems that empower development teams.

In this role, you'll design and maintain Kubernetes-based multi-cloud infrastructure, implement monitoring and observability solutions, and collaborate directly with product and engineering teams to solve complex technical challenges. You'll own critical systems end-to-end, drive automation initiatives, and mentor engineers who want to elevate their infrastructure expertise.

Responsibilities

  • Design, build, and maintain Kubernetes-based multi-cloud platform architecture with a focus on availability, scalability, and fault tolerance
  • Establish configuration best practices and network services that engineering teams depend on daily
  • Implement and continuously improve monitoring, alerting, and observability tools that provide meaningful visibility into system health and performance
  • Participate in on-call rotations and serve as the technical expert for incident response and resolution
  • Create comprehensive runbooks and automation that transform complex operational challenges into manageable, repeatable processes
  • Collaborate cross-functionally with product engineering, product management, and support teams to define and deliver infrastructure features
  • Identify repetitive operational work and build automation to eliminate toil across the organization
  • Mentor less experienced engineers on complex infrastructure topics, breaking down technical problems into clear, actionable steps
  • Conduct root cause analysis following incidents and implement preventive measures

What We're Looking For

Required Qualifications:

  • Deep, hands-on production experience building, deploying, and maintaining Kubernetes clusters at scale, including workload management, networking, and storage
  • Strong expertise with infrastructure as code tools such as Terraform or similar IaC platforms, including versioning, testing, and safe deployment practices
  • Proven experience implementing monitoring and observability solutions using tools like Prometheus, Grafana, or comparable platforms
  • Demonstrated 3rd-level support and incident response capabilities, with experience diagnosing complex production issues and communicating effectively with stakeholders under pressure
  • Passion for automation and continuously raising quality standards across infrastructure systems
  • Responsible use of AI tools for research, code review, documentation, and automation, with strong judgment about validation, confidentiality, and accountability

Preferred Qualifications:

  • Experience with major cloud providers such as AWS (EKS), Google Cloud Platform (GKE), or similar managed Kubernetes services
  • Hands-on experience with ArgoCD or GitOps workflows for declarative infrastructure management
  • Proficiency in Python, Go, or similar programming languages for building automation tools
  • Experience defining service level objectives (SLOs) and implementing effective alerting frameworks

How We Work With You

As a specialized recruiting firm, Get A Job.ai connects exceptional technical talent with leading organizations. When you apply through our platform, one of our experienced recruiters will review your background and schedule a screening conversation to understand your goals, experience, and what you're looking for in your next role.

If there's a strong mutual fit, our talent team will submit your profile directly to our client and guide you through each stage of their interview process. We'll provide feedback, prepare you for conversations, and advocate on your behalf throughout. Please note that all applications must come through Get A Job.ai—direct contact with the client is not part of this process.

Pay

Compensation details will be discussed during the screening process and are dependent on experience, qualifications, and geographic location.

Equal Employment Opportunity: Get A Job.ai is committed to creating an inclusive recruitment process. We welcome applications from candidates of all backgrounds and provide equal opportunity without regard to race, color, religion, sex, national origin, age, disability, or any other protected characteristic under applicable law.

Apply with Get A Job.ai

A recruiter will review your profile and submit you to the client. Do not contact the client directly.

More options

Apply with Get A Job.ai

Apply through Get A Job.ai. A recruiter will review your profile and submit you. Do not contact the client directly.

Apply through Get A Job.ai. A recruiter will review your profile and submit you.

Terms used in this posting

on-call
You may be required to be reachable and available to work outside normal scheduled hours, typically for a set rotation.

About this role & career path

Traits that fit this role

  • Achievement Orientation
  • Intellectual Curiosity
  • Cautiousness
  • Integrity
  • Attention to Detail

Source: O*NET Work Styles (Distinctiveness Rank).

Typical preparation needed: Job Zone 4: Considerable Preparation Needed. Most of these occupations require a four-year bachelor's degree, but some do not. — via O*NET

Industry news

Source: O*NET (public-domain bulk data)

Salary & compensation

Workers in Architecture & Engineering occupations earn a national median of $95,541via US Census ACS / Data USA

Market context

  • Similar listings in Remote 1

Job details above are provided by the employer/source. The sections on this page are compiled from public data sources with AI assistance.

Accommodations: if you need a workplace accommodation to apply for or perform this job, see ADA.gov or EEOC.gov for guidance on your rights and how to request one.

Add application deadline to calendar

Listing facts

  • Role Senior Site Reliability Engineer
  • Employer Get A Job.ai
  • Location Remote · Remote-friendly
  • Type Full Time
  • Pay (from listing) Pay not listed
  • Posted September 22, 2026
  • Apply by October 22, 2026
  • Overview Full job description on this page (581 words)

Facts above come from this job record on Get A Job.AI — not copied from third-party review sites.

Typical work in Senior Site Reliability Engineer

Independent occupational context from O*NET (U.S. public-domain labor data). This is about the occupation, not a rewrite of this employer's posting.

  • Study product characteristics or customer requirements to determine validation objectives and standards.
  • Analyze validation test data to determine whether systems or processes have met validation criteria or to identify root causes of production problems.
  • Develop validation master plans, process flow diagrams, test cases, or standard operating procedures.
  • Prepare detailed reports or design statements, based on results of validation and qualification tests or reviews of procedures and protocols.
  • Maintain validation test equipment.
  • Conduct validation or qualification tests of new or existing processes, equipment, or software in accordance with internal protocols or external standards.

Source: O*NET

Limited public data for this employer

We only show facts we can ground in public sources (Wikidata, O*NET, news/discussion links, or this listing). We do not invent Glassdoor-style ratings, salaries, or testimonials when data is thin. Use the listing facts, occupation context, and related openings below while we continue researching.

Explore related openings

Occupation family: Senior Site Reliability Engineer

Keep exploring on Get A Job.ai

Not quite the right fit? Your next opportunity is a click away.

Hiring instead? Post a job and reach candidates searching right now.

Senior Site Reliability Engineer Get A Job.ai · Remote