Loading...

Senior Site Reliability Engineer

  • Company: Get A Job.ai
  • Salary: Pay not listed
  • Work type: Remote
  • Full Time
  • Remote

Website Get A Job.ai

Represented by Get A Job.ai

About This Opportunity

We are representing a confidential enterprise software organization seeking a Senior Site Reliability Engineer to join their distributed systems team. This is a remote position offering the opportunity to work on large-scale platform infrastructure that supports critical business operations.

Responsibilities

  • Run production environments by monitoring availability and maintaining a holistic view of system health
  • Build software and systems to manage platform infrastructure and applications
  • Improve reliability, quality, and time-to-market of distributed software solutions
  • Measure and optimize system performance, pushing capabilities forward and anticipating scaling needs
  • Provide primary operational support and engineering for multiple large distributed applications
  • Gather and analyze metrics from operating systems and applications to assist in performance tuning and fault finding
  • Partner with development teams to improve services through rigorous testing and release procedures
  • Participate in system design consulting, platform management, and capacity planning
  • Create sustainable systems and services through automation and infrastructure improvements
  • Balance feature development speed and reliability with well-defined service level objectives
  • Drive incident response efforts during outages and critical incidents, including resolution and cross-functional communication
  • Conduct blameless postmortems to continuously improve system reliability

What We're Looking For

Required Experience and Skills:

  • 3-6 years of working experience in site reliability engineering, systems engineering, automation, and reliability
  • Proficiency in at least one programming language such as Python, Go, Java, or C#
  • Experience with scripting languages including Bash or PowerShell
  • Deep understanding of cloud computing platforms, particularly AWS services such as EC2, ECS, Lambda, and DynamoDB
  • Experience with infrastructure as code tools such as CloudFormation or Terraform
  • Strong knowledge of CI/CD concepts and tools such as Jenkins, GitLab CI/CD, or CircleCI
  • Deep experience with containerization technologies including Docker and Kubernetes, plus microservices architecture
  • Experience with monitoring and observability tools such as Prometheus, Grafana, ELK stack, or CloudWatch
  • Excellent problem-solving skills and ability to troubleshoot complex issues in distributed systems
  • Proven experience in incident management and conducting blameless postmortems

Preferred Additional Qualifications:

  • Hands-on experience working with large Kubernetes clusters; certification is a plus
  • Working experience with Grafana Observability Suite including Loki, Mimir, and Tempo
  • Administration or development experience with monitoring and automation tools such as Splunk, Datadog, PagerDuty, or Rundeck
  • Familiarity with configuration management tools like Ansible, Puppet, or Chef
  • Cloud platform certifications such as AWS Certified DevOps Engineer or Google Cloud Professional DevOps Engineer

Personal Attributes:

  • Strong communication skills and ability to collaborate effectively with cross-functional teams
  • Team player with ability to work well in a collaborative environment
  • Fast learner with ability to self-educate on relevant technologies
  • Ability to multitask and prioritize work effectively
  • Ability to remain focused and calm under pressure

How We Work With You

Our talent team at Get A Job.ai partners with leading technology organizations to connect exceptional candidates with confidential opportunities. When you apply through our platform, a dedicated recruiter will review your profile and conduct an initial screening. If there's a strong match, we will submit your profile to our client for consideration. Please do not attempt to contact the client directly, as all communications will be coordinated through Get A Job.ai to ensure a smooth and professional process.

Pay

Compensation details will be discussed with qualified candidates during the screening process.

Equal Employment Opportunity: Get A Job.ai is committed to creating an inclusive recruitment process. We welcome applications from candidates of all backgrounds and ensure fair consideration throughout our evaluation process.

Apply with Get A Job.ai

A recruiter will review your profile and submit you to the client. Do not contact the client directly.

More options

Apply with Get A Job.ai

Apply through Get A Job.ai. A recruiter will review your profile and submit you. Do not contact the client directly.

Apply through Get A Job.ai. A recruiter will review your profile and submit you.

About this role & career path

Traits that fit this role

  • Achievement Orientation
  • Intellectual Curiosity
  • Cautiousness
  • Integrity
  • Attention to Detail

Source: O*NET Work Styles (Distinctiveness Rank).

Typical preparation needed: Job Zone 4: Considerable Preparation Needed. Most of these occupations require a four-year bachelor's degree, but some do not. — via O*NET

Industry news

Source: O*NET (public-domain bulk data)

Salary & compensation

Workers in Architecture & Engineering occupations earn a national median of $95,541via US Census ACS / Data USA

Market context

  • Similar listings in Remote 1

Job details above are provided by the employer/source. The sections on this page are compiled from public data sources with AI assistance.

Accommodations: if you need a workplace accommodation to apply for or perform this job, see ADA.gov or EEOC.gov for guidance on your rights and how to request one.

Add application deadline to calendar

Listing facts

  • Role Senior Site Reliability Engineer
  • Employer Get A Job.ai
  • Location Remote
  • Type Full Time
  • Pay (from listing) Pay not listed
  • Posted September 17, 2026
  • Apply by October 17, 2026
  • Overview Full job description on this page (559 words)

Facts above come from this job record on Get A Job.AI — not copied from third-party review sites.

Typical work in Senior Site Reliability Engineer

Independent occupational context from O*NET (U.S. public-domain labor data). This is about the occupation, not a rewrite of this employer's posting.

  • Study product characteristics or customer requirements to determine validation objectives and standards.
  • Analyze validation test data to determine whether systems or processes have met validation criteria or to identify root causes of production problems.
  • Develop validation master plans, process flow diagrams, test cases, or standard operating procedures.
  • Prepare detailed reports or design statements, based on results of validation and qualification tests or reviews of procedures and protocols.
  • Maintain validation test equipment.
  • Conduct validation or qualification tests of new or existing processes, equipment, or software in accordance with internal protocols or external standards.

Source: O*NET

Limited public data for this employer

We only show facts we can ground in public sources (Wikidata, O*NET, news/discussion links, or this listing). We do not invent Glassdoor-style ratings, salaries, or testimonials when data is thin. Use the listing facts, occupation context, and related openings below while we continue researching.

Explore related openings

Occupation family: Senior Site Reliability Engineer

Keep exploring on Get A Job.ai

Not quite the right fit? Your next opportunity is a click away.

Hiring instead? Post a job and reach candidates searching right now.

Senior Site Reliability Engineer Get A Job.ai · Remote