Loading...

Site Reliability Engineers (SRE)

  • Company: Get A Job.ai
  • Salary: Pay not listed
  • Work type: Remote

Website Get A Job.ai

Represented by Get A Job.ai

About This Opportunity

Our talent team is representing a confidential cloud infrastructure organization seeking Site Reliability Engineers to join their remote team. This role focuses on observability, incident response, chaos engineering, and driving reliability best practices across cloud deployments. You'll work with a team dedicated to building resilient, fault-tolerant systems while mentoring colleagues and shaping cultural change around SRE principles.

Responsibilities

  • Implement and enforce reliability best practices and cloud resiliency standards across infrastructure
  • Design and execute chaos engineering experiments to improve fault tolerance and system resilience
  • Monitor and support cloud migration initiatives to ensure seamless transitions with minimal disruption
  • Design, implement, and optimize observability solutions for cloud infrastructure including monitoring, APM, and logging
  • Coordinate with cross-functional IT teams to establish observability standards for both individual and organizational needs
  • Assess cloud deployments regularly for compliance with organizational standards and industry best practices
  • Investigate and remediate gaps in observability and system reliability
  • Participate in service capacity planning, performance analysis, and system tuning activities
  • Provide technical mentorship to colleagues on SRE practices, tooling, and methodologies
  • Stay current with emerging technologies and deliver training on relevant tools and practices
  • Participate in 24/7 on-call rotation for production support

What We're Looking For

Required Qualifications:

  • BSc or MSc degree in Computer Science or related technical field
  • 5+ years of cloud services experience with at least 3 years focused on AWS
  • 3+ years in an SRE role or similar reliability-focused position
  • Hands-on experience with monitoring, APM, logging, and alerting platforms
  • Familiarity with ITIL practices including incident, problem, and change management
  • Advanced knowledge of SRE methodologies and service level management
  • Strong troubleshooting abilities with a track record of mentoring others
  • Extensive experience with Kubernetes and container orchestration ecosystems
  • Advanced proficiency with CI/CD pipelines and Infrastructure as Code tools, especially Terraform and CloudFormation
  • Experience with version control systems such as Git
  • Strong documentation and organizational skills
  • Excellent time management and research capabilities
  • Advanced Linux administration, networking, and scripting expertise

Preferred Qualifications:

  • Experience with streaming platforms like Kafka
  • RDBMS experience, particularly with Postgres and MySQL
  • Proficiency in Python or Go scripting languages

How We Work With You

When you apply through Get A Job.ai, one of our specialized recruiters will review your profile and conduct an initial screening call. If there's a strong match, we'll submit your candidacy to our client for consideration. The client's interview process typically includes an online technical challenge, an introductory conversation with their talent team, technical interviews with the engineering team, and a final interview.

Please apply exclusively through Get A Job.ai—do not attempt to contact the client directly, as all applications are managed confidentially through our platform.

Pay

Compensation details will be discussed during the screening process based on your experience and qualifications.

Equal Opportunity: Get A Job.ai is committed to inclusive hiring practices. We encourage applications from candidates of all backgrounds and work with our clients to ensure fair and equitable recruitment processes.

Apply with Get A Job.ai

A recruiter will review your profile and submit you to the client. Do not contact the client directly.

More options

Apply with Get A Job.ai

Apply through Get A Job.ai. A recruiter will review your profile and submit you. Do not contact the client directly.

Apply through Get A Job.ai. A recruiter will review your profile and submit you.

Terms used in this posting

on-call
You may be required to be reachable and available to work outside normal scheduled hours, typically for a set rotation.

Market context

  • Similar listings in Remote 1

Job details above are provided by the employer/source. The sections on this page are compiled from public data sources with AI assistance.

Accommodations: if you need a workplace accommodation to apply for or perform this job, see ADA.gov or EEOC.gov for guidance on your rights and how to request one.

Add application deadline to calendar

Listing facts

  • Role Site Reliability Engineers (SRE)
  • Employer Get A Job.ai
  • Location Remote · Remote-friendly
  • Type Full Time
  • Pay (from listing) Pay not listed
  • Posted September 17, 2026
  • Apply by October 18, 2026
  • Overview Full job description on this page (481 words)

Facts above come from this job record on Get A Job.AI — not copied from third-party review sites.

Limited public data for this employer

We only show facts we can ground in public sources (Wikidata, O*NET, news/discussion links, or this listing). We do not invent Glassdoor-style ratings, salaries, or testimonials when data is thin. Use the listing facts, occupation context, and related openings below while we continue researching.

Explore related openings

Keep exploring on Get A Job.ai

Not quite the right fit? Your next opportunity is a click away.

Hiring instead? Post a job and reach candidates searching right now.

Site Reliability Engineers (SRE) Get A Job.ai · Remote