- Company: Get A Job.ai
- Salary: Pay not listed
- Work type: Remote
Website Get A Job.ai
Represented by Get A Job.ai
About This Opportunity
We are representing a confidential enterprise software organization seeking a Senior Site Reliability Engineer to join their distributed systems team. This is a remote position offering the opportunity to work on large-scale platform infrastructure that supports critical business operations.
Responsibilities
- Run production environments by monitoring availability and maintaining a holistic view of system health
- Build software and systems to manage platform infrastructure and applications
- Improve reliability, quality, and time-to-market of distributed software solutions
- Measure and optimize system performance, pushing capabilities forward and anticipating scaling needs
- Provide primary operational support and engineering for multiple large distributed applications
- Gather and analyze metrics from operating systems and applications to assist in performance tuning and fault finding
- Partner with development teams to improve services through rigorous testing and release procedures
- Participate in system design consulting, platform management, and capacity planning
- Create sustainable systems and services through automation and infrastructure improvements
- Balance feature development speed and reliability with well-defined service level objectives
- Drive incident response efforts during outages and critical incidents, including resolution and cross-functional communication
- Conduct blameless postmortems to continuously improve system reliability
What We're Looking For
Required Experience and Skills:
- 3-6 years of working experience in site reliability engineering, systems engineering, automation, and reliability
- Proficiency in at least one programming language such as Python, Go, Java, or C#
- Experience with scripting languages including Bash or PowerShell
- Deep understanding of cloud computing platforms, particularly AWS services such as EC2, ECS, Lambda, and DynamoDB
- Experience with infrastructure as code tools such as CloudFormation or Terraform
- Strong knowledge of CI/CD concepts and tools such as Jenkins, GitLab CI/CD, or CircleCI
- Deep experience with containerization technologies including Docker and Kubernetes, plus microservices architecture
- Experience with monitoring and observability tools such as Prometheus, Grafana, ELK stack, or CloudWatch
- Excellent problem-solving skills and ability to troubleshoot complex issues in distributed systems
- Proven experience in incident management and conducting blameless postmortems
Preferred Additional Qualifications:
- Hands-on experience working with large Kubernetes clusters; certification is a plus
- Working experience with Grafana Observability Suite including Loki, Mimir, and Tempo
- Administration or development experience with monitoring and automation tools such as Splunk, Datadog, PagerDuty, or Rundeck
- Familiarity with configuration management tools like Ansible, Puppet, or Chef
- Cloud platform certifications such as AWS Certified DevOps Engineer or Google Cloud Professional DevOps Engineer
Personal Attributes:
- Strong communication skills and ability to collaborate effectively with cross-functional teams
- Team player with ability to work well in a collaborative environment
- Fast learner with ability to self-educate on relevant technologies
- Ability to multitask and prioritize work effectively
- Ability to remain focused and calm under pressure
How We Work With You
Our talent team at Get A Job.ai partners with leading technology organizations to connect exceptional candidates with confidential opportunities. When you apply through our platform, a dedicated recruiter will review your profile and conduct an initial screening. If there's a strong match, we will submit your profile to our client for consideration. Please do not attempt to contact the client directly, as all communications will be coordinated through Get A Job.ai to ensure a smooth and professional process.
Pay
Compensation details will be discussed with qualified candidates during the screening process.
Equal Employment Opportunity: Get A Job.ai is committed to creating an inclusive recruitment process. We welcome applications from candidates of all backgrounds and ensure fair consideration throughout our evaluation process.
Apply with Get A Job.ai
A recruiter will review your profile and submit you to the client. Do not contact the client directly.
Apply through Get A Job.ai. A recruiter will review your profile and submit you. Do not contact the client directly.
Apply through Get A Job.ai. A recruiter will review your profile and submit you.
Local insights for this role are preparing — this section updates automatically in a few seconds (or refresh).
Listing facts
- Role Senior Site Reliability Engineer
- Employer Get A Job.ai
- Location Remote
- Type Full Time
- Pay (from listing) Pay not listed
- Posted September 17, 2026
- Apply by October 17, 2026
- Overview Full job description on this page (559 words)
Facts above come from this job record on Get A Job.AI — not copied from third-party review sites.
Typical work in Senior Site Reliability Engineer
Independent occupational context from O*NET (U.S. public-domain labor data). This is about the occupation, not a rewrite of this employer's posting.
- Study product characteristics or customer requirements to determine validation objectives and standards.
- Analyze validation test data to determine whether systems or processes have met validation criteria or to identify root causes of production problems.
- Develop validation master plans, process flow diagrams, test cases, or standard operating procedures.
- Prepare detailed reports or design statements, based on results of validation and qualification tests or reviews of procedures and protocols.
- Maintain validation test equipment.
- Conduct validation or qualification tests of new or existing processes, equipment, or software in accordance with internal protocols or external standards.
Limited public data for this employer
We only show facts we can ground in public sources (Wikidata, O*NET, news/discussion links, or this listing). We do not invent Glassdoor-style ratings, salaries, or testimonials when data is thin. Use the listing facts, occupation context, and related openings below while we continue researching.
Explore related openings
Occupation family: Senior Site Reliability Engineer
Keep exploring on Get A Job.ai
Not quite the right fit? Your next opportunity is a click away.
- Browse all jobs
- More jobs by category
- Remote jobs you can do from anywhere
- Research typical pay for this role
- Set a job alert so new matches reach you first
- Upload your resume to apply faster
Hiring instead? Post a job and reach candidates searching right now.
