- Company: Get A Job.ai
- Location: Europe
- Salary: Pay not listed
- Work type: Remote
Website Get A Job.ai
Represented by Get A Job.ai
About This Opportunity
We are representing a confidential technology organization in their search for a Senior Site Reliability Engineer to join a lean, high-performing infrastructure team. This is a deeply hands-on role at the core of a high-traffic system handling 5–7k requests per second, where you'll own reliability, performance, and stability in a fast-paced production environment.
You'll tackle real-time production challenges, manage critical incidents, and participate in on-call rotations. This position demands resilience, strong decision-making under pressure, and a proactive approach to continuous improvement at scale.
If you thrive in high-load environments and want direct impact on systems serving millions of users, we'd like to hear from you.
Responsibilities
- Own system reliability through active platform monitoring, alert management, and real-time incident response
- Participate in 24/7 on-call rotations with full ownership of production stability in high-traffic environments
- Investigate incidents, perform root cause analysis, and implement sustainable fixes to prevent recurrence
- Build and continuously improve monitoring, alerting, and observability across the Kubernetes (EKS) ecosystem
- Deploy, manage, and optimize infrastructure using Terraform, Helm, and GitOps tools (Flux/ArgoCD)
- Drive automation initiatives and proactively enhance system resilience to reduce manual intervention
- Maintain and evolve CI/CD pipelines and infrastructure-as-code practices
- Collaborate with engineering teams to support deployments and minimize user impact in live environments
- Introduce new tools and technologies to enhance scalability, reliability, and performance
- Handle environment-specific requests and ensure smooth platform operations under constant load
What We're Looking For
- Strong hands-on experience with Kubernetes (deployment, scaling, troubleshooting) in high-load production environments
- Experience with GitOps tools such as FluxCD or ArgoCD
- Proven track record in incident response, root cause analysis, and postmortem documentation
- Solid experience with AWS, Terraform, Docker, and CI/CD pipelines
- Proficiency with monitoring and observability tools: Datadog, Prometheus, Grafana, and logging stacks like ELK or CloudWatch
- Strong understanding of networking concepts and protocols
- Proficiency in at least one scripting language (Python, Go, or Node.js)
- Experience with version control systems (Git)
- Familiarity with incident management tools like PagerDuty or Opsgenie
- Ability to operate effectively in fast-paced, high-pressure environments with strong ownership and accountability
- Proactive, resilient mindset focused on continuous improvement and system stability
How We Work With You
Candidates apply directly through Get A Job.ai. Our talent team will conduct an initial screening, and qualified candidates will be submitted to our client for consideration. The recruitment process typically includes an HR interview (30-45 minutes), a technical interview (90 minutes), and a final interview with senior leadership (60 minutes).
Important: Please do not attempt to contact the client directly. All communication and coordination will be managed through Get A Job.ai to ensure a smooth and professional process.
Pay
Compensation details will be discussed during the screening process and are competitive within the European market for this level of expertise.
Location: Europe (Remote)
Get A Job.ai is an equal opportunity recruiter. We celebrate diversity and are committed to creating an inclusive environment for all candidates we represent.
Apply with Get A Job.ai
A recruiter will review your profile and submit you to the client. Do not contact the client directly.
Apply through Get A Job.ai. A recruiter will review your profile and submit you. Do not contact the client directly.
Apply through Get A Job.ai. A recruiter will review your profile and submit you.
Terms used in this posting
- on-call
- You may be required to be reachable and available to work outside normal scheduled hours, typically for a set rotation.
Explore Get A Job.ai online
Working in Europe
Weather right now in Europe: checking… · Local time: · Air quality: · Daylight: · UV index: · Wind: · Pollen:
Europe is a continent located entirely in the Northern Hemisphere and mostly in the Eastern Hemisphere. It is bordered by the Arctic Ocean to the north, the Atlantic Ocean to the west, the Mediterranean Sea to the south, and Asia to the east. Europe shares the landmass of Eurasia with Asia, and of Afro-Eurasia with both Africa and Asia. Europe is commonly considered to be separated from Asia by the watershed of the Ural Mountains, the Ural River, the Caspian Sea, the Greater Caucasus, the Black Sea, and the Turkish straits.
- Elevation 379m (1,243 ft)
About this role & career path
Traits that fit this role
- Achievement Orientation
- Intellectual Curiosity
- Cautiousness
- Integrity
- Attention to Detail
Source: O*NET Work Styles (Distinctiveness Rank).
Typical preparation needed: Job Zone 4: Considerable Preparation Needed. Most of these occupations require a four-year bachelor's degree, but some do not. — via O*NET
Industry news
- Inaugural Architecture & Engineering Industry Day - portseattle.org
- Mechanical engineering scholarship delivers immediate impact - Texas A&M
- Call for Proposals: Engineering X Skills for Safety Impact Grants - fundsforNGOs
Source: O*NET (public-domain bulk data)
Salary & compensation
Workers in Architecture & Engineering occupations earn a national median of $95,541 — via US Census ACS / Data USA
Job details above are provided by the employer/source. The sections on this page are compiled from public data sources with AI assistance.
Accommodations: if you need a workplace accommodation to apply for or perform this job, see ADA.gov or EEOC.gov for guidance on your rights and how to request one.
Listing facts
- Role Senior Site Reliability Engineer
- Employer Get A Job.ai
- Location Europe · Remote-friendly
- Type Full Time
- Pay (from listing) Pay not listed
- Posted September 14, 2026
- Apply by October 14, 2026
- Overview Full job description on this page (484 words)
Facts above come from this job record on Get A Job.AI — not copied from third-party review sites.
Typical work in Senior Site Reliability Engineer
Independent occupational context from O*NET (U.S. public-domain labor data). This is about the occupation, not a rewrite of this employer's posting.
- Study product characteristics or customer requirements to determine validation objectives and standards.
- Analyze validation test data to determine whether systems or processes have met validation criteria or to identify root causes of production problems.
- Develop validation master plans, process flow diagrams, test cases, or standard operating procedures.
- Prepare detailed reports or design statements, based on results of validation and qualification tests or reviews of procedures and protocols.
- Maintain validation test equipment.
- Conduct validation or qualification tests of new or existing processes, equipment, or software in accordance with internal protocols or external standards.
Limited public data for this employer
We only show facts we can ground in public sources (Wikidata, O*NET, news/discussion links, or this listing). We do not invent Glassdoor-style ratings, salaries, or testimonials when data is thin. Use the listing facts, occupation context, and related openings below while we continue researching.
Explore related openings
Occupation family: Senior Site Reliability Engineer
Keep exploring on Get A Job.ai
Not quite the right fit? Your next opportunity is a click away.
- Browse all jobs
- More jobs by category
- Remote jobs you can do from anywhere
- Research typical pay for this role
- Set a job alert so new matches reach you first
- Upload your resume to apply faster
Hiring instead? Post a job and reach candidates searching right now.
