Loading...

Site Reliability Engineer

  • Company: Get A Job.ai
  • Location: Paris
  • Salary: Pay not listed
  • Full Time
  • Paris

Website Get A Job.ai

Represented by Get A Job.ai

About This Opportunity

Our talent team is representing a confidential AI technology organization seeking a Site Reliability Engineer to join their Platform team in Paris. This is an opportunity to work on cutting-edge infrastructure supporting large-scale machine learning systems and customer-facing applications in a fast-paced, innovation-driven environment.

In this role, you'll be instrumental in shaping the reliability, scalability, and performance of production systems. You'll balance hands-on operational work with strategic engineering improvements, directly impacting the stability of web services, inference environments, and ML workloads across distributed infrastructure.

Responsibilities

  • Design, build, and maintain scalable, highly available, and fault-tolerant infrastructures supporting web services and ML workloads
  • Ensure platform, inference, and model training environments maintain high availability and enable seamless replication across HPC clusters
  • Operate production systems and troubleshoot issues, including on-call responses, incident management, and infrastructure scaling
  • Implement and improve monitoring, alerting, and incident response systems to minimize downtime and optimize performance
  • Develop and maintain workflows and tools for CI/CD, containerization, orchestration, monitoring, and logging
  • Participate in on-call rotations, respond to incidents, and perform thorough root cause analysis
  • Drive continuous improvement in infrastructure automation, deployment, and orchestration using Kubernetes, Flux, Terraform, and similar tools
  • Collaborate with AI/ML researchers to enable safe and reproducible model-training experiments
  • Build cloud-agnostic platform solutions that abstract infrastructure complexities for engineering teams
  • Work with security teams to ensure infrastructure adheres to best practices and compliance requirements
  • Document processes and procedures to ensure consistency and knowledge sharing

What We're Looking For

  • Master's degree in Computer Science, Engineering, or related field
  • 7+ years of experience in DevOps or SRE roles with strong expertise in cloud computing and distributed systems
  • Hands-on experience with site reliability issues, including root cause analysis, production troubleshooting, and on-call rotations
  • Proficiency working with reliability KPIs such as observability, alerting, and SLAs
  • Experience with CI/CD, containerization, and orchestration tools like Docker and Kubernetes
  • Knowledge of monitoring, logging, alerting, and observability tools (Prometheus, Grafana, ELK Stack, Datadog, or similar)
  • Familiarity with infrastructure-as-code tools like Terraform or CloudFormation
  • Proficiency in scripting languages including Python, Go, and Bash
  • Strong understanding of software development best practices
  • Solid grasp of networking, security, and system administration concepts
  • Excellent problem-solving and communication skills with ability to work effectively in collaborative environments
  • Experience in AI/ML environments, high-performance computing (HPC) systems, or modern cloud infrastructure solutions is a strong plus

How We Work With You

When you apply through Get A Job.ai, our experienced recruiting team will review your profile and conduct an initial screening to understand your background and career goals. If there's a strong match, we'll submit your candidacy to our client for consideration and guide you through their interview process.

Please apply exclusively through Get A Job.ai. Do not attempt to contact the client directly, as this may disqualify your application. We manage all communications and negotiations on your behalf to ensure the best possible outcome.

Pay

Compensation details will be discussed during the screening process based on your experience and qualifications.

Get A Job.ai is an equal opportunity recruiter. We celebrate diversity and are committed to creating an inclusive environment for all candidates regardless of race, color, religion, sex, sexual orientation, gender identity, national origin, veteran, or disability status.

Apply with Get A Job.ai

A recruiter will review your profile and submit you to the client. Do not contact the client directly.

More options

Apply with Get A Job.ai

Apply through Get A Job.ai. A recruiter will review your profile and submit you. Do not contact the client directly.

Apply through Get A Job.ai. A recruiter will review your profile and submit you.

Local insights for this role are preparing — this section updates automatically in a few seconds (or refresh).

Listing facts

  • Role Site Reliability Engineer
  • Employer Get A Job.ai
  • Location Paris
  • Type Full Time
  • Pay (from listing) Pay not listed
  • Posted September 5, 2026
  • Apply by October 6, 2026
  • Overview Full job description on this page (533 words)

Facts above come from this job record on Get A Job.AI — not copied from third-party review sites.

Typical work in Senior Site Reliability Engineer

Independent occupational context from O*NET (U.S. public-domain labor data). This is about the occupation, not a rewrite of this employer's posting.

  • Study product characteristics or customer requirements to determine validation objectives and standards.
  • Analyze validation test data to determine whether systems or processes have met validation criteria or to identify root causes of production problems.
  • Develop validation master plans, process flow diagrams, test cases, or standard operating procedures.
  • Prepare detailed reports or design statements, based on results of validation and qualification tests or reviews of procedures and protocols.
  • Maintain validation test equipment.
  • Conduct validation or qualification tests of new or existing processes, equipment, or software in accordance with internal protocols or external standards.

Source: O*NET

Limited public data for this employer

We only show facts we can ground in public sources (Wikidata, O*NET, news/discussion links, or this listing). We do not invent Glassdoor-style ratings, salaries, or testimonials when data is thin. Use the listing facts, occupation context, and related openings below while we continue researching.

Explore related openings

Occupation family: Senior Site Reliability Engineer

Keep exploring on Get A Job.ai

Not quite the right fit? Your next opportunity is a click away.

Hiring instead? Post a job and reach candidates searching right now.

Site Reliability Engineer Get A Job.ai · Paris