Loading...

Director, Site Reliability Engineering

  • Company: Get A Job.ai
  • Salary: Pay not listed
  • Work type: Remote

Website Get A Job.ai

Represented by Get A Job.ai

About This Opportunity

We are representing a confidential privacy technology company seeking a Director of Site Reliability Engineering to lead infrastructure operations for millions of users worldwide. This is a fully remote position with a distributed team.

In this role, you'll oversee the reliability and scalability of large-scale systems, leading an SRE team responsible for maintaining world-class infrastructure. You'll tackle complex operational challenges involving software, systems, automation, and process optimization while working with technologies including Perl, Go, TypeScript, and Python.

Responsibilities

  • Lead and manage a team of Site Reliability Engineers supporting high-traffic distributed systems
  • Participate in and oversee 24x7 on-call rotation for large-scale production deployments
  • Drive complex, high-impact projects from initial proposal through postmortem analysis
  • Develop and implement tools, services, alerts, and automated responses to identify and mitigate reliability risks
  • Root-cause instability issues in distributed systems handling significant traffic volumes
  • Design and implement infrastructure automation around provisioning and configuration management
  • Partner closely with software engineering teams to triage production issues and recommend remediation strategies
  • Leverage cloud-native services and architectures to enhance system reliability and scalability
  • Guide the technical direction of deployment infrastructure with focus on reliability and performance improvements

What We're Looking For

  • 10+ years of professional experience in reliability, platform, infrastructure, or software engineering
  • 4+ years leading SRE or infrastructure teams
  • Proven experience participating in 24x7 on-call rotations for large-scale systems
  • Proficiency in AI-driven development, including designing and implementing agentic workflows
  • Advanced programming skills in languages such as Go, Python, Perl, or TypeScript
  • Deep expertise administering and troubleshooting Linux and web technologies
  • Strong experience with infrastructure automation, configuration management, and cloud-native architectures
  • Hands-on experience with containerization technologies like Docker and Docker Compose
  • Demonstrated ability to translate vague problems into innovative solutions with measurable outcomes
  • Investigative mindset with strong root-cause analysis capabilities in distributed environments
  • Excellent collaboration skills with ability to lead cross-functional technical initiatives

How We Work With You

Get A Job.ai serves as your advocate throughout the hiring process. When you apply through our platform, one of our experienced recruiters will conduct an initial screening to understand your background and ensure alignment with the role requirements. We then present qualified candidates to our client and coordinate all subsequent interview stages.

Please apply exclusively through Get A Job.ai. Direct contact with the client is not permitted and may disqualify your application. Our team will keep you informed at every step and provide guidance to help you succeed.

Pay

Pay via Get A Job.ai has not been specified for this role. Compensation details will be discussed during the screening process with our recruitment team.

Equal Employment Opportunity: Get A Job.ai is committed to providing equal employment opportunities to all applicants regardless of race, color, religion, sex, national origin, age, disability, genetics, veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by law.

Apply with Get A Job.ai

A recruiter will review your profile and submit you to the client. Do not contact the client directly.

More options

Apply with Get A Job.ai

Apply through Get A Job.ai. A recruiter will review your profile and submit you. Do not contact the client directly.

Apply through Get A Job.ai. A recruiter will review your profile and submit you.

Terms used in this posting

on-call
You may be required to be reachable and available to work outside normal scheduled hours, typically for a set rotation.

Job details above are provided by the employer/source. The sections on this page are compiled from public data sources with AI assistance.

Accommodations: if you need a workplace accommodation to apply for or perform this job, see ADA.gov or EEOC.gov for guidance on your rights and how to request one.

Add application deadline to calendar

Listing facts

  • Role Director, Site Reliability Engineering
  • Employer Get A Job.ai
  • Location Remote · Remote-friendly
  • Type Full Time
  • Pay (from listing) Pay not listed
  • Posted September 22, 2026
  • Apply by October 23, 2026
  • Overview Full job description on this page (468 words)

Facts above come from this job record on Get A Job.AI — not copied from third-party review sites.

Limited public data for this employer

We only show facts we can ground in public sources (Wikidata, O*NET, news/discussion links, or this listing). We do not invent Glassdoor-style ratings, salaries, or testimonials when data is thin. Use the listing facts, occupation context, and related openings below while we continue researching.

Explore related openings

Keep exploring on Get A Job.ai

Not quite the right fit? Your next opportunity is a click away.

Hiring instead? Post a job and reach candidates searching right now.

Director, Site Reliability Engineering Get A Job.ai · Remote