Loading...

Director, Site Reliability Engineering

  • Company: Get A Job.ai
  • Salary: Pay not listed
  • Work type: Remote

Website Get A Job.ai

Represented by Get A Job.ai

About This Opportunity

We are representing a well-established, profitable technology company seeking a Director, Site Reliability Engineering to lead infrastructure operations for millions of users. This is a fully remote position with a distributed team across multiple countries.

Our client operates at significant scale, handling billions of requests and maintaining a global user base. As the engineering leader for site reliability, you will own complex operational challenges spanning software, systems, automation, and process optimization. You'll work closely with platform and product engineering teams to ensure world-class uptime and performance.

Responsibilities

  • Lead and grow an SRE team responsible for large-scale production infrastructure
  • Participate in and oversee 24x7 on-call rotations for critical systems
  • Drive reliability initiatives from proposal through implementation and postmortem analysis
  • Build automation around infrastructure provisioning and configuration management
  • Develop tooling, monitoring, alerting, and incident response processes to identify and mitigate reliability risks
  • Root-cause complex issues in high-traffic, distributed systems
  • Partner with software engineering teams to triage production incidents and implement code-level fixes
  • Scale infrastructure to support rapid growth while maintaining performance standards
  • Design and implement agentic workflows leveraging AI-driven development practices
  • Set technical direction for deployment architecture with a focus on reliability and efficiency

What We're Looking For

  • 10+ years of professional experience in reliability, platform, infrastructure, or software engineering
  • 4+ years leading SRE teams in production environments
  • Proven experience with 24x7 on-call rotations for large-scale deployments
  • Advanced programming skills in languages such as Perl, Go, TypeScript, or Python
  • Deep expertise administering and troubleshooting Linux and web technologies
  • Strong experience with containerization (Docker, Docker Compose) and cloud-native architectures
  • Track record of implementing infrastructure automation and configuration management
  • Proficiency in AI-driven development, including designing agentic workflows
  • Ability to analyze ambiguous problems, propose innovative solutions, and execute with metric-driven focus
  • Investigative mindset for diagnosing instability in distributed systems
  • Strong collaboration skills to lead high-impact, complex projects across teams

How We Work With You

Candidates apply directly through Get A Job.ai. Our talent team will conduct an initial screening to understand your background and career goals. If there's a strong match, we'll submit your profile to our client for consideration. Please do not attempt to contact the client directly—all communication and coordination will flow through our recruiting team to ensure a smooth process for everyone involved.

Pay

Pay via Get A Job.ai: $243,800 USD annually, plus stock options.

Equal Opportunity

Get A Job.ai provides equal employment opportunities to all qualified applicants without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability, or veteran status.

Apply with Get A Job.ai

A recruiter will review your profile and submit you to the client. Do not contact the client directly.

More options

Apply with Get A Job.ai

Apply through Get A Job.ai. A recruiter will review your profile and submit you. Do not contact the client directly.

Apply through Get A Job.ai. A recruiter will review your profile and submit you.

Terms used in this posting

on-call
You may be required to be reachable and available to work outside normal scheduled hours, typically for a set rotation.

Job details above are provided by the employer/source. The sections on this page are compiled from public data sources with AI assistance.

Accommodations: if you need a workplace accommodation to apply for or perform this job, see ADA.gov or EEOC.gov for guidance on your rights and how to request one.

Add application deadline to calendar

Listing facts

  • Role Director, Site Reliability Engineering
  • Employer Get A Job.ai
  • Location Remote · Remote-friendly
  • Type Full Time
  • Pay (from listing) Pay not listed
  • Posted September 23, 2026
  • Apply by October 23, 2026
  • Overview Full job description on this page (417 words)

Facts above come from this job record on Get A Job.AI — not copied from third-party review sites.

Limited public data for this employer

We only show facts we can ground in public sources (Wikidata, O*NET, news/discussion links, or this listing). We do not invent Glassdoor-style ratings, salaries, or testimonials when data is thin. Use the listing facts, occupation context, and related openings below while we continue researching.

Explore related openings

Keep exploring on Get A Job.ai

Not quite the right fit? Your next opportunity is a click away.

Hiring instead? Post a job and reach candidates searching right now.

Director, Site Reliability Engineering Get A Job.ai · Remote