- Company: Get A Job.ai
- Location: Berlin
- Salary: Pay not listed
Website Get A Job.ai
Represented by Get A Job.ai
About This Opportunity
We are representing a confidential e-commerce technology organization seeking a Senior Site Reliability Engineer to join their growing Berlin team. This is a rare opportunity to build a modern private cloud platform on Kubernetes and Harvester—running on dedicated hardware, not managed cloud consoles. You'll help define SRE practices from the ground up, working alongside development teams to enable reliable service delivery for production systems that handle billions of queries annually.
This is a hybrid role based in Berlin, requiring three office days per week.
Responsibilities
- Define and own service level objectives, indicators, and error budgets to drive data-informed reliability decisions
- Lead incident response end-to-end: fast detection, clear communication, blameless postmortems, and structural reduction of incident classes
- Eliminate operational toil through automation and GitOps practices
- Evolve observability capabilities including metrics, logs, traces, alerting, and runbooks across multiple technology stacks
- Contribute to building custom Kubernetes operators and CRDs that make stateful workloads declarative, self-healing, and safely upgradable
- Implement and roll out auto-scaling solutions including HPA, VPA, KEDA, and cluster autoscaling
- Plan capacity, performance, and cost across on-premises and cloud environments, accounting for peak-season loads
- Join the on-call rotation and own reliability topics within your first 90 days
What We're Looking For
Must-have qualifications:
- Production Kubernetes experience building and maintaining clusters on dedicated servers (e.g., kubeadm, RKE2, k3s)—managed-only experience is not sufficient
- Hands-on cluster lifecycle management and upgrade experience
- Lived SRE practice including SLOs, error budgets, incident management, and on-call responsibilities
- Practical experience with GitOps or comparable infrastructure/deployment automation
- Solid observability skills across metrics, logs, traces, and reliable alerting
- Strong automation instinct—you prefer fixing root causes over repeating workarounds
- Collaborative, enabling mindset that views SRE as a service to developers
- Fluent English required
Nice-to-have experience (genuinely optional):
- Harvester, KubeVirt, vSphere/ESXi, OpenStack, or similar virtualization/HCI platforms
- Container storage solutions such as Longhorn or Ceph
- Datacenter networking including load balancing, ingress, and VLAN configuration
- Auto-scaling implementations and capacity/cost planning
- Building Kubernetes operators and custom resource definitions
- German language skills
- Certifications (CKA, CKS) welcome but not required
If you've owned production systems, handled incidents, and worked deeply with Kubernetes—even if your title was never "SRE"—we encourage you to apply. Production experience and engineering mindset matter more than titles.
How We Work With You
When you apply through Get A Job.ai, our talent team will review your profile and conduct an initial screen. Qualified candidates will be submitted to our client for consideration. The client's process includes an introductory call, a take-home task (approximately two hours), a 90-minute technical interview with their development team, a leadership conversation, and a team meet.
Please apply exclusively through Get A Job.ai—do not contact the client directly.
Pay
Compensation details will be discussed during the screening process with our recruiting team.
Equal Employment Opportunity: Get A Job.ai is committed to inclusive hiring practices. We welcome applications from candidates of all backgrounds and work with our clients to ensure fair consideration throughout the recruitment process.
Apply with Get A Job.ai
A recruiter will review your profile and submit you to the client. Do not contact the client directly.
Apply through Get A Job.ai. A recruiter will review your profile and submit you. Do not contact the client directly.
Apply through Get A Job.ai. A recruiter will review your profile and submit you.
Terms used in this posting
- on-call
- You may be required to be reachable and available to work outside normal scheduled hours, typically for a set rotation.
- hybrid
- A work arrangement combining both in-office and remote/at-home work, typically on a set schedule.
Explore Get A Job.ai online
Working in Berlin, Deutschland
Weather right now in Berlin, Deutschland: checking… · Local time: · Air quality: · Daylight: · UV index: · Wind: · Pollen:
Berlin is the capital of Germany as well as its largest city by both area and population. With 3.7 million inhabitants, it has the highest population within its city limits of any city in the European Union. The city is also one of the states of Germany, being the third-smallest state in the country by area. Berlin is surrounded by the state of Brandenburg, bordering Brandenburg's capital Potsdam to the southwest. The urban area of Berlin has a population of over 5 million, making it the most populous in Germany. The Berlin-Brandenburg capital region has around 6 million inhabitants and is Ger
Note: Germany observes a public holiday on Oct 3 — German Unity Day.
🇩🇪 Relocation safety for Germany: Very Safe — via Warnely, CC BY 4.0
National unemployment rate in Germany: 3.7% — via World Bank
Wage growth in Germany (year over year): 2.9% — via Eurostat
GDP per capita in Germany: $60,496 — via World Bank
Consumer price inflation in Germany: 2.2% (annual) — via World Bank
Real GDP growth in Germany: 0.2% (annual) — via World Bank
Statutory minimum wage in Germany: €2,343/month — via Eurostat
Cost of living in Germany: 9.1% above the EU average — via Eurostat
Job vacancy rate in Germany: 2.8% — via Eurostat
Average hours worked per year in Germany: 1,332 — via OECD
Nearby green space: 10 parks within 1.5km — closest is Lustgarten (306m). via OpenStreetMap
Nearest public transit: Staatsoper (bus stop, 55m). via OpenStreetMap
- Elevation 36m (118 ft)
Job details above are provided by the employer/source. The sections on this page are compiled from public data sources with AI assistance.
Accommodations: if you need a workplace accommodation to apply for or perform this job, see ADA.gov or EEOC.gov for guidance on your rights and how to request one.
Listing facts
- Role Senior Site Reliability Engineer / SRE – Kubernetes & Hybrid Cloud (m/f/d)
- Employer Get A Job.ai
- Location Berlin
- Type Full Time
- Pay (from listing) Pay not listed
- Posted September 21, 2026
- Apply by October 21, 2026
- Country Germany
- Overview Full job description on this page (500 words)
Facts above come from this job record on Get A Job.AI — not copied from third-party review sites.
Limited public data for this employer
We only show facts we can ground in public sources (Wikidata, O*NET, news/discussion links, or this listing). We do not invent Glassdoor-style ratings, salaries, or testimonials when data is thin. Use the listing facts, occupation context, and related openings below while we continue researching.
Explore related openings
Keep exploring on Get A Job.ai
Not quite the right fit? Your next opportunity is a click away.
- Browse all jobs
- More jobs by category
- Remote jobs you can do from anywhere
- Research typical pay for this role
- Set a job alert so new matches reach you first
- Upload your resume to apply faster
Hiring instead? Post a job and reach candidates searching right now.
