- Company: Get A Job.ai
- Location: Köln
- Salary: Pay not listed
Website Get A Job.ai
Represented by Get A Job.ai
About This Opportunity
We are representing a confidential cloud infrastructure organization that is actively building an on-premise cloud platform. As part of their experienced engineering team, you'll work on OpenStack-based infrastructure and the Kubernetes and GitOps stack that powers their customer-facing cloud platform. AI-assisted engineering is a core part of their daily practice—from spec-driven development to agentic coding workflows, incident response, and automation. The platform is in active development, giving you direct influence on architecture, automation strategy, and the use of AI in platform engineering.
Responsibilities
- Design and develop OpenStack-based on-premise cloud infrastructure with the goal of rolling out complete cloud environments on bare metal in a highly automated manner
- Develop and operate Infrastructure as Code with Ansible and Terraform, as well as Kubernetes and GitOps workflows with FluxCD/ArgoCD—supported by LLMs, agentic workflows, and automated test and review processes
- Own the lifecycle of compute infrastructure—from bare metal, firmware, and provisioning to hypervisors and virtual compute nodes, including patching, migrations, host evacuations, capacity rebalancing, and automation of stable platform operations
- Advance AI substrate and self-healing strategies—from structured knowledge bases and agentic workflows for incident triage and capacity planning to gradual automation of current runbooks
- Design tests for non-regression, performance, and security; document and package solutions; continuously improve the platform based on telemetry, operational experience, and user feedback
- Serve as technical point of contact and sparring partner for colleagues on automation, platform engineering, and AI tooling
What We're Looking For
- Several years of hands-on experience as an SRE, Platform Engineer, or DevOps Engineer operating production infrastructure
- Deep experience with OpenStack, Kubernetes, and Linux, including bare metal environments
- End-to-end experience operating compute infrastructure—from firmware/BIOS rollouts, bare metal provisioning, and hardware diagnostics to hypervisors, migrations, host evacuations, graceful drains, and capacity rebalancing
- AI-assisted engineering as part of your daily workflow; you use LLMs and agentic tools strategically where they meaningfully support development, testing, reviews, or operations, and can assess where AI provides real value versus where deep engineering expertise remains critical
- Proficiency with Ansible, Terraform, and GitOps workflows like FluxCD or ArgoCD, with experience operating automated processes reliably and reproducibly in production
- Experience with Go and/or Python, plus familiarity with Claude Code, Cursor, Aider, or comparable agentic coding environments
- Technical background spanning observability, networking, compute tuning, auto-remediation, security-critical infrastructure, and multi-site cloud environments
- Strong ownership mindset with the drive to not just operate systems but continuously improve them; able to share knowledge and clearly communicate complex technical concepts
- Comfort working in English in an international environment, including discussing, documenting, and advancing technical topics collaboratively
How We Work With You
Candidates apply through Get A Job.ai. Our talent team conducts an initial screening to understand your background and ensure alignment with the role. We then submit qualified candidates to our client for their review and interview process. Please do not contact the client directly—all communication flows through our recruiting team to ensure a smooth, professional process.
Pay
Compensation details will be discussed during the screening process based on your experience and the client's budget.
Equal Opportunity: Get A Job.ai is committed to inclusive hiring practices. We welcome applications from all qualified candidates regardless of background.
Apply with Get A Job.ai
A recruiter will review your profile and submit you to the client. Do not contact the client directly.
Apply through Get A Job.ai. A recruiter will review your profile and submit you. Do not contact the client directly.
Apply through Get A Job.ai. A recruiter will review your profile and submit you.
Local insights for this role are preparing — this section updates automatically in a few seconds (or refresh).
Listing facts
- Role Site Reliability Engineer (m/w/d)
- Employer Get A Job.ai
- Location Köln
- Type Full Time
- Pay (from listing) Pay not listed
- Posted September 11, 2026
- Apply by October 11, 2026
- Country Germany
- Overview Full job description on this page (531 words)
Facts above come from this job record on Get A Job.AI — not copied from third-party review sites.
Typical work in Senior Site Reliability Engineer
Independent occupational context from O*NET (U.S. public-domain labor data). This is about the occupation, not a rewrite of this employer's posting.
- Study product characteristics or customer requirements to determine validation objectives and standards.
- Analyze validation test data to determine whether systems or processes have met validation criteria or to identify root causes of production problems.
- Develop validation master plans, process flow diagrams, test cases, or standard operating procedures.
- Prepare detailed reports or design statements, based on results of validation and qualification tests or reviews of procedures and protocols.
- Maintain validation test equipment.
- Conduct validation or qualification tests of new or existing processes, equipment, or software in accordance with internal protocols or external standards.
Limited public data for this employer
We only show facts we can ground in public sources (Wikidata, O*NET, news/discussion links, or this listing). We do not invent Glassdoor-style ratings, salaries, or testimonials when data is thin. Use the listing facts, occupation context, and related openings below while we continue researching.
Explore related openings
Occupation family: Senior Site Reliability Engineer
Keep exploring on Get A Job.ai
Not quite the right fit? Your next opportunity is a click away.
- Browse all jobs
- More jobs by category
- Remote jobs you can do from anywhere
- Research typical pay for this role
- Set a job alert so new matches reach you first
- Upload your resume to apply faster
Hiring instead? Post a job and reach candidates searching right now.
