# Site Reliability Engineer \(m/f/d\)

- Company: [Codesphere](<https://jobstar.asia/company/codesphere>)
- Location: Remote \(Germany\) · Karlsruhe · Munich
- Remote: Yes
- Team: Product & Engineering
- Employment type: Full Time
- Posted: April 29, 2026

## Job description

## About Codesphere

Codesphere is a *Virtual Cloud Provider* from Germany building the future of sovereign cloud infrastructure. Our platform gives enterprises and governments full sovereignty without giving up modern cloud capability – a vision recently validated by a series of multi-million European government tenders.

Since our founding in Karlsruhe in 2020, we’ve expanded into an international team of 60+ experts. Based in Karlsruhe and Munich and backed by top-tier investors, we are chasing a bold vision.

**We’re scaling fast and would love for you to join us and grow alongside us 🚀**

## What you'll drive

* You define and enforce SLOs, SLIs, and SLAs across production
* You monitor system health, plan capacity, and automate deployments, patching, and infrastructure provisioning
* You diagnose and resolve production incidents fast – including 24/7 on-call participation
* You lead post-mortems and turn findings into prevention; maintain runbooks and escalation procedures
* You manage cloud infrastructure via IaC and own CI/CD pipeline design and maintenance
* You drive scalability, fault tolerance, disaster recovery, and security compliance
* You partner with Dev teams on production readiness, Shift Left practices, and error budget management

## What makes you a great fit

* Proven experience in an SRE, DevOps, or platform engineering role with hands-on production ownership
* Strong knowledge of Kubernetes, Terraform, and Ansible
* Familiarity with Ceph or comparable distributed storage systems
* Experience with SLOs, SLIs, error budgets, and CI/CD pipeline design
* Degree in a relevant field or comparable qualification
* Calm, structured, and fast under pressure – strong debugging and incident response skills
* Good communicator, able to translate operational concerns into guidance for Dev teams
* Go development experience is a plus

## What's in it for you

* 32 days of paid time off – *30 regular vacation days plus Christmas Eve and New Year's Eve off*
* Meal allowance – *up to 15 digital vouchers per month, adding up to over €100 net for you*
* Flexibility – *hybrid work setup with mobile work options and flexibility around core hours*
* Steep learning curve – *fast-moving environment, real ownership, and a front-row seat to scaling a company*
* Job-Rad – *lease a bike through us, tax-free*
* Gym access – *stay active on site (Karlsruhe office only)*
* Employee events – *from team offsites to regular get-togethers*
* Company pension scheme – *company-supported pension to set you up for later*
* Great public transport links – *both offices are within walking distance of tram and metro stops*

## Apply

[Apply on Codesphere](<https://jobs.ashbyhq.com/codesphere/15a9b547-ec89-4e43-a0e3-d2d9e34c5357>)

Canonical job page: <https://jobstar.asia/job/site-reliability-engineer-m-f-d-codesphere-karlsruhe-7c7afe3f1b2a98c9>
