Site Reliability Engineer (m/f/d)

CodeSphere · Remote (Germany)

ExclusiveRemoteFull-timePublished Apr 29, 2026

Apply directly on CodeSphere’s careers site — no account needed.

About the role

About Codesphere

Codesphere is a Virtual Cloud Provider from Germany building the future of sovereign cloud infrastructure. Our platform gives enterprises and governments full sovereignty without giving up modern cloud capability – a vision recently validated by a series of multi-million European government tenders.

Since our founding in Karlsruhe in 2020, we’ve expanded into an international team of 60+ experts. Based in Karlsruhe and Munich and backed by top-tier investors, we are chasing a bold vision.

We’re scaling fast and would love for you to join us and grow alongside us 🚀

What you'll drive

  • You define and enforce SLOs, SLIs, and SLAs across production

  • You monitor system health, plan capacity, and automate deployments, patching, and infrastructure provisioning

  • You diagnose and resolve production incidents fast – including 24/7 on-call participation

  • You lead post-mortems and turn findings into prevention; maintain runbooks and escalation procedures

  • You manage cloud infrastructure via IaC and own CI/CD pipeline design and maintenance

  • You drive scalability, fault tolerance, disaster recovery, and security compliance

  • You partner with Dev teams on production readiness, Shift Left practices, and error budget management

What makes you a great fit

  • Proven experience in an SRE, DevOps, or platform engineering role with hands-on production ownership

  • Strong knowledge of Kubernetes, Terraform, and Ansible

  • Familiarity with Ceph or comparable distributed storage systems

  • Experience with SLOs, SLIs, error budgets, and CI/CD pipeline design

  • Degree in a relevant field or comparable qualification

  • Calm, structured, and fast under pressure – strong debugging and incident response skills

  • Good communicator, able to translate operational concerns into guidance for Dev teams

  • Go development experience is a plus

What's in it for you

  • 32 days of paid time off – 30 regular vacation days plus Christmas Eve and New Year's Eve off

  • Meal allowance – up to 15 digital vouchers per month, adding up to over €100 net for you

  • Flexibility – hybrid work setup with mobile work options and flexibility around core hours

  • Steep learning curve – fast-moving environment, real ownership, and a front-row seat to scaling a company

  • Job-Rad – lease a bike through us, tax-free

  • Gym access – stay active on site (Karlsruhe office only)

  • Employee events – from team offsites to regular get-togethers

  • Company pension scheme – company-supported pension to set you up for later

  • Great public transport links – both offices are within walking distance of tram and metro stops

Skills

  • Kubernetes
  • Terraform
  • Ansible
  • GitHub Actions
  • Go

Never be applicant #200 again

Every job here is indexed straight from company career pages — often hours after it opens, before it reaches the big boards. Create a free account and get your best matches in a twice-daily digest.

  • Your best matches, twice a day
  • No duplicates, no ghost jobs, no recruiter spam
  • Every job free to browse — pay only when you apply
Get my matched jobs

Free account — no card required

93 140 live jobs · 17 642 companies tracked · 26 added today

Similar jobs

Site Reliability Engineer (m/f/d) — CodeSphere · Real Job Offers