(Senior) Site Reliability Engineer (m/f/d) - Platform & Agentic Operations — Opportunihub
1KOMMA5° logo
Job Remote

(Senior) Site Reliability Engineer (m/f/d) - Platform & Agentic Operations

1KOMMA5° · Germany

At a glance

Type
Job
Organisation
1KOMMA5°
Location
Germany
Work mode
Remote
Deadline
Rolling / not stated
Posted
3 Jul 2026

About this job

<h2>1KOMMA5°</h2>At <strong>1KOMMA5°</strong>, we pursue a clear vision: <strong>Living on wind and sunlight forever for free</strong>. To make this a reality, we are building the energy system of the future with Heartbeat AI. Want to be part of it?We bring together regional craftsmanship and scalable software: We don't think of solar, batteries, heat pumps, and e-mobility as isolated components, but control them as an intelligent, integrated overall system in our virtual power plant. Directly connected to the electricity market – in real time, fully automated. This way, energy is used when it is available from renewables and particularly cost-effective. By 2030, our goal is to transition 1.5 million households to renewable energies. Over 3,000 people are working towards this every day, at more than 80 locations worldwide, from Finland to Australia.<br /><strong>Want to take responsibility and build solutions that truly matter? Apply now and help us shape the energy world of tomorrow.</strong><br />Learn <a href="https://1komma5.com/de/ueber-uns/jobs/tech/" rel="nofollow ugc noopener" target="_blank">more</a> about our Product &amp; Tech team!<h2>Deine Position</h2><p>1KOMMA5° is building Europe’s largest virtual power plant ("Heartbeat AI"). As a Senior SRE in our Platform team, you will bridge classic infrastructure with Agentic Engineering, specifically focusing on leveraging AI agents to eliminate developer friction, optimize CI/CD pipelines, and automate the resolution of code review and deployment bottlenecks.</p><h2>Tech Stack</h2><ul><li><p>Cloud &amp; Infra: GCP (CloudRun, GKE), Terraform, Terramate</p></li><li><p>Reliability: Incident.io, Datadog (OpenTelementry)</p></li><li><p>Agentic: Cursor</p></li><li><p>CI/CD &amp; DevEx: GitHub Actions, Backstage</p></li><li><p>Languages: Python, GoLang, TypeScript</p></li></ul><h2>Key Responsibilities Include but not limited to</h2><ul><li><p>Implement and improve monitoring, alerting, and incident response systems and processes to ensure high reliability for our customers and meet defined SLOs</p></li><li><p>Design, build, and maintain resilient, scalable infrastructure utilizing SRE principles and best practices</p></li><li><p>Attend post-incident reviews, detect patterns and contribute to continuous improvement efforts</p></li><li><p>Execute performance testing, analyze system bottlenecks, and formulate strategies for capacity planning to ensure our systems meet current and future demands effectively</p></li><li><p>Build systems where CI/CD test failures serve as immediate, real-time context for agents, enabling them to analyze logs, trace dependencies, and suggest or apply instant code fixes.</p></li></ul><h2>Dein Profil</h2><ul><li><p>6+ years in SRE, DevOps, or Platform Engineering</p></li><li><p>Strong understanding and practical application of Site Reliability Engineering (SRE) principles, methodologies, and best practices</p></li><li><p>Proficiency in programming/scripting languages such as Python, GoLang or TypeScript</p></li><li><p>Practical understanding of integrating LLMs into automated workflows. You know how to feed live system state (like a fresh CI test failure) into an agent as actionable context.</p></li><li><p>Prior experience in incident management, post-incident reviews, and implementing improvements to prevent future incidents</p></li><li><p>Ability to troubleshoot complex technical issues systematically and effectively</p></li><li><p>Good experience working with a public cloud provider, ideally Google Cloud Platform (GCP), and a solid understanding of its observability services</p></li><li><p>A proactive approach to spotting problems, areas for improvement, and performance bottlenecks</p></li><li><p>Excellent communication skills to convey technical concepts and collaborate effectively with diverse teams</p></li><li><p>Very good knowledge of spoken and written english, german is a plus</p></li><li><p>Residency in Germany</p></li></ul><p>Bonus points for:</p><ul><li><p>Interest in climate tech industry</p></li><li><p>Prior experience with IoT applications</p></li><li><p>Having worked in a scale up environment at a company of similar size</p></li></ul><h2>Benefits</h2><ul><li>You are part of an international, dynamic, and highly motivated team of people who have proven to make things happen</li><li>With your work, you accelerate the "energy transition" and hence have a direct impact on our climate</li><li>Work with and learn from other super-smart colleagues</li><li>You will enjoy direct contact with core decision-makers</li><li>You will enjoy the best chances of entering full-time in one of Europe’s most thriving scaleups</li><li>You work remotely (Germany-wide), with offices in Hamburg, Berlin or Munich</li><li>Create a healthy balance alongside your work and enjoy all the benefits of the EGYM Wellpass</li><li>Benefits and discounts are yours with Futurebens</li><li>Whether city bike or e-bike - be flexible with our job bike leasing and do something good for the environment at the same time</li></ul>

How to apply

  1. 1 Read the full details above and confirm you meet the eligibility criteria.
  2. 2 Prepare your documents — an updated CV, and any cover letter, proposal or certificates required.
  3. 3 Click Apply on official site to complete your application on 1KOMMA5°’s official page.
  4. 4 Submit as early as possible — many close once filled.
Apply on official site

Sourced from jobicy. Always verify details on the official website. Opportunihub never charges you to apply.

Frequently asked questions

How do I apply for (Senior) Site Reliability Engineer (m/f/d) - Platform & Agentic Operations?

Review the full details and eligibility on this page, prepare your documents, then use the “Apply on official site” button to complete your application on 1KOMMA5°’s official page.

Is this opportunity remote or location-based?

This opportunity is remote-friendly and open to applicants who can work from anywhere.

Is (Senior) Site Reliability Engineer (m/f/d) - Platform & Agentic Operations free to apply for?

Opportunihub lists this Job for free. Legitimate Jobs do not ask for payment to apply — never pay a fee to submit an application.