About this job
<h3>About the Role</h3><p style="min-height:1.5em">We're looking for a <strong>Senior Backend Engineer</strong> to own and scale the infrastructure behind one of the most widely adopted open-source LLM engineering platforms in the world — processing terabytes of AI tracing data per day for thousands of AI teams, including a significant portion of the Fortune 50.</p><p style="min-height:1.5em">We're a small, engineering-heavy team building what many consider the "Datadog" of LLM observability. We sit at the intersection of open source, high-throughput data infrastructure, and applied AI — and we're now part of a leading columnar database company, which means the engineers who built the database you'll be optimizing are a Slack message away.</p><p style="min-height:1.5em">This is a <strong>hybrid role</strong> based in EU timezones, with approximately <strong>one week per month</strong> in our Berlin office. You can be based anywhere in the EU. <em>Visa sponsorship is not available — you must be legally authorized to work in the EU.</em></p><h3>What You'll Do</h3><ul style="min-height:1.5em"><li><p style="min-height:1.5em"><strong>Own and optimize the ingestion pipeline:</strong> Work on throughput, latency, reliability, and cost efficiency across a high-volume data pipeline that flows from our API through to a ClickHouse-backed analytics store.</p></li><li><p style="min-height:1.5em"><strong>Make ClickHouse sing:</strong> Design and evolve schemas, write and tune complex analytical queries, and partner with ClickHouse engineers to push the limits of what's possible at petabyte scale.</p></li><li><p style="min-height:1.5em"><strong>Build backend abstractions:</strong> Create reusable, well-documented building blocks so that product engineers can ship data-heavy features without reinventing infrastructure each time.</p></li><li><p style="min-height:1.5em"><strong>Operate production systems:</strong> Run and improve cloud deployments across multiple production environments — handling capacity planning, incident response, and monitoring with SRE discipline.</p></li><li><p style="min-height:1.5em"><strong>Make self-hosting effortless:</strong> Maintain and improve deployment tooling from a single Docker Compose setup all the way to enterprise-grade Helm chart deployments on Kubernetes.</p></li><li><p style="min-height:1.5em"><strong>Contribute to open source:</strong> Everything you ship is public and immediately used by developers worldwide — participate in PR reviews, issue triage, and community engagement.</p></li></ul><h3>What We're Looking For</h3><p style="min-height:1.5em"><strong>Required:</strong></p><ul style="min-height:1.5em"><li><p style="min-height:1.5em">Based in the EU and able to work EU timezones; willing to be in Berlin approximately one week per month</p></li><li><p style="min-height:1.5em">Legally authorized to work in the EU without company visa sponsorship</p></li><li><p style="min-height:1.5em">Hands-on experience with <strong>ClickHouse or other columnar/analytical databases</strong> — schema design, query optimization, and managing large-scale analytical workloads</p></li><li><p style="min-height:1.5em">Experience designing, building, and operating <strong>high-throughput data pipelines</strong> (streaming and/or batch) capable of processing terabytes of data per day</p></li><li><p style="min-height:1.5em">Production experience with <strong>observability and distributed tracing systems</strong> (metrics, logs, traces) for large-scale, low-latency services</p></li><li><p style="min-height:1.5em">Experience deploying and operating services with <strong>Docker and Kubernetes</strong>, and applying SRE practices such as capacity planning, incident response, and monitoring</p></li><li><p style="min-height:1.5em">Proven ability to take <strong>end-to-end ownership</strong> of backend/platform components and work autonomously in a small, fast-moving engineering team</p></li><li><p style="min-height:1.5em">3+ years of backend or full-stack engineering experience</p></li></ul><p style="min-height:1.5em"><strong>Nice to Have:</strong></p><ul style="min-height:1.5em"><li><p style="min-height:1.5em">Experience contributing to or maintaining <strong>open-source projects</strong> — including PR reviews, issue triage, and community workflows</p></li><li><p style="min-height:1.5em">Background at companies working in AI/ML tooling, data infrastructure, or observability</p></li></ul><h3>Compensation & Benefits</h3><ul style="min-height:1.5em"><li><p style="min-height:1.5em"><strong>Salary:</strong> €90,000 – €160,000 per year, depending on experience</p></li><li><p style="min-height:1.5em">Meaningful equity in a well-backed, high-growth company (previously Series A/B stage investors; now part of a larger technology organization)</p></li><li><p style="min-height:1.5em">Small team = large scope and direct impact on product and architecture</p></li><li><p style="min-height:1.5em">Work with — and get direct access to — world-class database engineers at the infrastructure layer you depend on daily</p></li><li><p style="min-height:1.5em">Flexible remote-first culture with regular in-person time in Berlin</p></li></ul><h3>Location</h3><p style="min-height:1.5em">This role is open to candidates across the <strong>EU</strong> working in EU timezones. The team meets in person in <strong>Berlin, Germany</strong> approximately one week per month. Fully remote otherwise. <strong>Visa sponsorship is not available.</strong></p><p>Find <a href="https://www.arbeitnow.com">Jobs in Germany</a> on Arbeitnow</a>