About this job
<p>As one of the leading German HealthTech companies, we are reshaping discharge management – technology-driven, patient-centered, and free from bureaucracy. In addition to our market-leading SaaS platform, we develop AI solutions that radically simplify processes in hospitals and for aftercare providers, relieve healthcare professionals, and refocus attention on patients. Today, we already connect two-thirds of all German hospitals with over 650 rehabilitation clinics and 25,000 nursing and homecare providers. With currently around 100 employees, we continue to grow – and we are looking for people with character who want to help us improve the healthcare system and solve administrative complexity across care journeys in Europe.</p><p>Our AI portfolio spans model development through to production inference at scale, in a highly regulated healthcare context. This role exists to make sure that infrastructure runs reliably, scales predictably, and gets AI engineers' work safely into production.<br></p><h2>What to expect<br></h2><p>This is what you can expect from us as a Senior Forward Deployed Platform Engineer:</p><ul><li><p>Purposeful work – your work keeps our AI portfolio reliable and performant for hospitals and aftercare providers who depend on it daily.</p></li><li><p>Company culture – we believe in flat hierarchies that promote high performance and strong team dynamics. We foster an environment characterized by mutual respect, loyalty, and recognition. Together, we strive for our goals – and expect the same from you.</p></li><li><p>Flexibility – want to pick up your child from daycare? Like to exercise during lunch? We'll support you. We are a remote-friendly company offering flexible working hours. Workations are also possible by arrangement.</p></li><li><p>Edenred card – which you can use according to your needs.</p></li><li><p>Extra vacation day – so you can celebrate your birthday with your loved ones, you'll have the day off.<br></p></li></ul><p>This is how you will make an impact as a Senior Forward Deployed Platform Engineer:</p><ul><li><p>Our AI portfolio needs to run reliably and perform well. Thus it has to be properly monitored.</p></li><li><p>Capacity needs planning, support scaling initiatives, and ensure our systems remain stable as the company grows.</p></li><li><p>SageMaker models need to be continuously delivered, following AWS' best practices.</p></li><li><p>AI engineers need assistance with productionalizing AI workloads end-to-end, from code to container build, and rollouts using optimal deployment strategies.<br></p></li></ul><h2>How we build<br></h2><p>We run a modern cloud-native stack to build, ship, and operate AI workloads reliably in a regulated healthcare environment.</p><ul><li><p>Cloud: AWS (The European Sovereign Cloud partition)</p></li><li><p>Runtime: Kubernetes and Postgres</p></li><li><p>Inference: Bedrock, SageMaker, ClearML</p></li><li><p>Observability: Datadog, LangFuse, OTEL</p></li><li><p>CI/CD: GitHub Enterprise Cloud with Actions</p></li><li><p>Plus various auxiliary tools for security, compliance and operations: DockerHub, Google as IdP, SonarCloud, VPNs, QuickSight, CDN, Snowflake, ETLs, etc.</p></li></ul><p style="min-height: 1.7em;"></p> <br> <h2>Requirements<br></h2><p>Here's how we picture you as our Senior Forward Deployed Platform Engineer:<br></p><p>These are the knowledge and skills you need:</p><ul><li><p>Experience with AWS (CloudFormation, IAM, ECR, SageMaker, Bedrock, S3, SQS, DynamoDB, RDS, KMS).</p></li><li><p>Experience with Kubernetes (EKS) in production (kustomize, HPA and KEDA based autoscaling).</p></li><li><p>Experience deploying SageMaker models, setting up custom inference containers (HuggingFace and/or OSS model) and endpoints (provisioning, autoscaling, and blue/green rollouts).</p></li><li><p>Familiarity with observability & LLMOps best practices and solutions (Datadog, Langfuse, LLM/GenAI tracing, OTel, CloudWatch).</p></li><li><p>Proficiency in Python, specifically with FastAPI and async task management.</p></li><li><p>Knowing how to set up ML/LLM GPU inference on EKS (GPU-backed instances, Nvidia driver/plugin).</p></li><li><p>Ability to own technical decisions and collaborate directly with other teams.</p></li><li><p>Troubleshooting, RCA, and resolving problems. When things get stressful, you stay cool headed and work through issues step by step, asking for support from SMEs when needed.</p></li><li><p>Preference for writing over talking, resulting in clear, concise, yet complete documentation and asynchronous communications.</p></li><li><p>This is a cross-functional role bridging the gap between applications and the platform. Ability to both explain complex concepts clearly to non-domain experts, and extract information through well-formed questions during verbal communication, is a must.<br></p></li></ul><p>Bonus points:</p><ul><li><p>Healthcare or regulated-domain experience (e.g. ISO 27001, C5).</p></li><li><p>Self-hosting/home lab.</p></li><li><p>Nvidia Triton/vLLM or similar.</p></li><li><p>MLflow, Kubeflow, ClearML, DVC, Vertex AI.</p></li><li><p>Slack and GitHub automations (bots, workflows, apps, etc.).<br></p></li></ul><p>This will help you decide if you want to join us – our values are:</p><ul><li><p>Strong opinions, weakly held. We adapt, negotiate trade-offs, and reconsider decisions when necessary. We don't stick to one implementation – we re-evaluate and adjust as needed (including throwing away the old one).</p></li><li><p>Pragmatic choice of tools and avoidance of one-size-fits-all thinking.</p></li><li><p>Bringing professionalism and empathy to how you work with others.</p></li><li><p>Strong belief in solid technical fundamentals and preference for understanding and first principles over memorization.</p></li><li><p>Language proficiency (at least English, but the more, the better).</p></li></ul><p style="min-height: 1.7em;"></p><p>Find more <a href="https://www.arbeitnow.com/english-speaking-jobs">English Speaking Jobs in Germany</a> on Arbeitnow</a>