Skip to main content
ResumeKart
← Back to Jobs

Senior Platform Engineer AI & Observability (m/f/d)

Lrz•Garching bei München
Full-time3-7
👁️ 0 views•📝 0 applications•Posted 10/4/2026•Expires 11/3/2026
Tailor Resume for This JobCheck ATS Score

Get alerts for roles like this

More Senior Platform Engineer AI & Observability (m/f/d) roles in Garching bei München — straight to your inbox. No account needed.

Applying to this role? Tailor your résumé to this job description in one click, then download it clean — no watermark, no subscription.

Job Description

Description: Looking for an employer you can count on? Join us! Senior Platform Engineer AI & Observability (m/f/d) Your Role and Responsibilities: Design, deploy, and operate scalable AI, observability, and cloud-native platforms based on Kubernetes and HPC technologies.

Build and optimize AI services, LLM inference platforms, and GPU-enabled workloads. Develop and maintain monitoring, logging, tracing, and security solutions using open-source technologies. Create standardized deployment workflows, automation, and platform best practices.

Enable reliable, secure, and multi-tenant operation of federated research infrastructures. Collaborate with project partners and provide technical leadership in architecture, implementation, and operations.

Your Qualifications: Required/Minimum Qualifications Master’s degree (or equivalent) in Computer Science, Data Science, Computer Engineering, or a related field. Other Requirements Experience with Linux, Docker, Kubernetes, and cloud-native technologies.

Knowledge of observability, monitoring, logging, tracing, and security concepts. Programming and scripting skills, preferably in Python, Go, or Bash. Experience with DevOps, MLOps, platform engineering, or infrastructure automation. Strong communication, collaboration, and problem-solving skills.

Excellent written and spoken English. Additional or Preferred Qualifications: Experience with AI, machine learning, LLMs, or AI-assisted operations. Hands-on experience with inference frameworks such as vLLM, Triton, Ollama, or Llama. cpp.

Experience operating GPU-accelerated, large-scale Kubernetes, or HPC environments. Knowledge of Prometheus, Grafana, Helm, GitOps, and CI/CD. Familiarity with RAG architectures, vector databases, MCP-based services, or AI agents. Experience with security monitoring tools (such as Falco or Tracee).

Experience in research projects or open-source software development. Conditions: Working time full time (40 hrs) flexible working model with electronic time

Required Skills

LinuxDockerKubernetescloud-native technologiesPythonGoBashDevOpsMLOpsplatform engineeringinfrastructure automationAImachine learningLLMsvLLMTritonOllamaLlama.cppGPU-acceleratedHPCPrometheusGrafanaHelmGitOpsCI/CDRAG architectures

Partner picks for Senior Platform Engineer AI & Observability (m/f/d) in Garching bei München

Matched to the skills this page calls for and the candidate's location.

Partner
  • Partner course provider

    Structured programmes for software engineers, data science and DevOps.

    covers pythoncovers machine learningcovers devopscovers kubernetes
  • edXVerified partner
    Partner course provider

    Courses and programmes from universities and institutions worldwide.

    covers pythoncovers machine learning
  • UdemyVerified partner
    Partner course provider

    A marketplace of instructor-created courses across technology, business and creative skills.

    covers pythoncovers docker

Partners are ResumeKart affiliates or institutes it works with; ResumeKart may earn a commission when a candidate enrols. Placement is decided by relevance, not payment. How ResumeKart earns

The best-paying roles in your field. Every week. Free.

Join 10,000+ professionals getting job alerts and salary insights in their inbox

We respect your privacy. Unsubscribe anytime with one click.