Structured programmes for software engineers, data science and DevOps.
Senior Platform Engineer AI & Observability (m/f/d)
Get alerts for roles like this
More Senior Platform Engineer AI & Observability (m/f/d) roles in Garching bei München — straight to your inbox. No account needed.
Applying to this role? Tailor your résumé to this job description in one click, then download it clean — no watermark, no subscription.
Job Description
Description: Looking for an employer you can count on? Join us! Senior Platform Engineer AI & Observability (m/f/d) Your Role and Responsibilities: Design, deploy, and operate scalable AI, observability, and cloud-native platforms based on Kubernetes and HPC technologies.
Build and optimize AI services, LLM inference platforms, and GPU-enabled workloads. Develop and maintain monitoring, logging, tracing, and security solutions using open-source technologies. Create standardized deployment workflows, automation, and platform best practices.
Enable reliable, secure, and multi-tenant operation of federated research infrastructures. Collaborate with project partners and provide technical leadership in architecture, implementation, and operations.
Your Qualifications: Required/Minimum Qualifications Master’s degree (or equivalent) in Computer Science, Data Science, Computer Engineering, or a related field. Other Requirements Experience with Linux, Docker, Kubernetes, and cloud-native technologies.
Knowledge of observability, monitoring, logging, tracing, and security concepts. Programming and scripting skills, preferably in Python, Go, or Bash. Experience with DevOps, MLOps, platform engineering, or infrastructure automation. Strong communication, collaboration, and problem-solving skills.
Excellent written and spoken English. Additional or Preferred Qualifications: Experience with AI, machine learning, LLMs, or AI-assisted operations. Hands-on experience with inference frameworks such as vLLM, Triton, Ollama, or Llama. cpp.
Experience operating GPU-accelerated, large-scale Kubernetes, or HPC environments. Knowledge of Prometheus, Grafana, Helm, GitOps, and CI/CD. Familiarity with RAG architectures, vector databases, MCP-based services, or AI agents. Experience with security monitoring tools (such as Falco or Tracee).
Experience in research projects or open-source software development. Conditions: Working time full time (40 hrs) flexible working model with electronic time
Required Skills
Upskill for This Role
Courses from Udemy and edX matched to this role's skills.

Linux Administration & Bash Scripting: Practice Exams

Bash Scripting and Shell Programming (Linux Command Line)

Curso de Servidores Linux en CENTOS para principiantes

Bootcamp Terminal Linux dan CMD

Linux Professional Institute LPIC-1 102-500 Exam 2026

これだけでOK!CompTIA Linux+ (XK0-006) 合格講座:基礎から実務まで徹底解説
ResumeKart may earn a commission from these links at no extra cost to you.
Partner picks for Senior Platform Engineer AI & Observability (m/f/d) in Garching bei München
Matched to the skills this page calls for and the candidate's location.
- Partner course providercovers pythoncovers machine learningcovers devopscovers kubernetesBengaluruVisit partner →
- edXVerified partnerPartner course provider
Courses and programmes from universities and institutions worldwide.
covers pythoncovers machine learning - UdemyVerified partnerPartner course provider
A marketplace of instructor-created courses across technology, business and creative skills.
covers pythoncovers docker
Partners are ResumeKart affiliates or institutes it works with; ResumeKart may earn a commission when a candidate enrols. Placement is decided by relevance, not payment. How ResumeKart earns
Prepare to Win This Role
Everything you need to ace the interview and negotiate top-of-band compensation.