Senior AI Infrastructure & Platform Engineer (Basel)
AlbedisDescription of the company
- Design, implement and operate GPU-enabled infrastructure platforms for AI and machine learning workloads
- Administer and optimize Kubernetes and OpenShift platforms supporting enterprise AI services
- Manage high-performance computing (HPC) environments, including bare-metal and virtualized GPU systems
- Develop and maintain infrastructure automation using Ansible, Terraform and GitOps methodologies
- Troubleshoot complex platform issues involving Linux, networking, storage, Kubernetes and GPU architectures
- Support AI and data science teams with platform provisioning, resource allocation and operational excellence
- Implement platform monitoring and observability solutions
- Participate in incident management and on-call rotations
- Minimum 5 years of enterprise infrastructure engineering experience
- Minimum 3 years of hands-on experience administering Red Hat Enterprise Linux environments
- Proven experience operating Kubernetes and/or OpenShift platforms in production environments
- Strong hands-on experience with Terraform, ArgoCD and Ansible
- Practical experience supporting NVIDIA GPU infrastructures and GPU-enabled computing environments
- Knowledge of GPU architectures and technologies such as NVLink, PCIe switching and GPU virtualization
- Experience with Prometheus, Grafana and infrastructure monitoring solutions
- Strong Shell and Python scripting capabilities
- Experience working in highly regulated enterprise environments is advantageous
- Fluent in English; German advantageous
- Valid CH work permit or EU/EFTA citizenship required
- Willingness to work 100% on-site
- Central city office locations across Switzerland
- Above average insurance coverage fully borne our client
- Contribution to health insurance and meal allowance
- Excellent opportunities for further training and personal development
- International environment