Description of the companyOur client is a well-established financial institution known for its long-standing stability, international footprint, and commitment to operational excellence. The organisation fosters a collaborative culture, values expertise, and promotes continuous improvement across its global technology and data functions.
Role & ResponsibilitiesAs Senior AI Infrastructure & Platform Engineer, you will play a key role in designing, operating and continuously improving enterprise-grade AI infrastructure platforms supporting advanced machine learning and generative AI initiatives. You will work closely with infrastructure, platform engineering and data science teams to build highly available GPU-enabled environments and scalable container platforms capable of supporting large-scale model training and inference workloads.
Key Responsibilities:
- Design, implement and operate GPU-enabled infrastructure platforms for AI and machine learning workloads
- Administer and optimize Kubernetes and OpenShift platforms supporting enterprise AI services
- Manage high-performance computing (HPC) environments, including bare-metal and virtualized GPU systems
- Develop and maintain infrastructure automation using Ansible, Terraform and GitOps methodologies
- Troubleshoot complex platform issues involving Linux, networking, storage, Kubernetes and GPU architectures
- Support AI and data science teams with platform provisioning, resource allocation and operational excellence
- Implement platform monitoring and observability solutions
- Participate in incident management and on-call rotations
Candidate ProfileYou hold a university-level degree in Computer Science or a related discipline and bring several years of current, hands-on experience operating large-scale Linux and containerized infrastructure environments.
Additionally:
- Minimum 5 years of enterprise infrastructure engineering experience
- Minimum 3 years of hands-on experience administering Red Hat Enterprise Linux environments
- Proven experience operating Kubernetes and/or OpenShift platforms in production environments
- Strong hands-on experience with Terraform, ArgoCD and Ansible
- Practical experience supporting NVIDIA GPU infrastructures and GPU-enabled computing environments
- Knowledge of GPU architectures and technologies such as NVLink, PCIe switching and GPU virtualization
- Experience with Prometheus, Grafana and infrastructure monitoring solutions
- Strong Shell and Python scripting capabilities
- Experience working in highly regulated enterprise environments is advantageous
- Fluent in English; German advantageous
- Valid CH work permit or EU/EFTA citizenship required
- Willingness to work 100% on-site
Benefits:
- Central city office locations across Switzerland
- Above average insurance coverage fully borne our client
- Contribution to health insurance and meal allowance
- Excellent opportunities for further training and personal development
- International environment