Personeel.com

AI Infrastructure Systems Engineer (Amsterdam & London)

Amsterdam30+ dagen geleden

HybrideEngelstalig

Over de functie

Join Together AI to build the backbone of AI infrastructure, designing and automating systems for large-scale GPU clusters. This hybrid role in Amsterdam focuses on maximizing performance and reliability across hardware and software.

Wat ga je doen

  • Design and build fleet automation systems for GPU clusters
  • Build AI Infrastructure Agents for automated deployment and remediation
  • Develop Fleet Intelligence platforms for hardware health monitoring
  • Maximize GPU availability, utilization, performance, and reliability
  • Create automated validation systems for GPUs and networking
  • Build internal platforms and developer tools for infrastructure management

Wat je meebrengt

  • 3+ years building distributed systems or infrastructure platforms
  • Strong software engineering in Python, Go, or Rust
  • Experience with platforms, automation systems, or developer infrastructure
  • Experience with Linux, Kubernetes, Terraform, or Ansible
  • Strong systems thinking across hardware and software
  • Automation-first mindset

Mooi meegenomen

  • GPU infrastructure, CUDA, NCCL, NVLink/NVSwitch
  • InfiniBand or RoCE networking
  • Bare-metal provisioning and lifecycle management
  • Large-scale AI training or inference clusters
  • Hardware health monitoring and predictive failure detection
  • Distributed storage systems

Samenvatting door Personeel.com. De volledige vacaturetekst en het sollicitatieformulier staan op de website van Together AI (job-boards.greenhouse.io); met Solliciteer ga je er direct heen.

Meer vacatures bij Together AI

Alle 10 vacatures bij Together AI
Solliciteer