Ellison Institute of Technology (EIT) is seeking an experienced MLOps engineer to build and operate scalable GPU training and inference clusters. You’ll drive robust scheduling, isolation, and lifecycle management across high-performance hardware to accelerate research.
You will design data pipelines, optimise I/O and storage (Lustre), and ensure observability, security, and cost controls while collaborating with Research, Data, and Applied teams to forecast capacity and ML experimentation
#J-18808-Ljbffr