DeepMind in London seeks a Tech Lead for Gemini Inference Performance to guide a team of performance engineers and drive optimization across ML frameworks and hardware accelerators.
You will leverage transformer and Mixture-of-Experts concepts to improve latency and memory efficiency while collaborating with research teams on deployment trade-offs. A strong leadership track record is essential.
#J-18808-Ljbffr
Tech Lead, ML Inference Performance employer: Google LLC
DeepMind, as part of Google, offers an exceptional work environment that fosters innovation and collaboration among top-tier professionals in AI and materials science. Located in a vibrant tech hub, employees benefit from a culture that prioritises continuous learning and growth, alongside access to cutting-edge resources and interdisciplinary projects that drive meaningful advancements in semiconductor research.