Tech Lead, ML Inference Performance

Tech Lead, ML Inference Performance

Full-Time On-site
G

DeepMind in London seeks a Tech Lead for Gemini Inference Performance to guide a team of performance engineers and drive optimization across ML frameworks and hardware accelerators.

You will leverage transformer and Mixture-of-Experts concepts to improve latency and memory efficiency while collaborating with research teams on deployment trade-offs. A strong leadership track record is essential.

#J-18808-Ljbffr

Tech Lead, ML Inference Performance employer: Google LLC

DeepMind, as part of Google, offers an exceptional work environment that fosters innovation and collaboration among top-tier professionals in AI and materials science. Located in a vibrant tech hub, employees benefit from a culture that prioritises continuous learning and growth, alongside access to cutting-edge resources and interdisciplinary projects that drive meaningful advancements in semiconductor research.

G

Contact Details:

Google LLC Recruitment Team