Senior Inference Architect β€” Low-Latency, On-Device & GPU

Senior Inference Architect β€” Low-Latency, On-Device & GPU

Full-Time On-site
G

Genesis AI is seeking an expert in distributed systems and ML infrastructure to build and optimize low-latency inference pipelines for on-device robotics. The role focuses on GPU-accelerated architectures and integrating high-performance kernels into user-friendly frameworks.

Candidates should demonstrate production-grade Python skills with a strong background in C++/Rust/Go, and a track record of scaling workloads across both on-device and cluster environments.

#J-18808-Ljbffr

Senior Inference Architect β€” Low-Latency, On-Device & GPU employer: Genesis AI

Genesis is an exceptional employer for those looking to make a significant impact in the AI and robotics field. With a competitive salary, meaningful equity, and a hybrid work model, employees enjoy a dynamic work culture that fosters innovation and collaboration. The opportunity to shape the recruiting function from the ground up, alongside industry leaders, ensures that team members can grow professionally while tackling some of the most challenging problems in technology today.

G

Contact Details:

Genesis AI Recruitment Team