Computer Vision Researcher (VLM)

Computer Vision Researcher (VLM)

Full-Time 80000 - 100000 £ / year (est.) No working from home possible
D

At a Glance

  • Tasks: Lead groundbreaking research in 3D computer vision and language models.
  • Company: Innovative tech firm at the forefront of spatial intelligence.
  • Benefits: Hybrid work model, competitive salary, and opportunities for professional growth.
  • Other info: Dynamic London R&D hub with mentorship opportunities.
  • Why this job: Join a team shaping the future of AI and spatial reasoning.
  • Qualifications: PhD in relevant field and 4+ years of ML research experience.

The predicted salary is between 80000 - 100000 £ per year.

Building a living model of the world that people and machines can talk to. Powered by a proprietary database of over 30 billion posed images and a next-gen digital map, they are developing the spatial intelligence that helps humans and machines understand, navigate, and engage with the physical world.

As a Technical Anchor in their London R&D hub, you will bridge the gap between 3D computer vision and Vision-Language Models (VLMs), creating a unified framework where machines can reason about their surroundings.

What You’ll Be Doing:

  • Architect Semantic Grounding: Lead research into cross-modal grounding connecting 3D spatial features with language embeddings.
  • Scale 'Understand' Capabilities: Develop algorithms for continuous semantics, allowing 3D maps to evolve and improve situational awareness.
  • Agentic Frameworks: Build the 'spatial brain' for Embodied AI, enabling robots, drones, and machines to move into mission-level reasoning.
  • Multimodal Benchmarking: Define standards for measuring 'spatial common sense' in LLMs/VLMs.
  • Technical Mentorship: Act as the technical anchor for the London hub, guiding architecture and mentoring researchers.

What We Are Looking For:

  • Education: PhD (or equivalent) in Computer Vision, Machine Learning, or Robotics focusing on Multimodal/Semantic understanding.
  • Experience: 4+ years of ML research experience with a track record of shipping models bridging 3D Vision and Language.
  • Technical Depth: Expert knowledge of 3D Geometry (SfM, SLAM, VPS) and Transformer-based architectures (VLMs).
  • Research Impact: Multiple first-author publications at top-tier venues (CVPR, NeurIPS, ICLR).
  • Code Mastery: Production-quality research code in PyTorch or JAX + large-scale data pipeline management.
  • Location: Ability to work hybrid from their London office (3 days/week).

Computer Vision Researcher (VLM) employer: DeepRec.ai

At DeepRec.ai, we pride ourselves on being an exceptional employer, offering a dynamic work culture that fosters innovation and collaboration in the heart of Greater London. Our commitment to employee growth is evident through our focus on cutting-edge AI technologies and the opportunity to lead transformative projects that have a real impact on the future of science. With competitive compensation, a supportive environment, and the chance to work alongside industry leaders, we provide a unique platform for engineers passionate about making a difference in the world of AI.

D

Contact Details:

DeepRec.ai Recruitment Team

We think you need these skills to ace Computer Vision Researcher (VLM)

3D Computer Vision
Vision-Language Models (VLMs)
Semantic Grounding
Algorithm Development
Multimodal Understanding
3D Geometry (SfM, SLAM, VPS)
Transformer-based Architectures