Senior Staff Data Engineer

Senior Staff Data Engineer

Full-Time 72000 - 88000 £ / year (est.) Home office (partial)
Boehringer Ingelheim GmbH

At a Glance

  • Tasks: Lead data engineering architecture and standards for groundbreaking AI projects in healthcare.
  • Company: Join Boehringer Ingelheim, a top employer dedicated to innovative therapies and exceptional workplace culture.
  • Benefits: Enjoy a hybrid work model, competitive salary, and opportunities for professional growth.
  • Other info: Collaborate with a dynamic team and shape the future of healthcare technology.
  • Why this job: Make a real impact on disease understanding and therapeutic development through cutting-edge data engineering.
  • Qualifications: PhD or MSc in STEM with extensive data engineering experience and strong biomedical knowledge.

The predicted salary is between 72000 - 88000 £ per year.

Most diseases are still poorly understood at a biological level. Despite decades of research, the causal mechanisms driving many conditions remain unclear, limiting our ability to identify the right targets, design the right interventions and bring the right medicines to patients. The AI Accelerator exists to change that. Based in London and sitting within Computational Innovation, a global organisation spanning computational biology, human genetics, data excellence and AI, the Accelerator's mission is to build production-quality AI capabilities that deepen our understanding of disease biology and increase probability of success.

We do this by applying neural-based methods across the biomedical data landscape to integrate heterogeneous, multimodal data sources, infer biological relationships and embed causal thinking into what we build. The goal is not just to predict but to explain and understand why disease occurs. It could be electronic health records and medical imaging to support patient segmentation. It could be 'omics data to identify novel therapeutic targets. It could be predicting transcriptional change for a given disease-causing variant. It could be simulating the effect of modulating a target of interest.

None of this is possible without data. The quality, accessibility and engineering of biomedical data ultimately determine the quality of the models built upon them. The Data Engineering function provides the foundation on which the AI Accelerator operates by integrating diverse clinical, biological and real-world datasets into trustworthy, AI-ready assets that can support large-scale foundation model development and downstream scientific discovery.

Key Responsibilities

  • Set the technical direction, strategy and roadmap for data engineering, aligned to AI Accelerator priorities.
  • Own the data engineering architecture for the AI Accelerator, defining and evolving a layered architecture, feature and embedding provisioning patterns, and the harmonised multimodal data foundation that underpins model development.
  • Establish data engineering standards and engineering practices, including data quality controls, testing, CI/CD for data, reproducibility standards, data contracts, metadata management and dataset versioning.
  • Deliver hands-on leadership on the most difficult and highest-impact data engineering challenges, including integrating novel and complex modalities such as genomics, transcriptomics, imaging, clinical and real-world datasets.
  • Partner with the Data Excellence community and IT to align on governance, ontologies, metadata harmonisation, stewardship and shared enterprise data foundations.
  • Establish ways of working and coach other team members, onboarding, mentoring and technically leading data engineers as the team scales, while acting as the senior escalation point for complex data engineering challenges.

Requirements

  • PhD or MSc and equivalent experience in a STEM subject.
  • Extensive experience operating at the senior staff level within a data engineering function, including defining architecture, standards and technical direction, along with mentoring data engineers and establishing strong engineering principles within a growing team.
  • Deep expertise in large-scale data engineering, including distributed processing frameworks, pipeline orchestration, cloud data platforms and data integration at scale, with experience integrating internal and third-party datasets and working effectively with external data providers and technology partners.
  • Strong understanding of biomedical and healthcare data domains, such as genomics, transcriptomics, multi-omics, imaging, clinical data, electronic health records or real-world data, alongside an understanding of machine learning data requirements, including versioning, reproducibility and tensor-based workflows for large-scale AI systems.
  • Experience implementing data governance in practice, including metadata management, lineage, provenance, ontologies, cataloguing and FAIR principles and familiarity with Trusted Research Environments (TREs) and controlled-access research data environments.
  • Strong collaboration and influencing skills across technical and non-technical stakeholders, with the ability to communicate complex technical concepts clearly.

This is a hybrid role with approximately 4 days a week in the office.

Why This Is A Great Place To Work

Boehringer Ingelheim has been recognised as a Top Employer in the UK, demonstrating our commitment to building an exceptional workplace through strong people practices and supportive HR policies.

Senior Staff Data Engineer employer: Boehringer Ingelheim GmbH

Boehringer Ingelheim is an exceptional employer, recognised as a Top Employer in the UK, offering a supportive work culture that prioritises employee well-being and professional growth. As a Senior ML Engineer in London, you will be at the forefront of biomedical AI, collaborating with leading scientists and engineers to make impactful contributions to human health while enjoying a hybrid work model that promotes work-life balance.

Boehringer Ingelheim GmbH

Contact Details:

Boehringer Ingelheim GmbH Recruitment Team

StudySmarter Expert Advice🤫

We think this is how you could land Senior Staff Data Engineer

Get Involved in Data Science Meetups

Tap into local data science meetups or workshops to connect with fellow enthusiasts and professionals. These events are goldmines for networking, and sometimes even lead directly to job openings at companies like Boehringer Ingelheim GmbH!

Show Off Your Projects

Start building a public portfolio showcasing your data science projects on platforms like GitHub or personal websites. Highlight unique analyses or models you've developed. This not only demonstrates your skills but also gets your name out there for roles like Senior Staff Data Engineer at Boehringer Ingelheim GmbH.

Leverage Professional Networks

Join professional bodies related to data science, like the Data Science Society or similar organisations. Getting involved can lead to mentorship opportunities and insider knowledge about full-time positions at companies like Boehringer Ingelheim GmbH.

Apply Directly through Our Website

When you find a suitable opening like Senior Staff Data Engineer at Boehringer Ingelheim GmbH, make sure to apply directly through our website. It gives you an edge and shows you're keen to join our team. Plus, who doesn’t love a direct application? It’s easier than navigating through job boards!

We think you need these skills to ace Senior Staff Data Engineer

Python
SQL
Communication Skills
Automation
Data Engineering
Data Pipeline Development
API Integration

Some tips for your application 🫡

Show Off Your Projects:In the world of data science, your projects can speak volumes about your skills. Make sure to showcase a few key projects in your CV or portfolio, especially those that highlight your ability to work with data sets, build models, or use relevant tools like Python, R, or SQL. Don’t forget to include links to any GitHub repositories if applicable!

Quantify Your Achievements:Employers love numbers! When drafting your CV, highlight your achievements with quantifiable results. For instance, mention how your data analysis led to a certain percentage increase in efficiency or revenue at a previous job or project. These details can really make your application pop!

Craft a Tailored Cover Letter:For a full-time role at Boehringer Ingelheim GmbH, your cover letter should reflect your passion for data science and your excitement about the specific projects or values of the company. Dive into why you’re a good fit, how your skills align with their needs, and any unique perspectives you can bring to the team.

Stand Out with Relevant Courses and Certifications:Although experience talks, relevant courses or certifications can be your ticket to impressing hiring managers at Boehringer Ingelheim GmbH. Mention any standout courses you've completed that equipped you with essential skills, such as machine learning certifications or data visualisation courses. This shows your commitment to continuously developing your skills in the field!

How to prepare for a job interview at Boehringer Ingelheim GmbH

Brush Up on Your Statistics

For a data science role, we need to seriously sharpen our statistics skills. Get ready to tackle technical questions on probability distributions, hypothesis testing, and regression analysis. These are often the bread and butter of data science interviews, so don't just skim over them!

Showcase Your Projects

Prepare a killer portfolio showcasing your data science projects. We should include details about the datasets used, the tools and techniques applied, and the impact of your findings. If we can walk them through a particularly challenging project or a cool visualisation that had real-world implications, it’ll really make us stand out!

Get Comfortable with Python and R

Most data science positions require us to be proficient in programming languages like Python and R. We should practice common libraries like pandas, NumPy, and scikit-learn, and be ready for live coding exercises or algorithm questions. Showing off our coding chops can really impress the interviewers at Boehringer Ingelheim GmbH!

Prepare for Case Studies

Expect to encounter real-world case studies during the interview. We might be asked how we’d approach a data problem or analyse a dataset to extract insights. It's essential to think out loud and demonstrate our problem-solving process so that the interviewer can see our logical thinking in action.