Helsing logo
Posted Aug 20Berlin; London; Munich
Apply ↗

AI Research Engineer - AI Safety

MiddleOn-site (Berlin)€60,000 – €80,000 / yr
Required Skills
PythonRustMachine LearningLLMs / Generative AI
Job Description

Who we are

At Helsing we deliver AI-based capabilities and the enabling foundation that allow machines to perceive and assist human decision-making. You will have the unique opportunity to shape AI capabilities in one of the most challenging sectors, where high generalisation capabilities need to be paired with hardware constraints and robustness against adversarial attacks.

You will join a team focused on AI Assurance, where you will develop cutting-edge techniques for scalable evaluation of AI products across the company, design data collection and experimentation strategies to extract causal insights, and enhance responsible decision-making via uncertainty quantification and safety mechanisms.

The Role

At Helsing we deliver AI-based capabilities and the enabling foundation that allow machines to perceive and assist human decision-making. You will have the unique opportunity to shape AI capabilities in one of the most challenging sectors, where high generalisation capabilities need to be paired with robustness against adversarial attacks and the highest standards of operational safety.

You will be responsible for defining operational domains and evaluating the reliability of AI capabilities developed in-house. Your work will span the full assurance lifecycle: from characterising distribution shifts and failure modes, to developing and extending the state of the art in uncertainty quantification and calibration. You will interface deeply with our AI systems, design rigorous evaluation frameworks, and assess their robustness under real-world and adversarial conditions, collaborating across research, engineering, and product teams to translate assurance findings into actionable improvements.

You should apply if you

  • Hold an MSc in Mathematics, Statistics, Machine Learning, or a closely related field, with a strong mathematical and statistical foundation.

  • Have hands-on experience in model evaluation, uncertainty quantification, or calibration. You understand the difference between epistemic and aleatoric uncertainty and know how to measure and reduce them in deep learning models.

  • Are familiar with methods for distribution shift detection, out-of-distribution detection, and adversarial robustness evaluation, and can design experiments that surface genuine failure modes rather than benchmark artefacts.

  • Possess solid software engineering skills, writing clean and well-structured code in Python and/or languages like Rust or modern C++, and have experience deploying AI software to production including testing, QA, and monitoring.

  • Have excellent communication skills and the ability to report and present research findings clearly and efficiently, both internally and externally.

  • Are passionate about keeping up to date with current research and enjoy reimplementing and extending state-of-the-art approaches in deep learning evaluation and assurance.

Note: We operate at an inter

Ready to apply? Optimize your CV for this specific jobAI customizes your experience bullets and increases chances to get hired.

Similar Openings in Other

View all in category