Skills
Experience
Visual Inference Lab, TU Darmstadt
https://www.visinf.tu-darmstadt.deResearcher at Visual Inference Lab, working on robust semantic analysis of heterogeneous scenes from multi-modal data streams
IT-Consultant
http://www.patrip.orgfor PATRIP Foundation: Planning, implementation and maintaining of a data exchange platform
SysAdmin
at several companies: Administration of Linux/Unix servers, setup and management of high availability servers and loadbalancers
Publications
Doppio: A Dataset for Contactless Weight Estimation of Falling Particles
Simon Kiefhaber, Jan-Martin O. Steitz, Julia Grabinski, Christoph Reich, Paul Wagner, Max Zimmermann, Simone Schaub-Meyer, and Stefan Roth,
in
Proc. of the 48th DAGM German Conference on Pattern Recognition (GCPR),
2026.
Measuring the mass of powder, including falling particles, is a common task in industrial applications. While scales are effective for static measurements, many applications require contactless …
Adapters Strike Back
Jan-Martin O. Steitz, and Stefan Roth,
in
Proc. of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR),
2024.
Adapters provide an efficient and lightweight mechanism for adapting trained transformer models to a variety of different tasks. However they have often been found to be outperformed by other …
xGQA: Cross-Lingual Visual Question Answering
Jonas Pfeiffer, Gregor Geigle, Aishwarya Kamath, Jan-Martin O. Steitz, Stefan Roth, Ivan Vulić, and Iryna Gurevych,
in
Findings of the Association for Computational Linguistics (ACL),
2022.
Recent advances in multimodal vision and language modeling have predominantly focused on the English language, mostly due to the lack of multilingual multimodal datasets to steer modeling …
