Associate Manager, GenAI & Enterprise Solutions · Accenture, Düsseldorf
M.Sc. Informatik, RWTH Aachen University — Current building agentic and multimodal pipelines on industrial GenAI platform.
00:01 / About
Guten Tag! I am currently working as an Associate Manager at Accenture. Prior to joining Accenture, I served as an IDEA Research Grant Student in Prof. Dr. Manfred Claassen's group at ETH Zürich, and I obtained my Master's degree in Informatik from RWTH Aachen University and my Bachelor's degree from HFUT. During my Bachelor's studies I specialized in IoT and founded the HFUT RoboCup Lab. I also gained industry experience as a computer vision working student at the Siemens AG Aachen Gas Turbine Research Center during my Master's studies.
My research interests center on object recognition and segmentation, the deployment of machine learning systems, model interpretability, and multi-task & multimodal learning in industrial production.
00:02 / News
00:03 / Research
ETH Claassen Lab
Feedback from mentor at ETH Zürich →Sinovation Venture AI Institute
Feedback from mentor at Sinovation →RWTH Computer Vision Group
Feedback from mentor at RWTH →00:04 / Reviewing
Conferences
Journals
Review categories
00:05 / Honors
00:06 / Publications

ACL 2026 · Workshop on Customized and Personalized NLP (CustomNLP4U) · Poster Paper
A domain-adapted video-to-video clip generation framework combining Audio- and Vision-Language Models to produce highlight clips: a reproducible Cut & Merge algorithm with fade transitions and timestamp normalization, a role-based personalization mechanism for marketing/training/regulatory outputs, and a cost-efficient end-to-end pipeline. Evaluated on Video-MME and a proprietary set of 16,159 pharmacy videos across 14 disease areas, showing 3–4× speedup and 4× cost reduction over VLM baselines such as Gemini 2.5 Pro.

ACL 2026 · Workshop on Advances in Language and Vision Research (ALVR) · 🌟 Spotlight Paper
An industrial GenAI framework processing 200,000+ PDFs, 25,326 videos across eight formats, and 888 multilingual audio files in 20+ languages. Contributions: a large-scale architecture for multimodal reasoning in pharma; empirical analysis of 40+ VLMs on Video-MME, MMBench, and a proprietary 25,326-video set across 14 disease areas; and findings on multimodality, attention trade-offs, temporal reasoning limits, and video-splitting under GPU constraints.

NAACL 2025 · Industry Track
A cross-lingual benchmark of nearly 99,869 real traffic incident records from Vienna (2013–2023) assessing LLM robustness across spatial and temporal domains. Explores three hypotheses — sentence indexing, date-to-text conversion, and German-to-English translation — and incorporates RAG to further examine hallucination in both domains.
.png)
CVPR 2024 · Synthetic Data for Computer Vision Workshop · Full paper
An explainable visual dataset of 200K+ images with 15,410 manual semantic annotations across 12 ImageNet categories representing everyday objects. Incorporates six real-world scenarios (overexposure, blurring, color shifts, etc.) and a quantitative robustness criterion, with particular attention to background influence.
.png)
INDIN 2023 · IEEE 21st Int'l Conference on Industrial Informatics · Full paper
A self-supervised training method for visual transformers that learns meaningful image/video representations without large labeled sets, based on exact solutions of generated representations. The learned features fine-tune effectively on industrial downstream tasks.

ICASSP 2023 · Full paper
SleepHGNN, a deep model for sleep-stage classification using a Heterogeneous Graph Transformer to capture interactivity and heterogeneity across multimodal signals — the first attempt to apply heterogeneous graph neural networks to this task, with a graph-level classification framework generalizable to domains like protein and molecular graph classification.

NeurIPS 2022 · DMML Workshop
A comparison of ML deployment solutions spanning manual training/deployment through automated continuous-integration workflows, with practical evaluation metrics and a look at how real-world requirements diverge from academic settings.

ICLR 2022 · CSS Workshop
An exploration of how image backgrounds can help object recognition, building on the "noise or signal" baseline work by Xiao et al.
BMW Group · Industrial deployment
AIQX integrates machine learning and deep learning algorithms for visual inspection directly into BMW's production processes — a central standard for AI-based quality inspection across the global production system, enabling more robust defect detection and order verification.

ICLR 2021 · Workshop on AI for Public Health
DeeCamp 2020 · Medical Track of AI in Public Healthcare — Challenge winner

Sinovation Ventures AI Institute (创新工场), 2020
Designed three GPT-based generative models for real-world business scenarios.
00:07 / Notes
Hobbies: vlogger who loves museums and cooking — find me on TikTok / Red (小红书) / WeChat Channel as "Jonas的新鲜感" (Jonas' curiosity). The channel has received millions of views and likes 👍, and thousands of followers, across 100+ vlogs. Keeping learning — let's move on together!
I'm also a fan of hackathons — they've given me valuable experience and sharpened my critical thinking, problem-solving, and leadership skills.