Multimodal AI & medical imaging
Vision–language models that read a scan and produce clinically faithful text. My MSc thesis generated region-guided clinical reports from chest X-rays using LLMs — the PhD at the Pattern Recognition Lab carries that line of work forward.