Writing

Notes on multimodal models, retrieval-augmented generation, computer vision and whatever the PhD throws at me.

First post coming soon

I’m starting to write up what I learn — vision–language models, RAG systems that actually stay grounded, and notes from the lab. Check back, or follow along on GitHub.

Follow on GitHub