Writing
Notes on multimodal models, retrieval-augmented generation, computer vision and whatever the PhD throws at me.
First post coming soon
I’m starting to write up what I learn — vision–language models, RAG systems that actually stay grounded, and notes from the lab. Check back, or follow along on GitHub.
Follow on GitHub