Kyle Harrison
research-paper

Semantic Composition in Visually Grounded Language Models

Rohan Shrirang Pandey 2023 View original ↗

TL;DR — A Carnegie Mellon honours undergraduate thesis on whether vision-language models compose meaning the way humans do.

How much weight it carries: An undergraduate thesis. Careful work, but a thesis, not a peer-reviewed result.

Where this came from

50 pages. A copy is archived locally against link rot; the header links the original source.