AI engineering scientists that turn drawings into data for process design & engineering (YC S24)
Joined July 2024
- Structured Attention Matters to Multimodal LLMs in Document Understanding arxiv.org/abs/2506.21600
- doctors hate himAdding horizontal lines to images improves VLM (vision language model) performance of tasks like counting, visual search, spatial understating, scene understanding, and more


