PERSON DIRECTORY
Will Hardman
Will Hardman appears in 1 indexed conversation across The Cognitive Revolution. This directory brings every appearance, source, TL;DR, digest, and transcript into one searchable feed.
Teaching AI to See: A Technical Deep-Dive on Vision Language Models with Will Hardman of Veratai
Vision-language models are becoming a platform layer for medical assistance, insurance verification, document indexing and robotics, although multimodality’s necessity for AGI remains unresolved.The strongest competitive mechanisms are high-quality interleaved data, synthetic instruction tuning and modular fusion, with Chinese open models such as InternVL 2.5 and Qwen2-VL highly competitive with proprietary systems.Benchmark leadership is fragmented: o1 led cited MMMU results near 78%, Qwen2-VL reached 96.5% on DocVQA, while BLINK exposed persistent gaps in counting, visual IQ and illumination.
