VDInstruct: Zero-Shot Key Information Extraction via Content-Aware Vision Tokenization

Open in new window