Hi everyone! I have a question for folks who broke above .91, I am not looking for exact setups, just curious about the general direction people went. Are you mostly using OCR-specific models, general-purpose VLMs, or fully custom-trained architectures? Even a rough breakdown would help. Thanks
@sdv I would appreciate your take on this :)