Loading the SOTA2 catalog…
olmOCR: Unlocking Trillions of Tokens in PDFs with Vision Language Models · SOTA2 Research