Excited to share Nanonets-OCR2, a state-of-the-art suite of models designed for advanced image-to-markdown conversion and Visual Question Answering (VQA).
Wow, OCR is now basically a general domain. I remember when I spent like a year trying to create one for receipts. Took me 6 months of data curation to prepare.
3 comments of 4
[ 3.6 ms ] story [ 23.6 ms ] threadLive Demo -> https://docstrange.nanonets.com/
Blog -> https://nanonets.com/research/nanonets-ocr-2/
Nice job, the scores are superb.