Hacker Newsnew | past | comments | ask | show | jobs | submitlogin
Nanonets-OCR2-3B – OCR model that transforms documents into structured markdown (huggingface.co)
13 points by PixelPanda 11 months ago | hide | past | favorite | 4 comments


Wow, OCR is now basically a general domain. I remember when I spent like a year trying to create one for receipts. Took me 6 months of data curation to prepare.

Nice job, the scores are superb.


Yes, and its not just OCR (Optical Character Recognition), it understands layouts, captures signatures, charts, watermarks etc so way beyond just characters


Excited to share Nanonets-OCR2, a state-of-the-art suite of models designed for advanced image-to-markdown conversion and Visual Question Answering (VQA).

Live Demo -> https://docstrange.nanonets.com/

Blog -> https://nanonets.com/research/nanonets-ocr-2/


wow. so many use cases. nice job.




Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: