Popular repositories Loading
-
docstrange
docstrange PublicExtract and convert data from any document, images, pdfs, word doc, ppt or URL into multiple formats (Markdown, JSON, CSV, HTML) with intelligent structured data extraction and advanced OCR.
-
nanonets-ocr-sample-python
nanonets-ocr-sample-python PublicNanoNets OCR API Example for Python
-
RaspberryPi-ObjectDetection-TensorFlow
RaspberryPi-ObjectDetection-TensorFlow PublicObject Detection using TensorFlow on a Raspberry Pi
-
ocr-with-tesseract
ocr-with-tesseract PublicA comprehensive tutorial for OCR in python using Tesseract-OCR and OpenCV
-
ocr-python
ocr-python PublicOCR library to extract text & tables from PDF files and images. Convert any image or PDF to CSV / TXT / JSON / Searchable PDF.
Repositories
- idp-leaderboard-benchmarks Public
Prediction cache generation and evaluation pipeline for idp-leaderboard.org
- docstrange-python Public
- nanoindex Public
Agentic RAG Harness for long documents, Tree and Graph based reasoning. Cited answers down to the pixel
- docext Public
An on-premises, OCR-free unstructured data extraction, markdown conversion and benchmarking toolkit. (https://idp-leaderboard.org/)
- n8n-nodes-nanonets Public
- docstrange Public
Extract and convert data from any document, images, pdfs, word doc, ppt or URL into multiple formats (Markdown, JSON, CSV, HTML) with intelligent structured data extraction and advanced OCR.
- llm-data-converter Public
Convert any document format into LLM-ready data format (markdown) with advanced intelligent document processing capabilities powered by pre-trained models.
Top languages
Loading…
Most used topics
Loading…