Marketplace
By Use Case Marketplace
Compare By Use Case products across curated subcategories and trusted providers
- Products
- 15,460
- Subcategories
- 699
Selected subcategory
AI Document Extraction
Compare By Use Case products across curated subcategories and trusted providers
Explore By Use Case with structured category paths, practical filters, and independent product data
Showing 1–12 of 17
by Docling Project
SDK and CLI for parsing PDF, DOCX, HTML, and more, to a unified document representation for powering downstream workflows such as gen AI applications.
by Yang
A Chinese information extraction tool.
by opendataloader-project
PDF Parser for AI-ready data. Automate PDF accessibility. Open-source.
by NVIDIA
NeMo-Retriever is a Python library for scalable document content and metadata extraction.
by Microsoft
Large-scale Self-supervised Pre-training Across Tasks, Languages, and Modalities
by Fastino Labs
GLiNER2: Unified Schema-Based Information Extraction and Text Classification
by Nanonets
Extract and Convert PDF, Word, PowerPoint, Excel, images, URLs into multiple formats (Markdown, JSON, CSV, HTML) with intelligent content extraction and advanced OCR.
by magicrew
Turn documents into AI-ready Markdown with visual understanding
Entity and Relation Extraction Based on TensorFlow and BERT. 基于TensorFlow和BERT的管道式实体及关系抽取,2019语言与智能技术竞赛信息抽取任务解决方案。Schema based Knowledge Extraction, SKE 2019
by topduke
OpenOCR: An Open-Source Toolkit for General-OCR Research and Applications, integrates a unified training and evaluation benchmark, commercial-grade OCR and Document Parsing systems, and faithful reproductions of the core implementations from a wide range of academic papers.
by LandingAI
The official Python library for the ade API
by gal kahana
Node.js module for high performance creation, modification and parsing of PDF files and streams