Model card
NuExtract3 is a specialized image-to-text model designed for structured information extraction. Unlike general-purpose VLMs that often struggle with precision or hallucinate during data parsing, NuExtract3 focuses on transforming unstructured visual data into machine-readable formats. It is particularly effective for developers building automated pipelines for invoice processing, form digitization, and document analysis where schema adherence is critical. With an Apache-2.0 license, it offers the flexibility for commercial integration without restrictive overhead. Developers can integrate it into existing OCR workflows to replace brittle rule-based parsing with a more robust, neural extraction layer that maintains high fidelity to the source document.
Model files and versions
Download this model
We recommend using the ModelScope CLI or SDK. Install ModelScope first, then choose a full snapshot, single file, SDK or Git LFS workflow.
numind/NuExtract3Install the CLI and SDK dependency before downloading.
pip install modelscopeDownload the complete weights, configuration and model card.
modelscope download --model numind/NuExtract3README.md is used as an example; replace it with another repository file when needed.
modelscope download --model numind/NuExtract3 README.md --local_dir ./dirUseful in Python projects and automation scripts.
from modelscope import snapshot_download
model_dir = snapshot_download('numind/NuExtract3')Make sure Git LFS is installed correctly.
git lfs install
git clone https://www.modelscope.cn/numind/NuExtract3.gitFetch the repository structure first, then pull large files when needed.
GIT_LFS_SKIP_SMUDGE=1 git clone https://www.modelscope.cn/numind/NuExtract3.gitHow to use
- 01Step 1
Read the model card and source information.
- 02Step 2
Start with a small, non-sensitive evaluation.
- 03Step 3
Review quality, licensing and usage limits.
- 04Step 4
Adopt it only after validation.
Discussions
Use this space to keep checking source information, usage experience and maintenance status.
Open source page