Quatized nanonets/Nanonets-OCR-s with llama.cpp commit fb85a288

Multiple ways to use:

  • run with llama.cpp: ./llama-server -m "Nanonets-OCR-s-Q4_K_M.gguf" --mmproj "mmproj-Nanonets-OCR-s.gguf"
  • use with lmstudio: just pull from benhaotang/Nanonets-OCR-s-GGUF (Warning‼️: change chat template to chatml in model settings)
  • use with ollama: ollama run benhaotang/Nanonets-OCR-s

Suggested system prompt:

Extract the text from the above document as if you were reading it naturally.
Return the tables in html format. Return the equations in LaTeX representation.
If there is an image in the document and image caption is not present,
add a small description of the image inside the <img></img> tag;
otherwise, add the image caption inside <img></img>.
Watermarks should be wrapped in brackets. Ex: <watermark>OFFICIAL COPY</watermark>.
Page numbers should be wrapped in brackets. Ex: <page_number>14</page_number> or <page_number>9/22</page_number>.
Prefer using ☐ and ☑ for check boxes.
Downloads last month
59
GGUF
Model size
3.09B params
Architecture
qwen2vl
Hardware compatibility
Log In to view the estimation

4-bit

6-bit

8-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for benhaotang/Nanonets-OCR-s-GGUF

Quantized
(8)
this model