With a one-line change, you can switch between open-weight OCR VLMs (DeepSeek OCR 2, GLM-OCR, dots.mocr, Paddle OCR VL, PP-OCRv6, etc.) and process 100K+ pages for under $60 on VLM Run Gateway
> one OpenAI-compatible endpoint for open-weight OCR and VLM models
Neat! Looks like an updated version of OCR Arena[0], which doesn't seem to have kept up with new models for a bit.
Would be nice to see the inclusion of Mistral OCR 4, Baidu's Unlimited-OCR, and OlmOCR2.
[0]OCR Arena: https://www.ocrarena.ai/
thank you for the model suggestions!
With a one-line change, you can switch between open-weight OCR VLMs (DeepSeek OCR 2, GLM-OCR, dots.mocr, Paddle OCR VL, PP-OCRv6, etc.) and process 100K+ pages for under $60 on VLM Run Gateway
> one OpenAI-compatible endpoint for open-weight OCR and VLM models
Try it out quickly via OpenAI SDK:
``` client = OpenAI( base_url="https://gateway.vlm.run/v1/openai", api_key="<VLMRUN_API_KEY>", )
response = client.chat.completions.create( model="rednote-hilab/dots.mocr", messages=[{ "role": "user", "content": [{ "type": "document_url", "document_url": {"url": "https://.../invoice.pdf"}, }], }], extra_body={"document_dpi": 72}, ) ```
or via our CLI:
``` pip install vlmrun vlmrun gw models vlmrun config set --api-key 'vlmrun' # anon-user, rate-limited vlmrun gw chat <doc>.pdf -m zai-org/glm-ocr vlmrun gw chat <doc>.pdf -m zai-org/glm-ocr --json-mode vlmrun gw chat <doc>.pdf -m deepseek-ai/deepseek-ocr-2 vlmrun gw chat <doc>.pdf -m rednote-hilab/dots.mocr vlmrun gw chat <doc>.pdf -m paddleocr/pp-ocrv6 ```
Docs: https://docs.vlm.run/gateway
Catalog: https://docs.vlm.run/gateway/models
MCP: https://docs.vlm.run/gateway/mcp-server
Colab Quickstart: https://colab.research.google.com/drive/1RkuVIyuc5Po-UlcSlFy...
Read the full post here: https://huggingface.co/blog/vlm-run/intro-to-vlmrun-gateway
What free and open models are people using for segmentation tasks these days?