Jerry Liu 发布 OpenDocRouter,一个面向文档解析的统一 API,聚合前沿与开源权重模型并转写为 markdown。上线时提供 10 个模型,包括 Claude Opus 5.5、Gemini 3.8 Flash、GPT-6 Luna、MinerU2.5-Pro 和 PaddleOCR-VL-1.6;一行代码即可切换模型,无需新提示词或集成。
Today I’m excited to introduce OpenDocRouter - a unified API for document parsing with the latest frontier and open-weight models.
There are a lot of VLMs and OCR models that can be used for document parsing: we have over 130+ models on ParseBench, and a HuggingFace search for “ocr” turns up thousands of results.
It’s extremely time consuming to choose between OCR vendors. You need to figure out the right prompts, handle rate limits, manage deployments, integrations with all models you’re using, and benchmark new models as they come out.
OpenDocRouter provides a comprehensive, transparent set of models along the price-performance frontier. It manages a unified API to transcribe documents to markdown. It serves all frontier and open-weight models “at-cost”, with a small transaction cut. It handles rate limits with all models to ensure you can put massive volume through. It even offers bounding boxes and layout as a service, so that you can add grounding to any model that you’re using.
When new OCR candidate models ship, we will benchmark them on ParseBench and immediately add them to OpenDocRouter.
We’re adding a lot more models very quickly, and also adding some extremely exciting feature improvements (e.g. automated routing) as we speak.
We welcome your feedback!
Check it out: https://www.opendocrouter.ai/
Blog: https://www.llamaindex.ai/blog/introducing-opendocrouter
Today we're announcing OpenDocRouter: every model for document parsing under one API. There are ~4,000 OCR models on Hugging Face, and the frontier labs ship a new one nearly every month. Whatever's best for your docs today won't be by Q1. So stop asking which model to use for doc parsing. Ask how fast you can switch. ✅ Switch models in one line: same request, same markdown output, no new prompts or integrations ✅ Any model you want: 10 frontier and open-source models at launch, including Claude Opus 5.5, Gemini 3.8 Flash, GPT-6 Luna, MinerU2.5-Pro and PaddleOCR-VL-1.6 ✅ Choose with receipts: every model scored on ParseBench for quality and cost ✅ Traceable output from any model: "layout: true" adds grounded bounding boxes and the same layout classes, even for models that don't support it natively ✅ Pay only for what works: per-token pricing, failed pages never charged, top up from $25 The spread is the point: $0.86 to $48.82 per 1,000 pages, depending on what your documents actually need. Live now → http://opendocrouter.ai在 X 查看被引用的帖子
来源:Jerry Liu · x.com