pp-structurev3

by Baidu

PP-StructureV3 is an efficient and comprehensive document parsing solution that can effectively convert document images and PDF files into structured content (such as Markdown format). It features powerful capabilities including layout area detection, table recognition, formula recognition, chart understanding, and multi-column reading order recovery. This tool performs excellently across various document types and can handle complex document data.

API Pricing

Input$2 / 1M tokens

FAQ

What is pp-structurev3?

PP-StructureV3 is an efficient and comprehensive document parsing solution that can effectively convert document images and PDF files into structured content (such as Markdown format). It features powerful capabilities including layout area detection, table recognition, formula recognition, chart understanding, and multi-column reading order recovery. This tool performs excellently across various document types and can handle complex document data.

How much does pp-structurev3 cost?

On AIHubMix, pp-structurev3 costs $2 per million input tokens.

How do I call pp-structurev3 via API?

pp-structurev3 is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to pp-structurev3 — no other code changes needed.

Who develops pp-structurev3?

pp-structurev3 is developed by Baidu. AIHubMix aggregates it alongside models from other providers behind one API and one bill.

More from Baidu

ernie-5.1

by Baidu

ERNIE 5.1 is the latest model in the Wenxin series, with comprehensive upgrades to its…

$0.56/1M in · $2.54/1M out
119,000 tokens context

ernie-5.0

by Baidu

ERNIE 5.0 is the next-generation natively multimodal foundation model in the ERNIE…

$0.82/1M in · $3.29/1M out
119,000 tokens context

ernie-image-turbo

by Baidu

The Ernie-image-Turbo model is an 8-step distilled version of the Ernie-image model, also…

$2/1M in

musesteamer-air-image

by Baidu

musesteamer-air-image is a text-to-image model developed by the Baidu Search team aimed…

$2/1M in

qianfan-ocr

by Baidu

Qianfan-OCR-Fast is a multimodal large model specialized for OCR, trained primarily on…

$0.06/1M in · $0.25/1M out
32,000 tokens context

qianfan-ocr-fast

by Baidu

Qianfan-OCR-Fast is a multimodal large model specialized for OCR, trained primarily on…

$0.66/1M in · $2.74/1M out
32,000 tokens context

Use pp-structurev3 via the AIHubMix unified API — one interface for every major LLM.