Endpoints: 28,729MCP servers: 18,413Payout addresses: 2,070Paid calls: 1,526Letters: 13Defects: 1,322counted just now
teppi

MCP servercom.stagenth/doc-parse

Parse PDF/Word/PPT/HTML to Markdown; tables as JSON, image extraction, RAG chunking, page ranges.
UNRATEDActivestreamable-httpstagenth.com

Overview

Score?
UNRATED 0.284
of what a free look can see, on 25 looks
Looks
26
last 2 days ago
Tools
4

More info

URL
stagenth.com/mcp/doc-parse/
streamable-http
Says it is
stagenth-doc-parse 1.28.1
protocol 2025-06-18
In the record since
32 days ago

Among servers18,413 with a card

0median 0.606 · this server 0.284 · highest on record 0.8561

Toolsfrom sha256:29d76b3afa…837df4

The tools this server lists, read out of the definition it returned
ToolSchema
doc_chunk
把文档切成适合 RAG / 向量嵌入的语义块(按标题层级切,超长块按段落细分)。 返回 [{index, heading, text, chars}],喂检索/嵌入无需再自己写切块逻辑。
input · no output
doc_images
抽取文档内嵌的图片(PDF / .docx / .pptx),打包 ZIP 落文件中转站并返下载 URL。 过滤 <1KB 的图标噪声;多模态 AI 可配合 doc_parse 的文本一起理解图文文档。
input · no output
doc_parse
把 PDF / Word(.docx) / PPT(.pptx) / HTML / 纯文本 解析成干净 Markdown + 结构化(按实际解析页数计费,每页 1 credit)。 纯解析 + 格式转换、不生成任何 AI 内容;扫描件 OCR 暂未支持。 支持 page_range 只解析指定页码段(只按解析的页数扣费)。
input · no output
doc_tables
抽取文档中的表格(PDF / .docx / .pptx),转成 headers+rows 的 JSON 可直接计算。 与 doc_parse 的区别:表格不再压平成 markdown 竖线文本,而是结构化行列, AI 可直接算合计/对比;PDF 用版面识别找表格。
input · no output
Verify it yourselfnpx teppi-check https://stagenth.com/mcp/doc-parse/curl -s https://api.teppi.xyz/v1/trust/mcp/mcs_01M1FZ294CB5T518NFEPHZBX00