MCP servercom.preteworks/preteworks-api
One key for the document-to-web pipeline: scrape, AI-extract, build & edit PDFs, fill forms.
Overview
Score?
UNRATED 0.669
of what a free look can see, on 28 looks
Looks
29
last 6 hr ago
Tools
29
changed 25 days ago
More info
URL
api.preteworks.com/mcp
streamable-http
Says it is
preteworks 1.1.0
protocol 2025-06-18
In the record since
27 days ago
Among servers18,413 with a card
0median 0.606 · this server 0.669 · highest on record 0.8561
Toolsfrom sha256:52e767f2f1…9b9952 · +3 −0 25 days ago
| Tool | Schema |
|---|---|
| answer_from_document added Answer a question grounded ONLY in a source document — a web page, PDF, Office file, or raw text. Provide 'question' plus ONE source: 'url', 'pdf', 'file' (base64), or 'text'. Says |
input · no output |
| batch_scrape Fetch up to 10 public URLs and return each as clean Markdown, in one call — for research/RAG over several pages at once. Private/internal hosts are blocked. |
input · no output |
| crawl_site Fetch a URL plus up to 7 more same-origin pages it links to (≤8 total), each as clean Markdown. Bounded and synchronous. Private/internal hosts are blocked. |
input · no output |
| data_to_spreadsheet added Turn rows of data into a downloadable XLSX (default) or CSV. 'rows' is an array of objects (keys → header row) or an array of arrays (first row is the header). Returns a file_id (f |
input · no output |
| docx_to_pdf Convert a .docx document (base64) to a PDF. Returns a file_id and a ~1h URL. |
input · no output |
| docx_to_text Extract the text of a .docx document (base64). Returns the text inline. |
input · no output |
| extract_data Extract specified fields from a web page, PDF, or Office file as JSON, using AI. Provide 'fields' plus ONE source: a 'url', a 'pdf' (file_id or base64), or a 'file' (base64 .docx/. |
input · no output |
| extract_pdf_text Extract the text content of a PDF — for RAG, summarization, or search. Accepts a file_id (from a prior tool) or a base64-encoded PDF, and returns the text inline. Not OCR: a scanne |
input · no output |
| file_to_markdown Convert a Word (.docx), Excel (.xlsx) or CSV file (base64) into clean Markdown — for RAG, agents and pipelines. DOCX keeps headings/lists/tables; spreadsheets become Markdown table |
input · no output |
| fill_pdf_form Fill a PDF's form fields from a map of field -> value; optionally flatten. Returns a file_id and a ~1h URL. |
input · no output |
| images_to_pdf Combine PNG/JPEG images (base64) into a PDF, one image per page. Returns a file_id and a ~1h URL. |
input · no output |
| map_site Discover the same-origin URLs linked from a page — the site map, with no page content fetched. Private/internal hosts are blocked. |
input · no output |
| markdown_to_pdf Render Markdown to a clean, print-styled PDF. Returns a file_id and a ~1h URL. |
input · no output |
| merge_pdfs Combine 2+ PDFs into a single PDF, in the order given. Inputs are file_ids (from prior tools) or base64 PDFs. Returns a file_id (for chaining) and a ~1h download URL. |
input · no output |
| number_pdf Stamp a page number on every page. position: bottom-center (default) | bottom-left | bottom-right | top-center | top-left | top-right. format supports {n} and {total}. Returns a fi |
input · no output |
| pdf_info Read a PDF's metadata (title, author, subject, keywords, creator, producer, dates) plus page count and page size, as JSON. |
input · no output |
| read_pdf_form List a PDF's form fields (name, type, value, options) as JSON. |
input · no output |
| read_url Fetch a public http/https URL and return its main content as clean Markdown — ideal for giving an agent readable web content for research or RAG. Private/internal hosts are blocked |
input · no output |
| render_document Render a structured document to a PDF from a named template (invoice, receipt, report). Money/totals are computed for you. Returns a file_id (for chaining) and a ~1h download URL. |
input · no output |
| render_html_to_pdf Render a full HTML document to a PDF. External subresources are blocked for safety — inline images/fonts as data: URIs. Returns a file_id (for chaining) and a ~1h download URL. |
input · no output |
| rotate_pdf Rotate every page of a PDF by 90, 180, or 270 degrees (clockwise). Returns a file_id (for chaining) and a ~1h download URL. |
input · no output |
| scrape_page Fetch a public http/https URL and return its rendered HTML, the links on the page, and page metadata (title/description/OpenGraph/canonical/favicon) — in one render. For clean Mark |
input · no output |
| select_pages Keep only the specified pages of a PDF (e.g. [1,3,5]) and drop the rest, preserving order. Returns a file_id (for chaining) and a ~1h download URL. |
input · no output |
| set_pdf_metadata Set a PDF's title, author, subject, keywords (array or comma string) and/or creator. Only the fields you provide change. Returns a file_id + ~1h URL. |
input · no output |
| split_pdf Split a PDF into multiple PDFs — one per page by default, or by page ranges like '1-3;4-6'. Returns a file_id + URL per part. |
input · no output |
| summarize_document added Summarize a web page, PDF, Office file (.docx/.xlsx/.csv), or raw text using AI. Provide ONE source: 'url', 'pdf' (file_id or base64), 'file' (base64), or 'text'. Optional 'max_wor |
input · no output |
| url_to_pdf Fetch a public http/https URL and render the live page to a PDF. Returns a file_id (chainable into the PDF tools) and a ~1h download URL. Private/internal hosts are blocked. |
input · no output |
| url_to_screenshot Fetch a public http/https URL and capture a PNG screenshot. Set full_page for the entire scroll height. Returns a file_id and a ~1h download URL. |
input · no output |
| watermark_pdf Stamp a diagonal grey text watermark (e.g. "DRAFT", "CONFIDENTIAL") across every page of a PDF. Returns a file_id (for chaining) and a ~1h download URL. |
input · no output |
Verify it yourself
npx teppi-check https://api.preteworks.com/mcpcurl -s https://api.teppi.xyz/v1/trust/mcp/mcs_01M1X6NRH28MAWZJ7M8MZNH56X