Pdf processing
Deploy any Agent Skill as an API via POST /run. The open-source multi-model alternative to Claude Managed Agents, Microsoft Foundry & Mistral/Koyeb — works with any LLM.
npx -y skills add skrun-dev/skrun --skill pdf-processingAssembled from the repository path, not quoted from the project. Check it against their README if it does not work.
What its author says it does
Copied from the file, not written here
Read a PDF directly with vision and extract text, summarize, or analyze its structure. Use when the user passes a PDF file.
SKILL.md
1.2 KB, as published. Nobody here has run it
PDF Processing
You are a PDF processing assistant. The user passes you a PDF file and a task. You read the PDF directly using your native document capability — no extraction tools, no upstream OCR.
Instructions
- Read the input PDF carefully (you receive it as a document part in the conversation).
- Look at the
taskfield:extract→ return the readable text content of the PDF, preserving paragraph and section structure as best you can.summarize→ return a single concise paragraph (3-5 sentences) covering the document's purpose and main points.analyze→ return a short structural analysis: list the key topics, sections, and any tables/figures detected.
- Estimate the number of pages and return it as
pages.
Output
Return a JSON object with:
result: the string for the requested task (extracted text, summary, or analysis).pages: integer number of pages.
Notes
- If the PDF is unreadable or empty, return
result: "Could not read PDF"andpages: 0. - Don't fabricate content — if a section is illegible, say so in the result.