pdf 技能编写
- 作者仓库星标 239
- 作者仓库 python-agentframework-demos
PDF to Markdown Conversion
This skill converts PDF files to Markdown format using Microsoft's markitdown package.
When to use
- User asks to convert a PDF to Markdown
- User wants to extract text content from a PDF
- User needs to read or parse a PDF document
- User asks to summarize or analyze a PDF file
How to use
Use uvx to run markitdown directly. Pick the dependency group matching the file type:
| File type | Dependency group |
|---|---|
pdf |
|
| PowerPoint | pptx |
| Word | docx |
| Excel (.xlsx) | xlsx |
| Excel (.xls) | xls |
uvx 'markitdown[pdf]' <path-to-file> -o output.md
Or install all optional dependencies at once:
uvx 'markitdown[all]' <path-to-file> -o output.md
Examples
uvx 'markitdown[pdf]' report.pdf -o report.md
uvx 'markitdown[pptx]' slides.pptx -o slides.md
uvx 'markitdown[docx]' document.docx -o document.md
Output
- If you were asked to save the output to a specific file, save it to the requested file using
-o. - If no output file was specified, use the source filename with a
.mdsuffix.
- 流狐分类
- 数据
- 作者声明 Agent
- 未找到明确声明;不据此推断已兼容或已测试
- 静态检查
- 88 / 100 · 启发式扫描,不代表运行安全
- 作者 / 版本 / 许可
- @Azure-Samples · 未声明 license
- 流狐 Token 估算
- 低消耗
- 流狐接入估算
- 即装即用
- 是否需要外部 API Key
- 未发现要求
- 检测到的系统要求
- 未声明
- 底层运行要求
- 未声明
- 检测到的文件与系统行为
-
- 只读
- 检测到的网络行为
- 仅限本地
- 安装命令数
- 无(仅作为资料)
档案由构建时根据 SKILL.md 与安装命令自动衍生,可能与作者实际意图存在差异。
需要注意: 未限定 allowed-tools,默认拥有全部工具权限。
作者没有在当前 SKILL.md 中定义固定输出样例。 User asks to convert a PDF to Markdown User wants to extract text content from a PDF User needs to read or parse a PDF document
Use uvx to run markitdown directly. Pick the dependency group matching the file type: File type · Dependency group PDF · pdf
Examples
If you were asked to save the output to a specific file, save it to the requested file using -o. If no output file was specified, use the source filename with a .md suffix.
# PDF to Markdown Conversion
This skill converts PDF files to Markdown format using [Microsoft's markitdown](https://github.com/microsoft/markitdown) package.
## When to use
- User asks to convert a PDF to Markdown
- User wants to extract text content from a PDF
- User needs to read or parse a PDF document
- User asks to summarize or analyze a PDF file
## How to use
Use `uvx` to run markitdown directly. Pick the dependency group matching the file type:
| File type | Dependency group |
|-----------|-----------------|
| PDF | `pdf` |
| PowerPoint | `pptx` |
| Word | `docx` |
| Excel (.xlsx) | `xlsx` |
| Excel (.xls) | `xls` |
```bash
uvx 'markitdown[pdf]' <path-to-file> -o output.md
```
Or install all optional dependencies at once:
```bash
uvx 'markitdown[all]' <path-to-file> -o output.md
```
## Examples
```bash
uvx 'markitdown[pdf]' report.pdf -o report.md
uvx 'markitdown[pptx]' slides.pptx -o slides.md
uvx 'markitdown[docx]' document.docx -o document.md
```
## Output
- If you were asked to save the output to a specific file, save it to the requested file using `-o`.
- If no output file was specified, use the source filename with a `.md` suffix. 作者原文负责流程事实;流狐只索引当前章节、要点、文件与命令。
章节 -> When to use → How to use → Examples → Output
要点 -> This skill converts PDF files to Markdown format using [Microsoft's markitdown](https://github.com/microsoft/markitdown) package. · Use uvx to run markitdown directly. · - If you were asked to save the output to a specific file, save it to the requested file using -o.
文件/命令 -> uvx · pdf · pptx · docx · xlsx · xls · .md
内容 SHA-256 -> 11a09c5ac0dd
原文结构
适用与边界
原文中的明确线索
uvx、pdf、pptx、docx、xlsx、xls、.md