pdf-structured-extractor
从 PDF 文档中结构化提取文字、表格与图片,保留标题层级和阅读顺序,支持双栏;扫描页自动渲染为图片交由视觉能力识别;适用于“提取 PDF 文字”“PDF 转 Markdown”“导出 PDF 表格为 CSV/Markdown”“导出 PDF 图片”“识别扫描 PDF/扫描件”等场景。
Install this skill
or
pdf-structured-extractor3 files
Comments
Sign in to leave a comment.
No comments yet. Be the first to comment!