PDF 处理
原名:pdf
当用户涉及PDF文件操作时(如读取、合并、拆分、旋转、加水印、创建、填写表单、加密解密、提取图片及OCR识别等),使用此技能。
- 分类
- 开发提效
- 版本
- v1.0.0
- 作者
- 弈韬(@ra1nzzz)
- 下载
- 1
- 收藏
- 0
- 发布
- 2026-08-17
- 更新
- 2026-08-18
- TRACE 评分
- 2.8 / 5
内容概览
This guide covers essential PDF processing operations using Python libraries and command-line tools. For advanced features, JavaScript libraries, and detailed examples, see REFERENCE.md. If you need to fill out a PDF form, read FORMS.md and follow its instructions. IMPORTANT : Never use Unicode subscript/superscript characters (₀₁₂₃₄₅₆₇₈₉, ⁰¹²³⁴⁵⁶⁷⁸⁹) in ReportLab PDFs. The built-in fonts do not include these glyphs, causing them to render as solid black boxes. Instead, use ReportLab's XML markup tags in Paragraph objects: For canvas-drawn text (not Paragraph objects), manually adjust font the size and position rather than using Unicode subscripts/superscripts. Task Best Tool Command/Code ------ ----------- -------------- Merge PDFs pypdf writer.add page(page) Split PDFs pypdf One page per file Extract text pdfplumber page.extract text() Extract tables pdfplumber page.extract tables() Cr…