本文へ移動
cccskills
無料GitHub で公開

pdf-processing

Extract text and tables from PDF files, fill forms, merge documents. Use when working with PDF files or when the user mentions PDFs, forms, or document extraction.

インストール方法を見る

含まれるファイル(4)

  • SKILL.md1.1 KB
  • FORMS.md882 B
  • scripts/analyze_form.py1.3 KB
  • scripts/extract_text.py1.2 KB

SKILL.md(原文)

インストールする前に、エージェントに与えられる指示の中身を確認できます。

PDF Processing

This skill provides utilities for working with PDF documents.

Quick Start

Use pdfplumber to extract text from PDFs:

import pdfplumber

with pdfplumber.open("document.pdf") as pdf:
    text = pdf.pages[0].extract_text()
    print(text)

Available Operations

  1. Text Extraction: Extract text content from PDF pages
  2. Table Extraction: Extract tabular data from PDFs
  3. Form Filling: Fill PDF forms with provided data
  4. Document Merging: Combine multiple PDFs into one

Advanced Features

Form filling: See FORMS.md for complete guide

Utility scripts:

  • Run scripts/analyze_form.py to extract form fields
  • Run scripts/extract_text.py to extract text from a PDF

Best Practices

  1. Always validate PDF files before processing
  2. Handle password-protected PDFs gracefully
  3. Check for scanned PDFs that may require OCR

レビュー

まだレビューはありません。使ってみた感想をお寄せください。

同じリポジトリのスキル

概要と使いどころ

Use when retrieving from or asking questions against a WeKnora knowledge base via the `weknora` CLI — and especially when unsure whether to use `chat`, `session ask`, or `search chunks` for a given goal.

日本語の概要は準備中です。原文の説明を表示しています。

Tencent/WeKnora3.3万2026年10月11日 更新

Use when driving a WeKnora RAG server through the `weknora` CLI as an agent — authenticating, managing knowledge bases / documents / sessions / agents, running search or chat, or interpreting the CLI's JSON envelopes and exit codes. Read this before any other weknora-* skill.

日本語の概要は準備中です。原文の説明を表示しています。

Tencent/WeKnora3.3万2026年10月11日 更新

Tencent のスキルをすべて見る

このスキルの問題を報告する