Ocr Document Processor
Extract text and structure from scans, images, and scanned PDFs. Use for OCR, searchable PDFs, table extraction, receipt parsing, and business card parsing.
- Скачивания
- 0
- В избранном
- 0
- Комментарии
- 0
- Просмотры
- 2
Установить скилл
Добавьте инструмент одной командой или скачайте проверенный архив версии.
npx skills add dkyazzentwatwa/chatgpt-skills --skill ocr-document-processor- Версия
- 1.0.0+103b4308f65d
- Автор
- Владимир Ломтев
- Репозиторий
- dkyazzentwatwa/chatgpt-skills
Как установить
- 1Скопируйте команду из блока установки.
- 2Запустите её в терминале из каталога проекта.
Документация
OCR Document Processor
Handle OCR-heavy inputs where text must be recovered from images or scanned pages.
Use This For
- OCR on images and scanned PDFs
- Searchable PDF export
- Structured extraction to text, markdown, JSON, or HTML
- Table extraction from scanned material
- Receipt parsing and business card parsing
Workflow
- Decide whether plain OCR, structured extraction, or document-specific parsing is needed.
- Preprocess noisy inputs before extraction when skew, blur, or shadows are present.
- Use
scripts/ocr_processor.pyfor core OCR tasks. - Use the focused helpers when the input is specialized:
scripts/business_card_scanner.pyscripts/receipt_scanner.py
- Return confidence caveats when the source is low quality, rotated, handwritten, or multilingual.
Guardrails
- Prefer explicit language selection when accuracy matters.
- Do not claim fields are exact when OCR confidence is weak.
- Route non-scanned digital PDFs to
document-converter-suiteinstead of OCR by default.
Требования и возможности
Источник пакета
https://github.com/dkyazzentwatwa/chatgpt-skills/tree/103b4308f65df393e76765cb1ce12395d9cf5106/ocr-document-processor
Файлы версии
| Путь | Размер | SHA256 |
|---|---|---|
| SKILL.md | 1239 | bad4ae6cd32e293d... |
| agents/openai.yaml | 182 | 2c552117fa931fda... |
| scripts/business_card_scanner.py | 3273 | 4e5a960ed30256c3... |
| scripts/ocr_processor.py | 27804 | 3a665a7763ce0bf8... |
| scripts/receipt_scanner.py | 3677 | 4448424823ce19e8... |
Частые вопросы
- Как установить Ocr Document Processor?
- Используйте команду
npx skills add dkyazzentwatwa/chatgpt-skills --skill ocr-document-processorили скачайте ZIP-архив. - Можно ли скачать Ocr Document Processor бесплатно?
- Да, опубликованную версию можно скачать из маркетплейса бесплатно.
Похожие инструменты
Смотреть всеStory Long Write长篇网文写作。从大纲到正文,辅助长篇网络小说的创作,包括世界观、人物、情节线管理。触发方式:/story-long-write、/写长篇、「帮我开书」「写大纲」「日更」「续写」「继续写」「修改第X章」「回炉」「重写第X章」。Parallel Deep ResearchONLY use when user explicitly says 'deep research', 'exhaustive', 'comprehensive report', or 'thorough investigation'. Slower and more expensive than parallel-web-search. For normal research/lookup requests, use parallel-web-search instead. Supports multi-turn: pass --previous-interaction-id from a prior research or enrichment to continue with context.LangfuseInteract with Langfuse and access its documentation: tracing, monitoring, creating datasets, running experiments, and evaluating AI applications. Use when needing to (1) query or modify Langfuse data, (2) look up Langfuse documentation, concepts, integration guides, a feature or SDK usage, or (3) do any AI engineering task (AI observability, prompt engineering/management, evaluation and evaluator management, experimentation, dataset management, evaluation-driven CI/CD, feedback collection). Invoke it for tasks in this scope even when Langfuse is not configured or explicitly mentioned.
Комментарии
Войдите, чтобы оставить комментарий.
Комментариев пока нет.