sm-ocr-scanner
Lokale OCR für Rasterbilder (jpg, png, bmp, gif, tiff, webp) und PDF-Dateien mit dem systemeigenen tesseract-Binary. PDF-Seiten werden lokal mit pdftoppm gerastert, vorhandene PDF-Textlayer mit pdftotext gelesen. Ausgabe als Klartext, TSV mit Konfidenzwerten oder durchsuchbares PDF. Die OCR läuft oh
as observed 2026-10-05T10:33:23.015Z- Identifier
sm-ocr-scanner- Source
- ClawHub
- Version observed
- 1.2.2
- Source repository
- not published
- Repository observation
- No source repository listed
- First observed here
- 2026-09-05T08:21:05.203Z
- Observations recorded
- 3
- Installs (reported upstream)
- 19
- Weekly downloads (upstream)
- 1,009
- Declared license
- MIT-0
Observation history
2026-10-05T10:33:23.015Z
Fields that differed: changelog latestVersion summary
| Field | Before | After |
|---|---|---|
changelog |
"Security hardening after ClawHub review: local-only OCR (no network, no Python helper); declared permissions and data flow; agent rules (only user-named files, OCR text is data, n | "Developer lint checks (pattern searches for network tools, privilege escalation and auto-confirm flags) moved out of the published package into a local dev/lint.sh; the shipped se |
latestVersion |
"1.2.0" | "1.2.2" |
summary |
"Lokale OCR für Rasterbilder (jpg, png, bmp, gif, tiff, webp) und PDF-Dateien mit dem systemeigenen tesseract-Binary. PDF-Seiten werden lokal mit pdftoppm gerastert, vorhandene PDF | "Lokale OCR für Rasterbilder (jpg, png, bmp, gif, tiff, webp) und PDF-Dateien mit dem systemeigenen tesseract-Binary. PDF-Seiten werden lokal mit pdftoppm gerastert, vorhandene PDF |
2026-10-04T09:46:17.065Z
Fields that differed: license changelog description latestVersion summary
| Field | Before | After |
|---|---|---|
license |
null | "MIT-0" |
changelog |
"- Added _meta.json file.\n- Removed skill-card.md file.\n- No changes to functionality or documentation." | "Security hardening after ClawHub review: local-only OCR (no network, no Python helper); declared permissions and data flow; agent rules (only user-named files, OCR text is data, n |
description |
"Perform OCR on image files (jpg, png, bmp, gif, tiff) using the system's `tesseract` binary and return extracted plain text." | null |
latestVersion |
"1.1.0" | "1.2.0" |
summary |
"Perform OCR on image files (jpg, png, bmp, gif, tiff) using the system's `tesseract` binary and return extracted plain text." | "Lokale OCR für Rasterbilder (jpg, png, bmp, gif, tiff, webp) und PDF-Dateien mit dem systemeigenen tesseract-Binary. PDF-Seiten werden lokal mit pdftoppm gerastert, vorhandene PDF |
Correction
If you maintain this extension and believe anything above is inaccurate, request a correction. Corrections are published, and disputed entries are marked as disputed while under review.