Categories: None [Edit]
pikuri-pdf
pikuri-pdf plugs PDF → text extraction into pikuri-core's
+Pikuri::Extractor+ registry. The bundled +Pikuri::Extractors::PDF+
extractor wraps the pure-Ruby pdf-reader gem and extracts lazily:
paged reads (the +read+ tool's windows) parse only the pages the
window needs, so the first page of a 500-page PDF never pays for
the other 499.
Shipped separately from pikuri-core so the core's dependency tree
stays minimal and auditable: pdf-reader and its transitive deps
(Ascii85, afm, hashery, ruby-rc4, ttfunk) ride along only for hosts
that opt into PDF support.
Registration is explicit — +Pikuri::Extractors::PDF.register+ — so
requiring the gem changes nothing by itself; the host script picks
which extractors it wires in. One registration extends the +read+
tool, +web_scrape+, and the pikuri-vectordb indexer simultaneously.
Total
Ranking: 191,785 of 196,250
Downloads: 603
Daily
Ranking: 29,007 of 196,218
Downloads: 3
Downloads Trends
Ranking Trends
Num of Versions Trends
Popular Versions (Major)
Popular Versions (Major.Minor)
Depended by
| Rank | Downloads | Name |
|---|---|---|
| 185,569 | 1,393 | pikuri |
Depends on
| Rank | Downloads | Name |
|---|---|---|
| 478 | 112,426,929 | pdf-reader |
| 186,607 | 1,229 | pikuri-core |
Owners
| # | Gravatar | Handle |
|---|---|---|
| 1 | mavi |