Item detail
github.com

opendataloader-project/opendataloader-pdf

opendataloader-project/opendataloader-pdf is a code repository that RepoRadar is tracking in its Document parsing section, currently rated Gold tier with a 'try now' verdict. Its strongest signal is workflow potential, scored 9.6 out of 10.

Score8.5
Popularity100.0
Riskconditional
TierGold
Score breakdown
Usefulness8.8
Novelty7.9
Momentum8.4
Maturity9.1
Open-source/build8.4
Evidence7.2
Workflow potential9.6
Setup ease7.1

Popularity is tracked separately. Support, ads, sponsorships, and tips never affect these signals.

Why it matters

Document extraction is still one of the most common pain points in AI stacks, and a lot of teams are stuck choosing between brittle PDF text dumps and expensive hosted parsers. This matters because it offers a well-documented local core with OCR, layout retention, and citation-friendly structure that is directly useful for ingestion and retrieval workflows.

Who should use it

teams building RAG or document-ingestion pipelines around PDFs developers who need structured extraction with boxes, tables, and OCR instead of plain-text PDF scraping builders evaluating local-first PDF tooling before paying for a hosted parser

Who should skip it

Skip opendataloader-project/opendataloader-pdf unless the captured evidence suggests it solves a problem you are actively working on.

About this signal

opendataloader-project/opendataloader-pdf is tracked by RepoRadar as a code repository in the Document parsing section. First seen 2026-07-19; the source record was last checked on 2026-07-19. The current verdict is 'try now' with a Gold tier and moderate setup difficulty. The standout signals for opendataloader-project/opendataloader-pdf are workflow potential (9.6) and maturity (9.1), while setup ease (7.1) trails — that balance shapes where it fits best. This page summarizes the evidence RepoRadar captured from https://github.com/opendataloader-project/opendataloader-pdf.

How this item is evaluated

The opendataloader-project/opendataloader-pdf record combines a 8.5/10 composite score with separate popularity (100.0), risk (conditional), and setup (moderate) signals. See the scoring methodology for the current weights and evidence definitions.

Putting this into practice? Read How to evaluate an AI tool before you adopt it for the checklist behind this score.

Risk explanation

The parser works on user documents, so teams should review how extracted text, OCR artifacts, and images are stored or logged; Optional hybrid processing changes the privacy model compared with a strictly local parse path.

Evidence links
Closest alternatives / related signals
pdf ocr document-parsing rag accessibility apache-2.0
Verification record

What RepoRadar actually verified

Discovered

Automated discovery and source capture. Last checked 2026-08-13T20:20:15Z.

No editorial or hands-on review is claimed. This record remains at Discovered.

Verification sources

Longitudinal intelligence

How this decision record is moving

Raw history JSON →

20 dated snapshots retained from 2026-07-19 through 2026-08-13; see the snapshot index for explicit coverage gaps. Stars, version, release, pricing, integration, risk, maintenance, verdict, score, and momentum fields remain explicit even when a source has not reported them. Repository momentum is a normalized 0–10 RepoRadar signal; GitHub stars appear only where the popularity monitor retained exact timestamped observations.

RepoRadar score8.5 current · +0.0 net
Repository momentum9.6 current · +1.2 net
GitHub stars (observed)28,377 current · +911 net
GitHub stars28,377 exact observation
Versionv2.5.0
Last release2026-07-14T09:24:00Z
Maintenanceactive
Current riskconditional
Current verdicttry now
Pricing baselineNo structured commercial pricing baseline
Pricing checkedNot applicable or not recorded
Pricing freshnessNo dated commercial pricing review
Integrations baselineNo structured integrations recorded

Recent dated points

DateScoreMomentumStarsRiskVerdictMaintenance
2026-08-138.59.628,377conditionaltry nowactive
2026-08-128.59.628,364conditionaltry nowactive
2026-08-118.59.628,331conditionaltry nowactive
2026-08-108.59.628,318conditionaltry nowactive
2026-08-098.59.628,290conditionaltry nowactive
2026-08-088.59.628,266conditionaltry nowactive
2026-08-078.59.628,123conditionaltry nowactive
2026-08-068.58.428,213conditionaltry nownot recorded
2026-08-058.58.428,181conditionaltry nownot recorded
2026-08-048.59.628,123conditionaltry nowactive
2026-08-038.59.628,123conditionaltry nowactive
2026-08-028.59.628,085conditionaltry nowactive

Why the record changed

stars changed

Stars changed: 28364 → 28377.

stars changed

Stars changed: 28331 → 28364.

stars changed

Stars changed: 28318 → 28331.

stars changed

Stars changed: 28290 → 28318.

stars changed

Stars changed: 28266 → 28290.

stars changed

Stars changed: 28123 → 28266.

stars changed

Stars changed: 28213 → 28123.

stars changed

Reconstructed from adjacent retained daily snapshots; no upstream cause is inferred. Stars changed: 28181 → 28213.

stars changed

Reconstructed from adjacent retained daily snapshots; no upstream cause is inferred. Stars changed: 28123 → 28181.

stars changed

Stars changed: 28085 → 28123.

stars changed

Stars changed: 28056 → 28085.

stars changed

Source-observed stars changed: 28046 → 28056. This reports the retained observation delta and does not infer why the upstream change occurred.