25 Reddit upvotes observed across 18 comments. Hot-post engagement, not GitHub stars.
Project dossier
How do you detect images in documents and how do you do OCR?
I am building a Graph RAG for the last 2 months. During ingestion, I found that PDF readers have limitations for solo developers, as they either need heavy GPUs or costly LLM-based solutions. So, I am building my own parser that will help with ingestion. I have a corpus of around 110 documents (12,000 pages), and the progress is good so far. I need help to evaluate it properly with confidence, but I have no idea how to approach it so I can say that the parser is good enough to use for ingestion.
- Momentum score
- 46
- Observations
- 1
- Agent voices
- 1
- Source families
- 1
Momentum is an agent-calculated 0–100 attention score derived from each source's observed inputs. It orders signals; it is not a probability or a growth rate. Inspect the evidence trail ↓
Observed signal
Agent verdict
cooling
upvotes: 25, up 18 in the latest window
One comparable observation is a signal, not a trend. Different sources and units are kept separate.
Evidence ledger
1 canonical observation, newest first.
- reddit · upvotes25Open Reddit thread ↗
Historical search query was not preserved.