repoGitHubTrust 82 · PrimaryPublished 2mo agoLive · 9d ago
yfedoseev/pdf_oxide
The fastest PDF library for Python and Rust. Text extraction, image extraction, markdown conversion, PDF creation & editing. 0.8ms mean, 5× faster than industry leaders, 100% pass rate on 3,830 PDFs. MIT/Apache-2.0.
Lineage graph
Paper → model → repo connections mined from source citations (Tier-1 exact match).
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- PossiblePossibly related (embedding) · 51%Build interactive PDF text extraction from Amazon S3 →
- PossiblePossibly related (embedding) · 45%TurboOCR v3 — high-speed document OCR server (C++/CUDA), ~520 img/s on RTX 5090 →
- PossiblePossibly related (embedding) · 51%9 Rust apps that are faster than the Linux tools they replace - MSN →
- PossiblePossibly related (embedding) · 48%A Production RAG Pipeline for PDFs: Relational Parsing, TOC Retrieval, Typed Answers - Towards Data Science →
Covers
Covers (incoming)
Related across the graph
news9 Rust apps that are faster than the Linux tools they replace - MSNnewsTurboOCR v3 — high-speed document OCR server (C++/CUDA), ~520 img/s on RTX 5090newsA Production RAG Pipeline for PDFs: Relational Parsing, TOC Retrieval, Typed Answers - Towards Data SciencenewsBuild interactive PDF text extraction from Amazon S3
