Rust

pdf-extract

A rust library for extracting content from pdfs

J

jrmuizel

Dernière activité 16 sept. 2026
jrmuizel/pdf-extract

598

étoiles

126

forks

73

issues ouvertes

Ce README est souvent en anglais.

pdf-extract

Build Status crates.io Documentation

A rust library to extract content from PDF files.

let bytes = std::fs::read("tests/docs/simple.pdf").unwrap();
let out = pdf_extract::extract_text_from_mem(&bytes).unwrap();
assert!(out.contains("This is a small demonstration"));

See also

Not PDF specific

Projets similaires

Rust library to read, manipulate and write PDF files.

Rustpdfpdf-filesrust
Ppdf-rs
1,7 k étoiles151

Fast Rust library for PDF inspection, classification, and text extraction. Intelligently detects scanned vs text-based PDFs to enable smart routing decisions.

Rustmarkdownnodejsocr-routing
Ffirecrawl
19,4 k étoiles1,3 k

A fast, helpful, and open-source document parser

Rustdocument-ocrdocument-processingocr
Rrun-llama
12,7 k étoiles865