Back to List
Notice:This resource is provided by a third-party author. Please review the code with AI tools or manually before use to ensure security and compatibility.
Rustyfedoseev/pdf_oxide

pdf_oxide

The fastest PDF library for Python and Rust. Text extraction, image extraction, markdown conversion, PDF creation & editing. 0.8ms mean, 5× faster than industry leaders, 100% pass rate on 3,830 PDFs. MIT/Apache-2.0.

80.6/100
922Forks: 113
View on GitHubHomepage →
Loading report...

Similar Projects

turbovec

78

A vector index built on TurboQuant, written in Rust with Python bindings

Rust14.6K

cocoindex

89

Incremental engine for long horizon agents 🌟 Star if you like it!

Rust11.2K

xberg

90

A polyglot document intelligence framework with a Rust core. Extract text, metadata, images, and structured data from 101 formats (115 file extensions) plus code intelligence for 371 code languages. 15 language bindings — Rust, Python, Ruby, Java, Go, PHP, Elixir, C#, TypeScript — plus CLI, REST API, and MCP server.

Rust8.9K

pyrefly

82

A fast type checker and language server for Python

Rust6.9K
Back to List