Home/Products/PDF Engine
Productiondocument ai

PDF Engine — High-accuracy OCR & Data Extraction.

Intelligent document parsing.

High-performance, stateless document generation, OCR extraction, and multi-file merging.

Engine Specifications & Telemetry

Real-time architecture and telemetry running against Rimnex engine benchmarks.

PDF Engine Neural Document Pipeline
> Document Parser: Ingestion stream active (PDF 2.0 / Native & Scanned)
> OCR Worker Pool: PaddleOCR + Tesseract v5 (Dual pass, 8 vCPUs)
> Table Structure Recovery: 99.4% confidence (Hierarchical JSON & Markdown)
> Vector Indexing: Chunking & embeddings generated (1536d pgvector)
> Pipeline Latency: 42 pages parsed, structured & cached in 380ms

Key Capabilities & Specifications

Production-grade features designed for high-concurrency developer workflows.

99.8% precision across invoices, contracts and tables
Multi-column OCR with PaddleOCR & Tesseract
Direct structured JSON output via high-throughput API
Stateless microservice compute with zero database dependencies
Architecture Blueprint

Decoupled Compute Architecture

PDF Engine operates on a standalone compute cluster with zero local databases. Authentication, OAuth2 client credential hashing, and metered quota billing are handled centrally by Rimnex.

Stateless Compute

Incoming jobs process in memory and stream back results. Zero database lag.

Unified SSO

Your Rimnex session cookie automatically authorizes your session on pdf-engine.rimnex.com.

Central Billing

Manage all subscription tiers, invoices, and Lemon Squeezy receipts in Rimnex Console.

Pricing & 7-Day Free Trial

Standardized plans (Starter, Growth, Pro, Pay-As-You-Go, Custom) and 7-day trials are hosted directly on PDF Engine's dedicated website.

View Plans on PDF Engine ↗