Beta Accessagentic sandbox

Omnix — Think Faster. Talk Deeper.

Multi-model AI, unified.

Access leading AI models, work with files and images, and have fast, context-aware conversations in one place.

Engine Specifications & Telemetry

Real-time architecture and telemetry running against Rimnex engine benchmarks.

Omnix Multi-Model & Client WebGPU Kernel
> Runtime Environment: Client WebGPU (Q4_K_M Tensor Quantization)
> Air-Gapped Privacy Mode: Zero cloud data transmission verified
> Unified Hub: Hybrid routing across local WebGPU, Claude 3.5 & GPT-4o
> Memory Buffer: 128k context window cached in local tensor memory
> In-Browser Execution Speed: 58.4 tokens/sec on modern GPU

Key Capabilities & Specifications

Production-grade features designed for high-concurrency developer workflows.

100% on-device inference via WebGPU option
Unified access to Claude, GPT-4o, and Gemini models
Context-aware file and image comprehension
Zero cloud data transmission privacy mode
Architecture Blueprint

Decoupled Compute Architecture

Omnix operates on a standalone compute cluster with zero local databases. Authentication, OAuth2 client credential hashing, and metered quota billing are handled centrally by Rimnex.

Stateless Compute

Incoming jobs process in memory and stream back results. Zero database lag.

Unified SSO

Your Rimnex session cookie automatically authorizes your session on omnix.rimnex.com.

Central Billing

Manage all subscription tiers, invoices, and Lemon Squeezy receipts in Rimnex Console.

Pricing & 7-Day Free Trial

Standardized plans (Starter, Growth, Pro, Pay-As-You-Go, Custom) and 7-day trials are hosted directly on Omnix's dedicated website.

View Plans on Omnix ↗