GPU-Native Infrastructure
Two platforms, built GPU-native: a Knowledge Platform that turns your documents into answers people trust, and a Real-Time Platform for state and keys that stay resident. All running where your data lives.
Knowledge Platform
Turn your documents into answers people trust — embed, retrieve, answer, publish.
Forge
GPU-native vector embedding API built on a proprietary CUDA engine. The Turbo tier is FREE FOREVER — forge production-grade vectors at zero cost, rate-limited, no card required. Three quality tiers (Turbo, Pro, Ultra) with 87ms median latency and zero data retention — hosted, or run the same engine yourself on bare metal or Kubernetes.
Voxell Answers
Drop in a document, ask a question, get a cited answer — not chunks, not vectors. The whole retrieval pipeline, managed, on Voxell's MTEB-leading embeddings.
Voxell Spaces
Publish a corpus as a private, hosted site your people can read and ask questions of. Signed links, a per-visitor answer budget, and an access allowlist — no app to build.
Coherence
GPU-native KNN retrieval database. Exact k-NN + BM25 hybrid at scale, with no ANN index to build, tune, or drift.
Learn more →Real-Time Platform
State and keys resident on the GPU — local-first sync and a Redis-fast cache.
Voxell Lux
Self-hosted, low-latency cross-device state synchronization. Local-first reads with real-time WebSocket push, flat predictable cost, deployed on your own cloud account.
ARC
GPU-native key-value + sorted-set store. Redis verbs served by a resident megakernel — 217M ops/s at 8.8µs p50, with the CPU out of the data path.
Forge and Answers are live and self-serve, and Lux is available now via private offer. Coherence and ARC ship through the Design Partner Program — direct founder access and a say in the roadmap.