Wholesale inference for open-weight LLMs.

RouteCortex operates dedicated GPU capacity for open-weight roleplay and creative-writing models. OpenAI-compatible endpoints, aggressive throughput, no prompt or completion logging.

Not a consumer product. We sell inference to platforms and developer tooling — aggregators, chat frontends, and character AI backends looking for models they can’t source elsewhere.

Endpoint

api.routecortex.io/v1

OpenAI-compatible · bearer-token auth · streaming · Prometheus metrics at /metrics

Spec
Interface
OpenAI-compatible (chat completions, models, metrics)
Available now
TheDrummer/Cydonia-24B-v4.3 · bfloat16 · 8k context
Hardware
NVIDIA A100 80GB class · region on request
Payout
USDC on Base, Monero, or Nano
Onboarding
[email protected]