Wholesale inference for open-weight LLMs.
RouteCortex operates dedicated GPU capacity for open-weight roleplay and creative-writing models. OpenAI-compatible endpoints, aggressive throughput, no prompt or completion logging.
Not a consumer product. We sell inference to platforms and developer tooling — aggregators, chat frontends, and character AI backends looking for models they can’t source elsewhere.
api.routecortex.io/v1
OpenAI-compatible · bearer-token auth · streaming ·
Prometheus metrics at /metrics
- Interface
- OpenAI-compatible (chat completions, models, metrics)
- Available now
TheDrummer/Cydonia-24B-v4.3· bfloat16 · 8k context- Hardware
- NVIDIA A100 80GB class · region on request
- Payout
- USDC on Base, Monero, or Nano
- Onboarding
- [email protected]