OpenAI- and Anthropic-compatible. GLM and DeepSeek hosted in the U.S.




















Point the SDK you already use at Corvex Token Factory.
Requests land on U.S. GPUs.
OpenAI- and Anthropic-compatible chat completions.
Stack is tuned for time-to-first-token and sustained tokens per second under concurrency.
Prompts and completions are processed in memory. Not logged, not stored, not used for training. Usage metadata for billing only. Details in the docs.
A simple migration gets you going.
Corvex takes data security seriously. With Corvex Token Factory, that means customer data is never logged, stored, or used for training. Workloads run on GPUs in the U.S., SOC 2 Type II and HIPAA certified compliant.
Quick answers on data handling, compliance, migration and support. Can't find what you need? Contact us.
Zero data retention: your prompts and outputs are never stored, and never used to train models. Requests aren't forwarded to third-party providers or to the model developers. SOC 2 Type II examined, HIPAA compliant.
Yes. Prompt and response content is not stored for any period of time. (Prompts must exist in memory momentarily for the model to process them.) We retain only operational metadata — timestamps, model name, token counts, request status, latency, and workspace/key IDs — never content.
Today, appropriately authorized operators may have technical access to production systems. That access is restricted and governed by audited controls, but standard endpoints don't yet make it cryptographically impossible. Making operator access impossible — not just forbidden — is exactly what our confidential-computing work delivers.
No. Corvex operates the open-weight models directly. Requests are never forwarded to the developers, and we never train on your prompts or responses.
On Corvex-owned and -managed hardware in the United States. Nothing is routed to third-party inference providers.
SOC 2 Type I and Type II complete; HIPAA compliant and Business Associate Agreements are available. Full controls in the Trust Center.
Sign up, generate a key, and point your OpenAI- or Anthropic-compatible client at our endpoint. Most developers make their first call within minutes. No credit card to start.
Yes. Update the base URL, API key, and model name. Same API structure as OpenAI's chat completions and Anthropic's endpoints — in most cases you change one line.
No. The open-weight models we serve are available from multiple providers or can be self-hosted. To switch, update the base URL, key, and model name. Test before moving production traffic — serving behavior and performance vary by provider.
A focused, optimized catalog — currently GLM 5.3 and DeepSeek V4 Flash 0731. Request models in the console or Discord.
We run open-weight models on our own infrastructure and price close to the cost of serving, plus a sustainable margin — instead of charging a premium for exclusive access to a closed model.
Connect with Corvex support engineers via our Discord channel or from the support tab in Corvex Token Factory.
Yes — SOC 2 Type II, HIPAA dedicated Slack, and named support engineers. Contact the team.
Get your API key and start building with assurance from U.S.-based hardware and zero data retention.