Corvex Launches Token Factory for Open Weight AI Inference
Corvex, Inc. launches Corvex Token Factory, a serverless inference service for open weight AI models, as it announced in a press release. The service initially supports GLM 5.3 from Z.ai and DeepSeek V4 Flash 0731 through an API.
Corvex says the service processes prompts and responses in memory without logging, storing, or using them for model training. Requests run on hardware managed by Corvex and are not sent to model developers or third party inference providers. The company retains operational metadata for security, billing, and service operations, but says this excludes prompt and response content.
The service offers OpenAI and Anthropic compatible APIs, allowing developers to connect existing clients by changing the base URL, API key, and model name. Customers pay for input and output tokens without charges for idle GPU capacity. Registration is available at no cost following a closed alpha period.
We hope you enjoyed this article
Consider subscribing to one of our newsletters like Silicon Brief or Daily AI Brief.
Also, consider following us on social media:
More from Data Centers
Sep 30 SWI Group Reports €631.6 Million Profit as AI Infrastructure Portfolio Grows Sep 30 Kao Data Names Bruce Claassen Chief Financial Officer Sep 30 Virtus Solis Signs 20 Year Space Solar Deal With Brae Systems Sep 30 Transmart Expands Magnetic Core Sales Push to North American AI Data Centers Sep 30 NGEN Launches Data Center Power and Cooling Supply ServiceSilicon Brief
Weekly coverage of AI hardware developments including chips, GPUs, cloud platforms, and data center technology.
Whitepaper
Tensordyne Napier: What If One Rack Could Do the Work of Nine?
Tensordyne
This Tensordyne whitepaper presents Napier, an inference-focused AI processor and rack-scale system based on the company’s TDN Math logarithmic number system. It examines infrastructure requirements for large mixture-of-experts and agentic models, compares major inference architecture approaches, and details the TDN AIP processor, TDN72 pod, TDN Link fabric, and Napier Ultra configuration. The paper reports simulation-based performance, cost, and accuracy-validation results, including Tensordyne’s projected comparison of one Napier rack with a nine-rack Nvidia Rubin plus Groq deployment; the chip is reported as taped out and in fabrication.
Read moreYou may also like
Crusoe to Run Perplexity Model Training and Inference
Gensyn Releases Auditable open-1b AI Model
Accels Launches AI Model Routing and Token Finance Platform
Abacus.AI Releases Three Smaug Open Weight Models
Liner Offers Model Routing API With OpenAI Compatibility
Daily AI Brief: the AI news that matters, in your inbox.