Data stays in your network
Prompts, documents and answers are processed on servers you own, inside your own data centre.
Cortex On-Premise
Run TextCortex on servers in your own data centre. Set your headcount, usage and open-weight model, and see the GPU servers, load and electricity cost it takes.
Estimates based on measured TextCortex enterprise usage. Ctrl/⌘ + scroll to zoom, drag to move.
Why on-premise
Prompts, documents and answers are processed on servers you own, inside your own data centre.
Compare DeepSeek, GLM, MiMo, Kimi, Gemma, Qwen and Nemotron on quality, memory and speed before you buy hardware.
See the server tier, the number of nodes and the electricity cost for your country before you ask for a quotation.
How the sizing works
Usage profiles come from real enterprise use of TextCortex, from occasional chat to agent-first power users working on long documents.
Servers are sized for five times the busiest-hour average, and never below the burst of a single power user.
Per-server throughput is either measured in public benchmarks or marked as an estimate. Final sizing is confirmed with a load test.
FAQ
It gives a sizing estimate. Demand comes from measured TextCortex enterprise usage. Server capacity is measured in public benchmarks where marked and estimated otherwise. Request a quotation to confirm the sizing for your organisation.
DeepSeek V4 Flash, DeepSeek V4.1 Flash, GLM-5.3, GLM-5.3-Flash, MiMo-V2.6-Pro, Kimi K3, Gemma 4 31B, Qwen3.8 27B and Nemotron 3 Super. Each model needs a minimum amount of GPU memory, so not every model fits every server tier.
Servers are assumed to draw 60% of their maximum power around the clock (30% for a standby node), plus 40% for cooling. The price is the average business electricity price for the selected country, excluding VAT.
Yes. Turn on the standby node in the server panel. It adds one extra node, and the estimate includes its idle power.
Select “Get a price quotation” in the configurator. You can email the configuration to our team or book a call.
Talk to our team about hardware, models and the rollout in your organisation.