Qwen/Qwen3-32B-TEE
32B parameter model optimized for reasoning, coding, and instruction following. 40K context window. Most popular TEE model with 33.2M runs in 7 days.
Pricing: $0.15 per million input tokens, $0.44 per million output tokens
14 TEE models (Trusted Execution Environment) running inside Intel TDX hardware enclaves. Prompts and responses are encrypted in memory using AES-256 with CPU-fused keys, even the infrastructure operator cannot read them. OpenAI-compatible REST API. Drop-in replacement for OpenAI and Anthropic for regulated workloads.
TEE models run inside an Intel TDX Trusted Domain, a hardware-isolated virtual machine verified by CPU microcode. This is the same confidential computing technology used by Microsoft Azure Confidential VMs and Google Cloud Confidential VMs. Every prompt and response is encrypted in memory during processing, making TEE models the correct choice for legal documents, patient records, financial data, and any workload under GDPR Article 28, HIPAA, DORA, or professional secrecy rules.
32B parameter model optimized for reasoning, coding, and instruction following. 40K context window. Most popular TEE model with 33.2M runs in 7 days.
Pricing: $0.15 per million input tokens, $0.44 per million output tokens
397B parameter mixture-of-experts model. 256K context window for long document processing, entire contracts or financial reports in one pass.
Latest DeepSeek model with strong performance on code and reasoning tasks.
Pricing: $0.20 per million input tokens, $0.89 per million output tokens
Reasoning model with chain-of-thought capabilities. Suitable for CFA-grade financial analysis, multi-step legal reasoning, and complex compliance analysis.
Pricing: $0.46 per million input tokens, $1.85 per million output tokens
4.6M runs in 7 days. Strong instruction following.
Moonshot AI model with large context window support.
24B parameter Mistral model. Efficient and capable, ideal for cost-sensitive workloads.
Pricing: $0.06 per million input tokens, $0.18 per million output tokens
ZhipuAI GLM models with bilingual (Chinese/English) support.
OpenAI open-weight models running inside Intel TDX enclaves.
Also available: DeepSeek-V3.1-TEE, DeepSeek-V3.1-Terminus-TEE, DeepSeek-V3-0324-TEE, DeepSeek-TNG-R1T2-Chimera-TEE, GLM-4.7-TEE, deepseek-ai/DeepSeek-V4-Flash-0731-TEE, Qwen3-Coder-Next-TEE.
All TEE models are accessible via an OpenAI-compatible REST API. Change your base URL to https://api.voltagegpu.com/v1 and use your VoltageGPU API key. Compatible with OpenAI SDKs (Python, Node.js, Go), LangChain, LlamaIndex, CrewAI, OpenClaw, and any other OpenAI-compatible client library.
curl https://api.voltagegpu.com/v1/chat/completions \
-H "Authorization: Bearer YOUR_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "Qwen/Qwen3-32B-TEE",
"messages": [{"role": "user", "content": "Analyze this contract..."}]
}'TEE models satisfy technical safeguards for GDPR Article 28, HIPAA, SOC 2 (audit in progress), DORA, NIS2 Directive, and French CNIL guidelines. Law firms, accounting firms, clinics, and fintech companies use VoltageGPU TEE models to analyze sensitive documents without violating professional secrecy or regulatory constraints.
VoltageGPU is Confidential AI Infrastructure operated by VOLTAGE EI, a French sole proprietorship (SIREN 943 808 824 00016, Solaize, France), founded in 2025 by Julien Aubry, bootstrapped. Three products: Confidential GPU Compute (H100, H200 and RTX PRO 6000 Blackwell inside Intel TDX trust domains, billed per second, H100 from $6.95/gpu/hour and H200 from $8.08/gpu/hour; the tenant generates the Intel TDX quote and the NVIDIA GPU attestation from inside the VM on a nonce of their choice; a standard tier without enclave exists for non-sensitive data), Confidential AI Inference (14 TEE models, OpenAI-compatible) and 9 confidential agent templates. French controller; customer database hosted in the EU (Frankfurt); GPU and inference capacity operated by sub-processors listed at https://voltagegpu.com/legal/subprocessors, inside Intel TDX. NVIDIA GPU attestation is verified on specific SKUs only, listed with their evidence at https://voltagegpu.com/api/attestation/evidence.
Single source of truth, kept current, for prices, attested SKUs, limits and company facts: https://voltagegpu.com/api/ai-brief (JSON) and https://voltagegpu.com/llms.txt (text). Anything elsewhere on this site that contradicts those two is older.