About Blackbox
Overview
Blackbox is a high-trust platform for frontier inference. It runs the open-weight model you choose as a dedicated deployment on Blackbox GPUs, or routes 300+ models through one endpoint. Prompts are encrypted before they leave your machine and decrypted only where the model runs, so nothing readable crosses the wire.
The gateway enforces zero data retention and keeps training opt-out on by default. Artificial Analysis independently measured Blackbox as the #1 Nemotron 3 Ultra provider, at 454 tokens per second and 2.7× lower cost than the #2 provider. Route across 300+ models and send only the data each task requires, or run workloads you cannot share on a deployment that serves only you.
Key Benefits
- Dedicated single-tenant deployments run your chosen open-weight model on Blackbox GPUs, isolated from every other customer.
- One endpoint connects you to 300+ hosted models with a single key, bill, and dashboard.
- End-to-end encryption seals prompts in the client, so the proxy, the host, and the operator never hold the key.
- The gateway enforces zero data retention and no training on routed traffic through provider terms and per-request flags.
- Artificial Analysis ranked Blackbox the #1 Nemotron 3 Ultra provider at 454 tokens per second.
- OpenAI-compatible REST and streaming let you change the base URL and keep your existing code.
How It Works
You commit to a token volume through a purchase order, and the balance decreases as your teams consume it at published per-model rates. A forward-deployed engineer then configures the account, including controls such as PII removal, at contract start. For API access, you point three lines of code at an OpenAI-compatible endpoint and keep your existing code.
Use Cases
- AI platform engineers route production traffic across 300+ models through one OpenAI-compatible endpoint.
- Security teams run regulated workloads on a dedicated single-tenant deployment with data residency.
- Enterprise developers build cloud coding agents with the Agents API and CLI on the same per-token commit.
- IT administrators manage access with SAML SSO, SCIM, RBAC, and audit logs.
- Machine learning teams compare open-weight models against published per-model token rates before committing.
Why Choose This Product
Blackbox operates its own GPUs end to end for open-weight models, which is why open-weight commits earn a larger discount than closed models. Closed models route to their upstream providers and bill through those relationships. There are no platform fees, no markups, and no per-seat charges, and metering runs per token rather than per seat.
Key Features
End-to-End Encryption
All plansPrompts are encrypted before they leave your machine and decrypted only where the model runs, and the proxy, host, and operator never hold the key.
Zero Data Retention
All plansThe gateway enforces zero data retention and no training on routed traffic, so nothing readable is kept after the answer.
Dedicated Deployment
All plansEnterprise Inference runs the open-weight model you choose on reserved Blackbox GPUs, single-tenant and isolated to you.
Model Router
All plansOne endpoint connects you to 300+ hosted models with one key, one bill, and one dashboard.
OpenAI-Compatible API
All plansAn OpenAI-compatible REST and streaming endpoint lets you change the base URL and keep your existing code.
Agents API & CLI
All plansThe Agents API and the CLI build cloud coding agents and draw on the same per-token commit as every other surface.
Speed-Optimized Inference
All plansArtificial Analysis ranked Blackbox the #1 Nemotron 3 Ultra provider at 454 tokens per second, 30% faster than the #2 provider.
PII Anonymization
All plansOn Enterprise, Blackbox removes PII before prompts reach a closed model.
Smart Routing
All plansSmart routing, failover, and prompt caching lower cost and extend the same commit by 10–20%, even on closed models.
Blackbox Pricing
View full pricing →- Dedicated single-tenant deployment isolated from every other customer
- Data residency and custom SLAs
- SAML SSO, SCIM, RBAC, and audit logs
- Zero data retention, contractual with DPA
- PII removed before closed models
- Dedicated forward-deployed engineer, implementation included at $0
Pricing extracted from the product website and may change. Check the source for current details.

