TB v2.1 Blackbox — GPT-5.6 Sol + Opus 4.8 (90% pass@1)
A two-model Blackbox reaches 90.2% pass@1 on Terminal-Bench v2.1 (AA basis) — leading the Artificial Analysis leaderboard — by using a sandbox-executing critic to gate exactly one graded answer per task.
READ ARTICLEBenchmark Performance: Faster Inference, Reference-Level Model Quality
Serving a model through the BLACKBOX AI API gives higher throughput and lower latency with no observed loss in quality. Our inference stack speeds up token generation without retraining or changing model weights, and our benchmark runs show no observed regression against the published GLM 5.2, NVIDIA Ultra, and Kimi K2.7 references.
READ ARTICLEArtificial Analysis: BLACKBOX AI Is the #1 Fastest Nemotron 3 Ultra Provider
Artificial Analysis independently benchmarks every API provider serving NVIDIA Nemotron 3 Ultra. In the latest snapshot, BLACKBOX AI holds #1 output speed — 454.4 tokens per second, 47% ahead of the runner-up, at roughly a third of its price.
READ ARTICLEDocker OpenClaw E2EE
Point OpenClaw at a BLACKBOX AI end-to-end encrypted model from inside a Docker container. All requests are sealed with ECDH + AES-256-GCM before they leave your machine and only decrypted inside the GPU enclave.
READ ARTICLEOrchestrator–Executor: A Two-Agent Split That Beats a Solo Model on SWE Tasks
Split the agent in two — a strong orchestrator that plans and verifies, a cheaper executor that implements — and Terminal-Bench 2.0 scores jump from 58.4 to 69.7.
READ ARTICLEBlackbox Encrypted AI: Confidential LLMs at 22× Lower Cost
End-to-end encrypted LLM inference with hardware attestation, at $0.45 per million output tokens — 22× cheaper than GPT-4o, with confidentiality guarantees TLS alone can't provide.
READ ARTICLEZero Data Retention: How Blackbox Enforces It By Default
Prompt and response content is not retained. Zero Data Retention is applied to every request by default, and where we can't meet it, we reject the request rather than serving it without protection — with usage metadata for billing retained transparently.
READ ARTICLE