OpenAI compatible API · Attested · Public status
Parasail
Parasail models on TrustedRouter with prices, routes, policy notes, and source links.
Onebase URL to migrate
100sof models and routes
0prompt or output logs. Always.
Parasailparasail
No logs
| Provider | Parasail |
|---|---|
| Provider website | https://www.parasail.io/ |
| Models | 26 public models |
| Prepaid routes | 26 |
| BYOK routes | 26 |
| Zero data retention | yes |
| Confidential compute | not claimed |
| Provider E2EE | not claimed |
| Policy note | Tracked as ZDR for serverless and dedicated inference. Parasail documents no storage or logging of submitted input on those service paths, retention only while generating and delivering output, and no training on input or output. Batch service is excluded from this claim; TrustedRouter does not route Parasail traffic through batch. Policy source |
Measured performance
27 samplesContinuously sampled across Parasail's routed models: p50 TTFT, effective throughput, and success rate. Effective throughput uses provider-reported output tokens over complete request time. Unsupported route and probe-configuration rows are separated from provider downtime. No prompt or output content stored.
| p50 TTFT | 1794 ms |
|---|---|
| Effective throughput | — |
| Uptime | 88.89% |
| Model | p50 TTFT | p50 TTFB | Effective throughput | Uptime | Config excluded | Availability samples |
|---|---|---|---|---|---|---|
| meta-llama/llama-3.3-70b-instruct | 900 ms | 900 ms | — | 100.00% | — | 2 |
| qwen/qwen2.5-vl-72b-instruct | 1439 ms | 1439 ms | — | 100.00% | — | 1 |
| google/gemma-3-27b-it | 1462 ms | 1462 ms | — | 100.00% | — | 2 |
| qwen/qwen3-vl-235b-a22b-instruct | 1577 ms | 1577 ms | — | 100.00% | — | 2 |
| deepseek/deepseek-v4-pro | 1633 ms | 1633 ms | — | 100.00% | — | 2 |
| google/gemma-4-26b-a4b-it | 1711 ms | 1711 ms | — | 100.00% | — | 2 |
| minimax/minimax-m3 | 1727 ms | 1727 ms | — | 100.00% | — | 1 |
| thedrummer/skyfall-36b-v2 | 1794 ms | 1793 ms | — | 100.00% | — | 1 |
| arcee-ai/trinity-large-thinking | 1807 ms | 1807 ms | — | 100.00% | — | 2 |
| z-ai/glm-5.2 | 1835 ms | 1835 ms | — | 100.00% | — | 1 |
| deepseek/deepseek-v4-flash | 1838 ms | 1838 ms | — | 100.00% | — | 2 |
| moonshotai/kimi-k2.7-code | 1923 ms | 1923 ms | — | 100.00% | — | 1 |
| openai/gpt-oss-20b | 1972 ms | 1972 ms | — | 100.00% | — | 1 |
| qwen/qwen3-coder-next | 2357 ms | 2357 ms | — | 100.00% | — | 1 |
| qwen/qwen3-next-80b-a3b-instruct | 4125 ms | 4125 ms | — | 100.00% | — | 1 |
| openai/gpt-oss-120b | 8417 ms | 8417 ms | — | 100.00% | — | 1 |
| qwen/qwen3-vl-8b-instruct | 5068 ms | 5068 ms | — | 50.00% | — | 2 |
| qwen/qwen3.5-35b-a3b | — | — | — | 0.00% | — | 1 |
| thedrummer/cydonia-24b-v4.1 | — | — | — | 0.00% | — | 1 |
Parasail performance history · Full provider & model leaderboard.
Provider models
Models served by Parasail.
Each row links to pricing, provider, benchmark, and API pages for the model.
| Model | AI IQ | Context | Endpoints | Prompt | Completion | Routes |
|---|---|---|---|---|---|---|
arcee-ai/trinity-large-thinkingArcee AI: Trinity Large Thinking |
IQ 97#79 | 262,144 | 2 | $0.231/1M | $0.8925/1M | prepaid BYOK |
bytedance/ui-tars-1.5-7bByteDance: UI-TARS 7B |
— | 128,000 | 2 | $0.105/1M | $0.21/1M | prepaid BYOK |
deepseek/deepseek-v4-flashDeepSeek: DeepSeek V4 Flash 0423 |
IQ 108#46 | 1,048,576 | 2 | $0.147/1M | $0.294/1M | prepaid BYOK |
deepseek/deepseek-v4-proDeepSeek: DeepSeek V4 Pro |
IQ 115#30 | 1,048,576 | 2 | $1.827/1M | $3.654/1M | prepaid BYOK |
google/gemma-3-27b-itGoogle: Gemma 3 27B |
— | 262,144 | 2 | $0.084/1M | $0.4725/1M | prepaid BYOK |
google/gemma-4-26b-a4b-itGoogle: Gemma 4 26B A4B |
IQ 96#81 | 262,144 | 2 | $0.1365/1M | $0.42/1M | prepaid BYOK |
google/gemma-4-31b-itGoogle: Gemma 4 31B |
IQ 101#66 | 262,144 | 2 | $0.1575/1M | $0.42/1M | prepaid BYOK |
meta-llama/llama-3.3-70b-instructMeta: Llama 3.3 70B Instruct |
— | 131,072 | 2 | $0.231/1M | $0.525/1M | prepaid BYOK |
meta-llama/llama-4-maverickMeta: Llama 4 Maverick |
IQ 90#91 | 1,048,576 | 2 | $0.3675/1M | $1.05/1M | prepaid BYOK |
minimax/minimax-m3MiniMax: MiniMax M3 |
IQ 114#33 | 1,048,576 | 2 | $0.315/1M | $1.26/1M | prepaid BYOK |
moonshotai/kimi-k2.6MoonshotAI: Kimi K2.6 |
IQ 119#19 | 262,144 | 2 | $0.7875/1M | $3.675/1M | prepaid BYOK |
moonshotai/kimi-k2.7-codeMoonshotAI: Kimi K2.7 Code |
IQ 118#22 | 262,144 | 2 | $0.7875/1M | $3.675/1M | prepaid BYOK |
openai/gpt-oss-120bOpenAI: gpt-oss-120b |
IQ 105#53 | 131,072 | 2 | $0.105/1M | $0.7875/1M | prepaid BYOK |
openai/gpt-oss-20bOpenAI: gpt-oss-20b |
IQ 100#70 | 131,072 | 2 | $0.042/1M | $0.21/1M | prepaid BYOK |
qwen/qwen2.5-vl-72b-instructQwen: Qwen2.5 VL 72B Instruct |
— | 128,000 | 2 | $0.84/1M | $1.05/1M | prepaid BYOK |
qwen/qwen3-coder-nextQwen: Qwen3 Coder Next |
— | 262,144 | 2 | $0.126/1M | $0.84/1M | prepaid BYOK |
qwen/qwen3-next-80b-a3b-instructQwen: Qwen3 Next 80B A3B Instruct |
— | 262,144 | 2 | $0.105/1M | $1.155/1M | prepaid BYOK |
qwen/qwen3-vl-235b-a22b-instructQwen: Qwen3 VL 235B A22B Instruct |
— | 262,144 | 2 | $0.2205/1M | $1.995/1M | prepaid BYOK |
qwen/qwen3-vl-8b-instructQwen: Qwen3 VL 8B Instruct |
— | 262,144 | 2 | $0.2625/1M | $0.7875/1M | prepaid BYOK |
qwen/qwen3.5-35b-a3bQwen: Qwen3.5-35B-A3B |
— | 262,144 | 2 | $0.1575/1M | $1.05/1M | prepaid BYOK |
qwen/qwen3.5-397b-a17bQwen: Qwen3.5 397B A17B |
— | 262,144 | 2 | $0.525/1M | $3.78/1M | prepaid BYOK |
qwen/qwen3.6-35b-a3bQwen: Qwen3.6 35B A3B |
IQ 100#71 | 262,144 | 2 | $0.1575/1M | $1.05/1M | prepaid BYOK |
thedrummer/cydonia-24b-v4.1TheDrummer: Cydonia 24B V4.1 |
— | 131,072 | 2 | $0.315/1M | $0.525/1M | prepaid BYOK |
thedrummer/skyfall-36b-v2TheDrummer: Skyfall 36B V2 |
— | 32,768 | 2 | $0.5775/1M | $0.84/1M | prepaid BYOK |
z-ai/glm-5.1Z.ai: GLM 5.1 |
IQ 114#32 | 204,800 | 2 | $1.47/1M | $4.62/1M | prepaid BYOK |
z-ai/glm-5.2Z.ai: GLM 5.2 |
IQ 120#16 | 1,048,576 | 2 | $1.47/1M | $4.62/1M | prepaid BYOK |