Provider: Parasail

Call Parasail models through Respan Gateway with unified logs, cost, and latency.
This page is for Respan LLM Gateway users.

Use Respan Gateway to call Parasail-hosted models (DeepSeek, Llama, Qwen, and other open-source deployments on Parasail’s serverless inference) while keeping unified observability (logs, cost, latency, reliability) in Respan.

Quick setup

1

Get a Respan API key

Sign up and create a key on the API keys page.

Send your first request

Pick the integration that matches your stack. The base URL is https://api.respan.ai/api and the only key needed is your RESPAN_API_KEY.

Parasail is OpenAI-compatible. Point the OpenAI SDK at the Respan gateway and call any Parasail model.

from openai import OpenAI
client = OpenAI(
api_key="YOUR_RESPAN_API_KEY",
base_url="https://api.respan.ai/api",
)
response = client.chat.completions.create(
model="parasail/deepseek-ai/DeepSeek-V3",
messages=[{"role": "user", "content": "Hello, Parasail!"}],
)
print(response.choices[0].message.content)

More integrations

Parasail models work with every Respan gateway integration:

Switch models

Change the model parameter to call any supported model through the same client. Use the parasail/ prefix to disambiguate when routing across providers. Browse the full list on the Models page.

client.chat.completions.create(model="parasail/deepseek-ai/DeepSeek-V3", messages=messages)
client.chat.completions.create(model="parasail/meta-llama/Llama-3.3-70B-Instruct", messages=messages)
client.chat.completions.create(model="parasail/Qwen/Qwen3-235B-A22B", messages=messages)
client.chat.completions.create(model="openai/gpt-5.5", messages=messages)
client.chat.completions.create(model="anthropic/claude-sonnet-4-5", messages=messages)

Use your own Parasail key (BYOK)

Credits are the default path. If you’d rather bill Parasail directly, attach your own provider key.

1

Open Providers

Go to the Providers page.

2

Add Parasail

Select Parasail and paste your parasail.api_key.

3

Load balancing (Optional)

Add multiple credential sets and use Load balancing weight to distribute traffic across them.

Override credentials per model (Optional)

Use credential_override when one model on a request should use a different Parasail key than the default.

{
"customer_credentials": {
"parasail": { "api_key": "YOUR_PARASAIL_API_KEY" }
},
"credential_override": {
"parasail/deepseek-ai/DeepSeek-V3": { "api_key": "ANOTHER_PARASAIL_API_KEY" }
}
}

Log without proxying (Optional)

Already calling Parasail directly? Send logs to Respan asynchronously to track cost, latency, and performance for those external calls.

import requests
requests.post(
"https://api.respan.ai/api/request-logs/create/",
headers={
"Authorization": "Bearer YOUR_RESPAN_API_KEY",
"Content-Type": "application/json",
},
json={
"model": "parasail/deepseek-ai/DeepSeek-V3",
"prompt_messages": [{"role": "user", "content": "Hello, how are you?"}],
"completion_message": {"role": "assistant", "content": "Hello from Parasail through Respan."},
"cost": 0.001,
"generation_time": 1.2,
"customer_params": {"customer_identifier": "user_123"},
},
)

See the logging guide for the full setup.