> I have to agree with random. I'm using it at Q6 on 32k context size and it has zero issues following the system prompt, following the character details and remembering what happened even in the higher context limits. I'm currently using it in a group chat with 3 characters where I have join character cards enabled and it plays each character really well and still includes the smaller character details. Only a few minor mishaps where it'll get a small detail mixed up and confused, but with a swipe it's fixed. Now using this as my main model because nothing has beaten it yet.
# Paste into Claude Code, Cursor, or Codex
Help me set up the Respan gateway: fetch and follow https:/
One provider serves this model through the Respan gateway. Copy a model ID to call that route directly.
| Provider | Input | Output |
|---|---|---|
Featherless featherless_ai/TheDrummer/Cydonia-24B-v4.3 | $0.70/M | $1.16/M |
USD per 1M tokens on Featherless, after gateway discounts.
Total cost from 0 to 1T tokens.
30-day averages per provider from requests through the Respan gateway, best first.
3tokens/son Featherless
3.0son Featherless
Use Cydonia 24B V4.3 in the tools you already run. Each guide points the app at the Respan gateway, then you pick this model.
Setup guides for more apps are in the integration docs.
Built to meet the security and privacy standards that enterprise and healthcare teams require.
ISO 27001
The internationally recognized standard for information security management.
SOC 2
Secure, compliant management of your data across all of our systems.
GDPR
Operated under GDPR, the world's strictest standard for data privacy.
HIPAA
HIPAA compliant, with a BAA available for healthcare teams.