blog · Aug 28, 2026 · 3 min · by the eroq team
How to use eroq with SillyTavern — uncensored API in five minutes
Point SillyTavern's Chat Completion source at eroq's OpenAI-compatible endpoint and your character cards run on an uncensored, roleplay-tuned engine with flat credit pricing. The full setup, plus which settings actually matter.
SillyTavern is the power tool of character chat — cards, lorebooks, group chats, total control — and it's bring-your-own-API by design. That makes it a perfect match for eroq: an OpenAI-compatible endpoint whose engines are tuned for exactly what SillyTavern users do all day, with no refusal wall in the middle of your scenes.
Here's the whole setup, start to finish.
1. Get a key
Create an account — it starts with 50 free credits, no card — and mint a key on the API keys page. The key is shown once; keep it somewhere safe.
2. Point SillyTavern at eroq
In SillyTavern:
- Open API Connections (the plug icon).
- API: select Chat Completion.
- Chat Completion Source: select Custom (OpenAI-compatible).
- Custom Endpoint (Base URL):
https://eroq.ai/v1 - API Key: paste your
eroq_sk_…key. - Model: enter
eroq-rp-plus(oreroq-rp-mini— see below). - Hit Connect — the status dot goes green, and you're running on eroq.
That's the entire integration. Streaming works out of the box (standard SSE), so replies render token by token like any other source.
3. Pick your model like you mean it
eroq-rp-plus(3 credits per reply) — the flagship: persona hold across long scenes, repetition damping tuned at RP temperatures, the prose texture your best cards deserve.eroq-rp-mini(1 credit per reply) — same persona handling, sized for speed and volume. The right default for fast back-and-forth and group chats where replies fly.
Both are uncensored within the written policy: adult roleplay between adult characters renders in character — no jailbreak prompt required, no sudden Thursday refusals. Prompts outside the hard limits return an explicit error instead of a lecture, and are never charged.
4. Settings that actually matter
- Temperature: RP+ is tuned around
0.8— start there rather than importing exotic presets. If you drag any model far from its tuning point, that's where loops begin. - Max tokens: eroq caps completions at 1,200; a
max_tokensaround 400–500 fits most card styles and keeps replies snappy. - Context size: credits are flat per reply — a huge context costs you nothing extra, but a curated one roleplays better. SillyTavern's summarize extension plus a reasonable context beats raw size.
- System prompt: your cards' system blocks replace eroq's neutral preamble entirely — full control, as SillyTavern intends.
5. What the flat pricing means for a SillyTavern habit
No token math, ever: a reply is 1 or 3 credits whether your card has a 200-token or 2,000-token context. A heavy evening — say 150 replies on RP mini — is 150 credits, about $1.50 at pack rates. The pricing page has the full table, and every response carries usage.credits_remaining so you always know where you stand. When the balance runs out, calls fail fast with a clear error and nothing is charged — top up and continue exactly where the scene stopped.
Troubleshooting
- 401 — key pasted with a space, or revoked. Mint a fresh one; it's instant.
- 402 — balance empty. Top up; the scene resumes where it stopped.
- A refusal? If a prompt crosses the hard limits (minors, real people, illegal content) you'll get an explicit
content_blockederror — deterministic, never charged. Lawful adult fiction between adult characters does not trigger it.
Five minutes, three strings, and your cards run on an engine that was tuned for them. If something in the setup fights you, support answers within a business day.
Build with the models behind this post — get an API key (50 free credits).