Chat apps

TypingMind with Spicy API: uncensored role-play models

TypingMind is a web chat front end for many AI models with your own keys, at typingmind.com. This guide connects it to Spicy API (spicyapi.com), so TypingMind can use spicy-companion-1, a role-play model that holds a persona and writes explicit adult scenes, or spicy-companion-1-flash, the same tuning for about a ninth of the cost of a long chat. You pay per token from a prepaid balance, with no subscription.

What you need

  • A Spicy API account: sign in at /auth and accept the terms.
  • A separate key for TypingMind: /dashboard/api-keys, then Create key. You can revoke it without breaking anything else. TypingMind calls the API from your browser and keeps the key in its settings, so never reuse a key your own apps depend on.
  • A monthly spend limit on the account (same page, Monthly spend limit). Past it, requests fail with a 402 and nothing is charged, so a leaked key cannot drain your balance.
  • A balance (top up from $50 by card or crypto at /dashboard/account), or a sandbox key to try the setup first: tick Sandbox when you create the key. It answers with a fixed reply and bills nothing.

Connect TypingMind

  1. Open the model menu above the chat box and choose Custom models, then Add Custom Model (Create Manually).
  2. Name: Spicy Companion 1. Model ID: spicy-companion-1. Context Length: 131072.
  3. API Type: OpenAI Chat Completions API. Endpoint URL: https://api.spicyapi.com/v1/chat/completions (the full URL).
  4. Authentication Type: API Key via HTTP Header, Header Key Authorization, Header Value Bearer sk-spicy-... with your key.
  5. Click Test & Save. TypingMind probes the model: system role, streaming and the sampling parameters pass; plugins, image and PDF input, thinking and caching fail, which is right for a text-only model. Click Confirm & Proceed, then Test & Save again.
  6. Pick Spicy Companion 1 in the model menu and chat.

Which model to pick

ModelBest forContext (tokens)Per 1M tokens (prompt / completion)100 messages at 4k / 16k context
spicy-companion-1Role-play and companions: holds a persona, writes explicit adult scenes, group scenes.131,072$1.00 / $2.80$0.47 / $1.71
spicy-companion-1-flashThe same role-play tuning for about a ninth of the cost of a long chat. Fast.43,000 tested$0.10 / $0.80$0.06 / $0.19
spicy-chat-1General assistant. Not tuned for role-play; use a companion model for characters.200,000$0.80 / $2.40$0.38 / $1.38

Chat apps resend the whole conversation with every message, so the context size you set drives the cost far more than the length of the replies. The last column assumes a 250 token reply at 4k context and 400 at 16k. Every response carries its exact cost_usd.

SettingValue
Context Length131072 for spicy-companion-1; for spicy-companion-1-flash add a second model with 43000.

If something goes wrong

StatusMeaningFix
400Wrong model id, or an option the chat models do not take (tools on spicy-companion-1-flash, images, more than 8,192 max tokens)Pick spicy-companion-1, spicy-companion-1-flash or spicy-chat-1; the error names the chat models
401Key missing, mistyped or revokedPaste the key again, or make a new one
402Balance empty, or the monthly spend limit reachedTop up, or raise the limit on the API keys page. Nothing is charged
422Blocked by moderationNothing is charged. See what is allowed below
429Too many requests in a minuteWait a minute; the app can retry
  • The capability test sends about a dozen short requests; on a live key the ones that pass bill the per-request minimum, well under a cent in total.
  • TypingMind names each new chat with a second request to the same model, a small extra charge on a live key.

What is allowed

Explicit role-play between adults is allowed, including non-consent between fictional adult characters (force, sleep, intoxication, framed as consensual non-consent or not) and other dark themes. Every message is screened before it reaches the model: minors in any form, real people (with or without consent), incest, real-world harm and the rest of the prohibited list are refused with a 422, before any charge. The screen reads the whole conversation, character card included, so a card that describes a minor or a real person is refused however the latest message is worded. Refusals in the most serious categories count as strikes. Send each user's id as user: the strikes then land on that user, who is refused after ten, and your account is only flagged. Without it they land on your account, flagged after three and suspended after ten. Drop a refused message from the history you send back, or the next message is refused too. A wrong refusal can be reported: POST its moderation_id to /v1/moderations/appeals. Full policy: /acceptable-use

Tested

typingmind.com (free web version) on 30 Sep 2026: custom model with header auth, the capability test, and a streamed chat turn with a sandbox key (plus the title request).

Other apps: /docs/integrations. Questions: contact@spicyapi.com.