Abliterated Models API Documentation
Send prompts to an uncensored large language model using standard OpenAI-compatible endpoints. The API serves a single abliterated model optimized for removing refusal patterns while maintaining a 100,000-token context window.
https://api.abliteratedmodelhub.com/v1uncensored
Base URL and Authentication
Use the base URL https://api.abliteratedmodelhub.com/v1 with any standard OpenAI-compatible client. Pass your API key in the Authorization header as a Bearer token. Every account starts with a unique key shown immediately after signup, which you can regenerate at any time to revoke the old one. No phone number or credit card is required to begin.
Authentication failures return a 401 status if the key is invalid or missing. Billing errors return a 402 status if your prepaid credit is exhausted. You can top up from $10 using crypto (USDT or USDC), and bonus credit is applied at higher tiers.
First Request
Send a basic chat request to the /v1/chat/completions endpoint. The model ID is uncensored. This is a text-in, text-out API designed for abliterated models that do not refuse lawful adult or controversial topics.
curl https://api.abliteratedmodelhub.com/v1/chat/completions \
-H "Authorization: Bearer $API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "uncensored",
"messages": [{"role": "user", "content": "Write a blunt product review of a cheap VPN."}]
}'The request body includes the standard messages array. The model responds with generated text. This is not a multi-model marketplace; you are interacting with one specific open-weight model tuned for minimal refusal.
Python SDK
Use the official OpenAI Python SDK or any compatible library. Set the base URL to our endpoint and provide your API key. The client treats our service as a standard OpenAI-compatible provider.
from openai import OpenAI
client = OpenAI(base_url="https://api.abliteratedmodelhub.com/v1", api_key="YOUR_KEY")
resp = client.chat.completions.create(
model="uncensored",
messages=[{"role": "user", "content": "Summarise this thread without softening it."}],
)
print(resp.choices[0].message.content)This approach works for simple text generation or complex prompt engineering. The model handles the abliterated weights on our GPU servers, so you do not need to manage local inference or VRAM constraints.
Node SDK
Initialize the OpenAI client in Node.js with the custom base URL. Pass your API key as the authorization token. The model parameter should be set to uncensored.
import OpenAI from "openai";
const client = new OpenAI({ baseURL: "https://api.abliteratedmodelhub.com/v1", apiKey: process.env.API_KEY });
const resp = await client.chat.completions.create({
model: "uncensored",
messages: [{ role: "user", content: "Draft a villain monologue for my game." }],
});
console.log(resp.choices[0].message.content);Node.js developers can use this setup for server-side generation or integration into larger applications. The API supports standard JSON payloads and returns structured completions.
Streaming Responses (SSE)
Enable streaming by setting stream: true in your request. The API returns Server-Sent Events (SSE) for real-time token delivery. This is useful for UIs that display text as it is generated.
stream = client.chat.completions.create(
model="uncensored",
messages=[{"role": "user", "content": "Tell the story in second person."}],
stream=True,
)
for chunk in stream:
if chunk.choices and chunk.choices[0].delta.content:
print(chunk.choices[0].delta.content, end="", flush=True)Streaming does not change the model behavior or context window. It only affects how the response is delivered. Each chunk contains a partial completion until the generation is complete.
Rate Limits, Errors, and Context
Each API key is limited to 300 requests per minute. The maximum request body size is 8 MB. The context window is 100,000 tokens for both input and output combined. If you exceed the limit, the API returns a 429 status code. Requests are blocked if they contain sexual content involving minors, which is a hard content limit. Prompts are not used for training your data.
What the API supports
One table with every limit, feature and price that applies to your key.
| Feature | Support |
|---|---|
| Protocol | OpenAI-compatible: any OpenAI SDK or client works — change the base URL and the key |
| Model | uncensored |
| Base URL | https://api.abliteratedmodelhub.com/v1 |
| Methods | POST /v1/chat/completions · GET /v1/models |
| API key | Authorization: Bearer YOUR_KEY |
| Completion length | 16,000 tokens max; 2,048 if max_tokens is not set |
| Tools / tool calls | Yes — tools, tool_choice; replies carry tool_calls, also when streaming; send results back as role: tool |
| Other parameters | temperature, top_p, stop, seed, presence_penalty, frequency_penalty |
| SSE streaming | Yes — server-sent events; the last chunk carries token usage |
| JSON mode | JSON object mode via response_format json_object |
| Context window | 100,000 tokens (prompt + completion together) |
| Max body | up to 8 MB per request |
| Headers | X-Request-Id, X-Balance-USD, X-RateLimit-Limit-Requests, X-RateLimit-Limit-Concurrency |
| Concurrency | up to 8 in parallel per key |
| Rate limit | 300 requests per minute per key |
| Trial credit | $0.50 for 7 days, no card |
| Billing | prepaid credit, charged by real token usage; errors and refusals are free |
| Volume bonus | +5% on $50+, +10% on $100+ |
| Payment | crypto: USDT on TRON or USDC on Base, $10–$500, any whole sum |
| Subscription | no monthly fee; paid credit does not expire |
| Token prices | input $0.25 / 1M tokens, output $1.00 / 1M tokens |
| Content policy | adult content allowed; sexual content involving minors is refused |
| Keys | one active key per account; a new key replaces the old one |
| Account | Google or e-mail and password |
Error reference
Every error is JSON with a type you can switch on. You are never charged for an error.
| Status | Type | Reason |
|---|---|---|
400 | bad_request | invalid JSON, empty messages, bad parameter, or prompt + max_tokens over the window — fix and resend |
401 | missing_key · invalid_key · key_revoked | no key, wrong key, or a key replaced by a newer one |
402 | no_credit | out of credit; add credit and retry |
403 | content_blocked | refused by the content policy |
404 | not_found | only /v1/chat/completions and /v1/models exist |
413 | request_too_large | body over 8 MB |
429 | rate_limited · concurrency | over 300/min or 8 parallel — back off and retry |
503 | upstream_busy | model busy — retry in a few seconds |
Questions and answers
What does "uncensored" mean for this model?
The model is tuned to answer without content refusals for lawful adult use, including controversial or fictional topics. It does not apply general political or social bias filters that typical models use. However, requests containing sexual content involving minors are always blocked.
Is this model the same as GPT or Claude?
No. The model ID is <code>uncensored</code>, an open-weight model run on our own GPU servers. It is not GPT, Claude, Gemini, Grok, or any other vendor's model. It is specifically abliterated to remove refusal patterns.
Do I need a credit card to start?
No. Every new account gets $0.50 of trial credit valid for 7 days. You can sign up with just an email and password. You only need crypto (USDT or USDC) when you want to top up your prepaid balance.
Your key is one form away
Create an account, copy the key, change the base URL. That is the whole setup.