Uncensored Model HubUncensored Models: API Quickstart
Uncensored Models: API Quickstart
Get started with the uncensored models API in minutes. This guide covers authentication, endpoints, and usage patterns using standard OpenAI-compatible clients.
https://api.uncensoredmodelhub.com/v1
Base URL and Authentication
Access the API by pointing your client to https://api.uncensoredmodelhub.com/v1. The service uses standard API key authentication. Generate your key on the Get API key page using just an email and password. Include this key in the Authorization header as a bearer token. Each account supports one active key at a time; you can regenerate it to revoke the previous one. The API serves a single open-weight model identified as uncensored, which does not refuse lawful adult, fictional, or controversial topics, though it blocks sexual content involving minors.
First Request
Send a standard chat completion request to test connectivity. The API expects a messages array containing your conversation history or a single user prompt. The model returns raw text without filtering lawful adult content, provided it does not violate the hard limit on minor sexual content. Use the official OpenAI SDK or any compatible HTTP client by setting the base URL and API key.
curl https://api.uncensoredmodelhub.com/v1/chat/completions \
-H "Authorization: Bearer $API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "uncensored",
"messages": [{"role": "user", "content": "Write a blunt product review of a cheap VPN."}]
}'
Python SDK
Use the official openai Python library for the fastest integration. Configure the client with your API key and the custom base URL. The SDK handles serialization and streaming automatically. This approach is ideal for backend services or scripts that need to generate large blocks of text without handling HTTP streams manually.
from openai import OpenAI
client = OpenAI(base_url="https://api.uncensoredmodelhub.com/v1", api_key="YOUR_KEY")
resp = client.chat.completions.create(
model="uncensored",
messages=[{"role": "user", "content": "Summarise this thread without softening it."}],
)
print(resp.choices[0].message.content)
Node SDK
For JavaScript environments, use the openai Node.js package. Set the baseURL to the uncensored API endpoint. This allows you to integrate unfiltered text generation directly into web servers or Node-based applications. The SDK manages connection pooling and error handling consistent with the OpenAI standard.
import OpenAI from "openai";
const client = new OpenAI({ baseURL: "https://api.uncensoredmodelhub.com/v1", apiKey: process.env.API_KEY });
const resp = await client.chat.completions.create({
model: "uncensored",
messages: [{ role: "user", content: "Draft a villain monologue for my game." }],
});
console.log(resp.choices[0].message.content);
Streaming Responses
Enable streaming by setting stream: true in your request. The API returns a Server-Sent Events (SSE) stream, delivering tokens as they are generated. This reduces perceived latency for interactive applications. Note that the model has a context window of 100,000 tokens for both prompt and completion combined. Streaming is useful for real-time display but does not change the underlying token pricing.
stream = client.chat.completions.create(
model="uncensored",
messages=[{"role": "user", "content": "Tell the story in second person."}],
stream=True,
)
for chunk in stream:
if chunk.choices and chunk.choices[0].delta.content:
print(chunk.choices[0].delta.content, end="", flush=True)
Limits, Errors, and Context
The API enforces a rate limit of 300 requests per minute per key and an 8 MB request body limit. Authentication failures return a 401 error if the key is invalid or revoked. Requests fail with a 402 error if your prepaid credit is insufficient. A 429 error indicates you have exceeded the per-minute rate limit. Pricing is transparent: $0.25 per 1M input tokens and $1.00 per 1M output tokens. Credit never expires, and new accounts receive $0.50 in trial credit for 7 days.
Under the hood: specs
Everything the endpoint can and cannot do, in one place — check it before you top up.
| Parameter | Details |
|---|---|
| API format | OpenAI Chat Completions schema; official openai SDKs work unchanged |
| Authentication | Authorization: Bearer YOUR_KEY |
| Endpoints | POST /v1/chat/completions · GET /v1/models |
| Base URL | https://api.uncensoredmodelhub.com/v1 |
| Model ID | uncensored |
| Streaming | Yes — server-sent events; the last chunk carries token usage |
| Tools / tool calls | Yes — tools, tool_choice; replies carry tool_calls, also when streaming; send results back as role: tool |
| Sampling parameters | temperature, top_p, stop, seed and the two penalties are passed through |
| JSON mode | response_format: {"type": "json_object"} |
| Max output | 16,000 tokens max; 2,048 if max_tokens is not set |
| Context window | 100,000 tokens (prompt + completion together) |
| Requests per minute | 300/min per key |
| Parallel requests | up to 8 in parallel per key |
| Headers | X-Request-Id, X-Balance-USD, X-RateLimit-Limit-Requests, X-RateLimit-Limit-Concurrency |
| Request size | up to 8 MB per request |
| Subscription | no monthly fee; paid credit does not expire |
| Free trial | $0.50 for 7 days, no card |
| How you pay | pay as you go from prepaid credit; nothing is charged for failed or refused requests |
| Volume bonus | +5% from $50, +10% from $100 |
| Price | input $0.25 / 1M tokens, output $1.00 / 1M tokens |
| Top-up | crypto: USDT on TRON or USDC on Base, $10–$500, any whole sum |
| Keys | one key per account, regenerate any time (the old one stops working) |
| Content policy | adult content allowed; sexual content involving minors is refused |
| Sign-in | Google or e-mail and password |
Errors and what to do
Every error is JSON with a type you can switch on. You are never charged for an error.
| Code | Type | Meaning |
|---|---|---|
400 | bad_request | malformed request or too long for the context window |
401 | missing_key · invalid_key · key_revoked | check the Authorization header or use your current key |
402 | no_credit | out of credit; add credit and retry |
403 | content_blocked | refused by the content policy |
404 | not_found | only /v1/chat/completions and /v1/models exist |
413 | request_too_large | body over 8 MB |
429 | rate_limited · concurrency | slow down: rate or parallel limit reached |
503 | upstream_busy | model busy — retry in a few seconds |
Questions and answers
Is the model really uncensored?
The model does not refuse lawful adult, creative, or controversial topics. It is tuned to answer directly without the typical corporate guardrails. However, it always blocks sexual content involving minors, which is a hard limit.
Can I use this with the OpenAI SDK?
Yes. The API is fully OpenAI-compatible. You only need to change the base URL and API key in your existing client configuration. It supports standard endpoints like <code>/v1/chat/completions</code> and streaming via SSE.
How does pricing work?
You pay as you go with prepaid credit. Input tokens cost $0.25 per million, and output tokens cost $1.00 per million. There are no monthly fees or subscriptions. You can top up with crypto (USDT or USDC), starting from $10.
Your key is one form away
Create an account, copy the key, change the base URL. That is the whole setup.
Get API key