Uncensored Model HubUncensored Models: API Quickstart

Uncensored Models: API Quickstart

Get started with the uncensored models API in minutes. This guide covers authentication, endpoints, and usage patterns using standard OpenAI-compatible clients.

https://api.uncensoredmodelhub.com/v1

Base URL and Authentication

Access the API by pointing your client to https://api.uncensoredmodelhub.com/v1. The service uses standard API key authentication. Generate your key on the Get API key page using just an email and password. Include this key in the Authorization header as a bearer token. Each account supports one active key at a time; you can regenerate it to revoke the previous one. The API serves a single open-weight model identified as uncensored, which does not refuse lawful adult, fictional, or controversial topics, though it blocks sexual content involving minors.

First Request

Send a standard chat completion request to test connectivity. The API expects a messages array containing your conversation history or a single user prompt. The model returns raw text without filtering lawful adult content, provided it does not violate the hard limit on minor sexual content. Use the official OpenAI SDK or any compatible HTTP client by setting the base URL and API key.

curl https://api.uncensoredmodelhub.com/v1/chat/completions \
  -H "Authorization: Bearer $API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "uncensored",
    "messages": [{"role": "user", "content": "Write a blunt product review of a cheap VPN."}]
  }'

Python SDK

Use the official openai Python library for the fastest integration. Configure the client with your API key and the custom base URL. The SDK handles serialization and streaming automatically. This approach is ideal for backend services or scripts that need to generate large blocks of text without handling HTTP streams manually.

from openai import OpenAI

client = OpenAI(base_url="https://api.uncensoredmodelhub.com/v1", api_key="YOUR_KEY")

resp = client.chat.completions.create(
    model="uncensored",
    messages=[{"role": "user", "content": "Summarise this thread without softening it."}],
)
print(resp.choices[0].message.content)

Node SDK

For JavaScript environments, use the openai Node.js package. Set the baseURL to the uncensored API endpoint. This allows you to integrate unfiltered text generation directly into web servers or Node-based applications. The SDK manages connection pooling and error handling consistent with the OpenAI standard.

import OpenAI from "openai";

const client = new OpenAI({ baseURL: "https://api.uncensoredmodelhub.com/v1", apiKey: process.env.API_KEY });

const resp = await client.chat.completions.create({
  model: "uncensored",
  messages: [{ role: "user", content: "Draft a villain monologue for my game." }],
});
console.log(resp.choices[0].message.content);

Streaming Responses

Enable streaming by setting stream: true in your request. The API returns a Server-Sent Events (SSE) stream, delivering tokens as they are generated. This reduces perceived latency for interactive applications. Note that the model has a context window of 100,000 tokens for both prompt and completion combined. Streaming is useful for real-time display but does not change the underlying token pricing.

stream = client.chat.completions.create(
    model="uncensored",
    messages=[{"role": "user", "content": "Tell the story in second person."}],
    stream=True,
)
for chunk in stream:
    if chunk.choices and chunk.choices[0].delta.content:
        print(chunk.choices[0].delta.content, end="", flush=True)

Limits, Errors, and Context

The API enforces a rate limit of 300 requests per minute per key and an 8 MB request body limit. Authentication failures return a 401 error if the key is invalid or revoked. Requests fail with a 402 error if your prepaid credit is insufficient. A 429 error indicates you have exceeded the per-minute rate limit. Pricing is transparent: $0.25 per 1M input tokens and $1.00 per 1M output tokens. Credit never expires, and new accounts receive $0.50 in trial credit for 7 days.

Under the hood: specs

Everything the endpoint can and cannot do, in one place — check it before you top up.

ParameterDetails
API formatOpenAI Chat Completions schema; official openai SDKs work unchanged
AuthenticationAuthorization: Bearer YOUR_KEY
EndpointsPOST /v1/chat/completions · GET /v1/models
Base URLhttps://api.uncensoredmodelhub.com/v1
Model IDuncensored
StreamingYes — server-sent events; the last chunk carries token usage
Tools / tool callsYes — tools, tool_choice; replies carry tool_calls, also when streaming; send results back as role: tool
Sampling parameterstemperature, top_p, stop, seed and the two penalties are passed through
JSON moderesponse_format: {"type": "json_object"}
Max output16,000 tokens max; 2,048 if max_tokens is not set
Context window100,000 tokens (prompt + completion together)
Requests per minute300/min per key
Parallel requestsup to 8 in parallel per key
HeadersX-Request-Id, X-Balance-USD, X-RateLimit-Limit-Requests, X-RateLimit-Limit-Concurrency
Request sizeup to 8 MB per request
Subscriptionno monthly fee; paid credit does not expire
Free trial$0.50 for 7 days, no card
How you paypay as you go from prepaid credit; nothing is charged for failed or refused requests
Volume bonus+5% from $50, +10% from $100
Priceinput $0.25 / 1M tokens, output $1.00 / 1M tokens
Top-upcrypto: USDT on TRON or USDC on Base, $10–$500, any whole sum
Keysone key per account, regenerate any time (the old one stops working)
Content policyadult content allowed; sexual content involving minors is refused
Sign-inGoogle or e-mail and password

Errors and what to do

Every error is JSON with a type you can switch on. You are never charged for an error.

CodeTypeMeaning
400bad_requestmalformed request or too long for the context window
401missing_key · invalid_key · key_revokedcheck the Authorization header or use your current key
402no_creditout of credit; add credit and retry
403content_blockedrefused by the content policy
404not_foundonly /v1/chat/completions and /v1/models exist
413request_too_largebody over 8 MB
429rate_limited · concurrencyslow down: rate or parallel limit reached
503upstream_busymodel busy — retry in a few seconds

Questions and answers

Is the model really uncensored?

The model does not refuse lawful adult, creative, or controversial topics. It is tuned to answer directly without the typical corporate guardrails. However, it always blocks sexual content involving minors, which is a hard limit.

Can I use this with the OpenAI SDK?

Yes. The API is fully OpenAI-compatible. You only need to change the base URL and API key in your existing client configuration. It supports standard endpoints like <code>/v1/chat/completions</code> and streaming via SSE.

How does pricing work?

You pay as you go with prepaid credit. Input tokens cost $0.25 per million, and output tokens cost $1.00 per million. There are no monthly fees or subscriptions. You can top up with crypto (USDT or USDC), starting from $10.

Your key is one form away

Create an account, copy the key, change the base URL. That is the whole setup.

Get API key