Quickstart: Use Unzensiertes LLM API
Use the uncensored AI via an OpenAI-compatible API interface. This quickstart guide shows you how to request text, stream, and integrate tools in minutes.
https://api.unzensiertesllm.com/v1uncensored
Base URL and Authentication
unzensiertesllm.com's API follows established standards to simplify integration. All requests are sent to the base URL https://api.unzensiertesllm.com/v1. You need an API key, which you receive immediately after registration on the "Get API key" page. This key is passed in the Authorization: Bearer <IHR_SCHLÜSSEL> header. If you use the official SDKs, you can configure the base URL and API key in the OPENAI_API_KEY and OPENAI_BASE_URL environment variables. The model you call has the ID uncensored. It is a standalone, specially tuned model that does not belong to well-known major providers like GPT or Claude.
Send First Request
To test the functionality, send a simple chat request. The interface expects a JSON object with the role user and a message. The response type is text. A successful request returns a response in the standard format. If the API key is invalid, you get a 401 error. If your credit is exhausted, the API reports 402. For your first use, you receive $0.50 in trial credit, valid for seven days.
curl https://api.unzensiertesllm.com/v1/chat/completions \
-H "Authorization: Bearer $API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "uncensored",
"messages": [{"role": "user", "content": "Write a blunt product review of a cheap VPN."}]
}'
Use Python SDK
Integration into Python projects is directly possible via the official openai package. Since the API is OpenAI-compatible, you can use the familiar syntax. Simply adjust the base_url and set the API key. The uncensored model processes inputs without the typical content filters of other providers, making it ideal for creative or adult content. Make sure to keep the library up to date to ensure compatibility.
from openai import OpenAI
client = OpenAI(base_url="https://api.unzensiertesllm.com/v1", api_key="YOUR_KEY")
resp = client.chat.completions.create(
model="uncensored",
messages=[{"role": "user", "content": "Summarise this thread without softening it."}],
)
print(resp.choices[0].message.content)
Node.js SDK Integration
Developers working with JavaScript or TypeScript can use the openai NPM package. Configure the client instance with the base URL of our infrastructure. This enables seamless integration into existing web applications or backend services. The response is returned as a stream or as a complete object, depending on the configuration. The uncensored model also supports tool calls, facilitating use in complex workflows. Ensure your Node.js environment is modern to fully utilize all features.
import OpenAI from "openai";
const client = new OpenAI({ baseURL: "https://api.unzensiertesllm.com/v1", apiKey: process.env.API_KEY });
const resp = await client.chat.completions.create({
model: "uncensored",
messages: [{ role: "user", content: "Draft a villain monologue for my game." }],
});
console.log(resp.choices[0].message.content);
Enable Streaming (SSE)
For real-time applications, streaming responses is essential. Set the parameter stream: true in your request. The API then sends Server-Sent Events (SSE) containing chunks of the generated text. This significantly improves perceived latency for end users. You can render the text incrementally while the model is still thinking. This is particularly useful for chat interfaces that require immediate feedback without waiting for full generation.
stream = client.chat.completions.create(
model="uncensored",
messages=[{"role": "user", "content": "Tell the story in second person."}],
stream=True,
)
for chunk in stream:
if chunk.choices and chunk.choices[0].delta.content:
print(chunk.choices[0].delta.content, end="", flush=True)
List Models and Limits
Use the endpoint GET /v1/models to query available models. You will see the model uncensored with its details. Pay attention to the rate limits: 300 requests per minute are allowed per key. The request body must not exceed 8 MB. The context window is 100,000 tokens (prompt plus response). If exceeded, you receive error codes such as 429. The API does not offer embeddings or image generation. It focuses purely on text completion. Prepaid credit never expires, and you can top up starting from $10.
Capabilities and limits
A quick checklist for developers: format, limits, features, billing.
| Feature | Support |
|---|---|
| API format | OpenAI-compatible: any OpenAI SDK or client works — change the base URL and the key |
| Model | uncensored |
| API key | Bearer token in the Authorization header |
| Methods | POST /v1/chat/completions · GET /v1/models |
| Base URL | https://api.unzensiertesllm.com/v1 |
| Structured output | response_format: {"type": "json_object"} |
| Completion length | prompt + completion fit within 100,000 tokens; max_tokens optional, no separate output cap |
| Other parameters | temperature, top_p, stop, seed, presence_penalty, frequency_penalty |
| Streaming | Supported (stream: true), usage included at the end |
| Context window | 100,000 tokens, input and output combined |
| Tools / tool calls | Yes — tools, tool_choice; replies carry tool_calls, also when streaming; send results back as role: tool |
| Request size | up to 8 MB per request |
| Response headers | X-Request-Id, X-Balance-USD, X-RateLimit-Limit-Requests, X-RateLimit-Limit-Concurrency |
| Rate limit | 300 requests per minute per key |
| Parallel requests | 8 requests at the same time per key |
| Trial credit | $0.50 of credit valid 7 days, no card needed · Trial key: 2 parallel requests, 60 req/min; full limits (8 and 300) after first top-up |
| Token prices | input $0.25 / 1M tokens, output $1.00 / 1M tokens |
| How you pay | pay as you go from prepaid credit; nothing is charged for failed or refused requests |
| Top-up | crypto: USDT on TRON or USDC on Base, $10–$500, any whole sum |
| Bonus credit | +5% on $50+, +10% on $100+ |
| Credit expiry | no monthly fee; paid credit does not expire |
| Content policy | adult content allowed; sexual content involving minors is refused |
| Account | sign in with Google or with e-mail + password |
| Keys | one active key per account; a new key replaces the old one |
HTTP errors
The type field is stable, the message is for humans. Errors cost nothing.
| HTTP | Type | What to do |
|---|---|---|
400 | bad_request | invalid JSON, empty messages, bad parameter, or prompt + max_tokens over the window — fix and resend |
401 | missing_key · invalid_key · key_revoked | no key, wrong key, or a key replaced by a newer one |
402 | no_credit | out of credit; add credit and retry |
403 | content_blocked | sexual content involving minors — refused, not billed |
404 | not_found | only /v1/chat/completions and /v1/models exist |
413 | request_too_large | body over 8 MB |
429 | rate_limited · concurrency | over 300/min or 8 parallel — back off and retry |
503 | upstream_busy | model busy — retry in a few seconds |
Frequently Asked Questions
How much does it cost to use the API?
You pay $0.25 per million input tokens and $1.00 per million output tokens. There are no monthly fees. You top up prepaid credit that never expires. Top-ups from $50 include bonus credit.
Is the API really uncensored?
Yes, the <code>uncensored</code> model provides answers without the typical filters for adult or controversial topics. However, there is a hard limit: sexual content involving minors is always blocked. This is the only fixed content restriction.
What error codes should I expect?
401 means an invalid API key. 402 signals that your credit is exhausted. 429 occurs when you exceed the limit of 300 requests per minute. Also note the 8 MB limit for the request body.