EN ▾

Quickstart: Use Unzensiertes LLM API

Use the uncensored AI via an OpenAI-compatible API interface. This quickstart guide shows you how to request text, stream, and integrate tools in minutes.

https://api.unzensiertesllm.com/v1uncensored

Base URL and Authentication

unzensiertesllm.com's API follows established standards to simplify integration. All requests are sent to the base URL https://api.unzensiertesllm.com/v1. You need an API key, which you receive immediately after registration on the "Get API key" page. This key is passed in the Authorization: Bearer <IHR_SCHLÜSSEL> header. If you use the official SDKs, you can configure the base URL and API key in the OPENAI_API_KEY and OPENAI_BASE_URL environment variables. The model you call has the ID uncensored. It is a standalone, specially tuned model that does not belong to well-known major providers like GPT or Claude.

Send First Request

To test the functionality, send a simple chat request. The interface expects a JSON object with the role user and a message. The response type is text. A successful request returns a response in the standard format. If the API key is invalid, you get a 401 error. If your credit is exhausted, the API reports 402. For your first use, you receive $0.50 in trial credit, valid for seven days.

curl https://api.unzensiertesllm.com/v1/chat/completions \
  -H "Authorization: Bearer $API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "uncensored",
    "messages": [{"role": "user", "content": "Write a blunt product review of a cheap VPN."}]
  }'

Use Python SDK

Integration into Python projects is directly possible via the official openai package. Since the API is OpenAI-compatible, you can use the familiar syntax. Simply adjust the base_url and set the API key. The uncensored model processes inputs without the typical content filters of other providers, making it ideal for creative or adult content. Make sure to keep the library up to date to ensure compatibility.

from openai import OpenAI

client = OpenAI(base_url="https://api.unzensiertesllm.com/v1", api_key="YOUR_KEY")

resp = client.chat.completions.create(
    model="uncensored",
    messages=[{"role": "user", "content": "Summarise this thread without softening it."}],
)
print(resp.choices[0].message.content)

Node.js SDK Integration

Developers working with JavaScript or TypeScript can use the openai NPM package. Configure the client instance with the base URL of our infrastructure. This enables seamless integration into existing web applications or backend services. The response is returned as a stream or as a complete object, depending on the configuration. The uncensored model also supports tool calls, facilitating use in complex workflows. Ensure your Node.js environment is modern to fully utilize all features.

import OpenAI from "openai";

const client = new OpenAI({ baseURL: "https://api.unzensiertesllm.com/v1", apiKey: process.env.API_KEY });

const resp = await client.chat.completions.create({
  model: "uncensored",
  messages: [{ role: "user", content: "Draft a villain monologue for my game." }],
});
console.log(resp.choices[0].message.content);

Enable Streaming (SSE)

For real-time applications, streaming responses is essential. Set the parameter stream: true in your request. The API then sends Server-Sent Events (SSE) containing chunks of the generated text. This significantly improves perceived latency for end users. You can render the text incrementally while the model is still thinking. This is particularly useful for chat interfaces that require immediate feedback without waiting for full generation.

stream = client.chat.completions.create(
    model="uncensored",
    messages=[{"role": "user", "content": "Tell the story in second person."}],
    stream=True,
)
for chunk in stream:
    if chunk.choices and chunk.choices[0].delta.content:
        print(chunk.choices[0].delta.content, end="", flush=True)

List Models and Limits

Use the endpoint GET /v1/models to query available models. You will see the model uncensored with its details. Pay attention to the rate limits: 300 requests per minute are allowed per key. The request body must not exceed 8 MB. The context window is 100,000 tokens (prompt plus response). If exceeded, you receive error codes such as 429. The API does not offer embeddings or image generation. It focuses purely on text completion. Prepaid credit never expires, and you can top up starting from $10.

Capabilities and limits

A quick checklist for developers: format, limits, features, billing.

FeatureSupport
API formatOpenAI-compatible: any OpenAI SDK or client works — change the base URL and the key
Modeluncensored
API keyBearer token in the Authorization header
MethodsPOST /v1/chat/completions · GET /v1/models
Base URLhttps://api.unzensiertesllm.com/v1
Structured outputresponse_format: {"type": "json_object"}
Completion lengthprompt + completion fit within 100,000 tokens; max_tokens optional, no separate output cap
Other parameterstemperature, top_p, stop, seed, presence_penalty, frequency_penalty
StreamingSupported (stream: true), usage included at the end
Context window100,000 tokens, input and output combined
Tools / tool callsYes — tools, tool_choice; replies carry tool_calls, also when streaming; send results back as role: tool
Request sizeup to 8 MB per request
Response headersX-Request-Id, X-Balance-USD, X-RateLimit-Limit-Requests, X-RateLimit-Limit-Concurrency
Rate limit300 requests per minute per key
Parallel requests8 requests at the same time per key
Trial credit$0.50 of credit valid 7 days, no card needed · Trial key: 2 parallel requests, 60 req/min; full limits (8 and 300) after first top-up
Token pricesinput $0.25 / 1M tokens, output $1.00 / 1M tokens
How you paypay as you go from prepaid credit; nothing is charged for failed or refused requests
Top-upcrypto: USDT on TRON or USDC on Base, $10–$500, any whole sum
Bonus credit+5% on $50+, +10% on $100+
Credit expiryno monthly fee; paid credit does not expire
Content policyadult content allowed; sexual content involving minors is refused
Accountsign in with Google or with e-mail + password
Keysone active key per account; a new key replaces the old one

HTTP errors

The type field is stable, the message is for humans. Errors cost nothing.

HTTPTypeWhat to do
400bad_requestinvalid JSON, empty messages, bad parameter, or prompt + max_tokens over the window — fix and resend
401missing_key · invalid_key · key_revokedno key, wrong key, or a key replaced by a newer one
402no_creditout of credit; add credit and retry
403content_blockedsexual content involving minors — refused, not billed
404not_foundonly /v1/chat/completions and /v1/models exist
413request_too_largebody over 8 MB
429rate_limited · concurrencyover 300/min or 8 parallel — back off and retry
503upstream_busymodel busy — retry in a few seconds

Frequently Asked Questions

How much does it cost to use the API?

You pay $0.25 per million input tokens and $1.00 per million output tokens. There are no monthly fees. You top up prepaid credit that never expires. Top-ups from $50 include bonus credit.

Is the API really uncensored?

Yes, the <code>uncensored</code> model provides answers without the typical filters for adult or controversial topics. However, there is a hard limit: sexual content involving minors is always blocked. This is the only fixed content restriction.

What error codes should I expect?

401 means an invalid API key. 402 signals that your credit is exhausted. 429 occurs when you exceed the limit of 300 requests per minute. Also note the 8 MB limit for the request body.

Your key is just one click away

Create account, copy key, adjust base URL. That's it.

Get API key