Pure API Focus
We provide only the chat-completions endpoint for text generation. No embeddings, images, audio, or video generation to clutter your integration.
We provide only the chat-completions endpoint for text generation. No embeddings, images, audio, or video generation to clutter your integration.
Supports SSE streaming with token usage in the last chunk. Enforce strict JSON output with response_format or use function calling with tools.
Process up to 64,000 tokens per request (prompt + completion). Generate up to 16,000 tokens in a single response for long-form tasks.
Our open-weight model answers without content refusals for lawful adult use. It is not GPT, Claude, or any other vendor's model.
Top up with USDT (TRC20) or USDC (Base) from $10 to $500. Credit never expires and errors/refusals are charged zero tokens.
Sign up with Google or email to get your API key immediately. No phone number required, and every new account receives $0.50 trial credit.
Sign up via Google or email on the dashboard to generate your unique API key instantly.
Configure your client to point to https://api.uncensoredai.top/v1 and inject your API key.
Send POST requests to /v1/chat/completions with model id "uncensored" and start generating text.
Generate creative writing, stories, or marketing copy without standard corporate filters blocking nuanced or adult themes. Ideal for media that requires fewer restrictions.
Prompt LLMs to break out of jailbreak patterns or analyze model weaknesses without them refusing due to standard safety refusals. Get raw model outputs.
Use JSON mode to reliably extract structured data from unstructured text. The dedicated model ensures consistent formatting without unexpected refusal interruptions.
Build chatbots that maintain character and context over long sessions. The 64k context window allows for deep, continuous conversations without losing track.
The Uncensored AI API provides a single, dedicated endpoint for text generation. Unlike aggregators that route through multiple models, we serve one open-weight model tuned for minimal refusals. This ensures consistent behavior and predictable latency for your applications. You interact with a standard chat-completions interface that accepts text and returns text.
Our architecture is simple: POST /v1/chat/completions and GET /v1/models. There are no embeddings, no image generation, and no fine-tuning endpoints to manage. This reduction in complexity means fewer points of failure and a faster integration process. You get exactly what you ask for, without the overhead of a multi-model platform.
Because we follow the OpenAI API specification, you can use the official SDKs or any compatible client with minimal configuration. Simply update the base_url to https://api.uncensoredai.top/v1 and provide your API key. The model identifier is "uncensored". This means existing code for GPT-4 or GPT-3.5 often works with zero code changes, only requiring a URL update.
This compatibility extends to standard parameters like temperature, top_p, stop sequences, and seed. You can leverage existing tooling and wrappers built for OpenAI without rewriting your logic. Check the official OpenAI documentation for the full list of supported parameters, as our implementation mirrors their standard behavior.
We support streaming via Server-Sent Events (SSE), allowing you to receive tokens as they are generated. The final chunk includes the total token usage for accurate billing and monitoring. For structured data needs, use response_format set to json_object to enforce strict JSON output, which is critical for programmatic pipelines.
Function calling is fully supported, allowing you to define tools and let the model decide when to invoke them. You can control tool choice behavior and pass multiple arguments. This makes the uncensored llm api suitable for complex agent workflows where reliable tool execution is required without standard refusal filters interfering with valid tool use.
This API is for developers who want a straightforward, pay-as-you-go solution for uncensored text generation. If you need a chat interface, image generation, or embeddings, this is not the right product. We do not offer SLA percentages, SOC2/HIPAA/ISO certifications, or on-prem deployment. If you require enterprise-grade compliance or model routing between multiple vendors, look elsewhere. We are a focused tool for a specific need: raw, uncensored text output via a simple API.
The context window is 64,000 tokens, counting both the prompt and the completion together. You can generate up to 16,000 tokens per request. If you do not set max_tokens, the default output limit is 2,048 tokens.
We accept crypto only: USDT (TRC20) or USDC (Base). You can top up any whole amount between $10 and $500. Credit never expires, and you receive a +5% bonus for $50+ and +10% for $100+ tops.
Yes, every new account gets $0.50 of trial credit valid for 7 days. No credit card is needed to sign up. You can start generating immediately after verifying your email or Google account.
The model is tuned to answer without refusals for lawful adult use, including controversial or fictional topics. However, we have a hard limit: sexual content involving minors is always refused. Requests containing that content will be blocked.
You are limited to 300 requests per minute per key and 8 concurrent requests per key. The maximum request body size is 8 MB. Each account is limited to one active key; generating a new key replaces the previous one.
Create an account, copy the key, change the base URL. That is the whole setup.