Uncensored AI Model For Coding: API Documentation
Get started with the uncensored ai model for coding by authenticating with your API key and sending requests to our OpenAI-compatible endpoint. Use the code samples below to integrate the uncensored llm into your existing workflows immediately.
https://api.uncensoredcodingai.com/v1uncensored
Base URL & Authentication
Our API follows the OpenAI chat-completions structure, allowing you to use existing SDKs with minimal configuration. The base URL is https://api.uncensoredcodingai.com/v1. Authentication is handled via a Bearer token in the Authorization header. You must include your API key exactly as provided during signup; there are no subscriptions or monthly fees, only pay-as-you-go prepaid credit.
Important: While the model is highly permissive, it does not remove all filters. A hard content limit always applies: requests containing sexual content involving minors will be blocked. All other lawful adult, controversial, or technical topics are processed without refusal.
First Request
Send a standard chat completion request to generate text. The model ID is uncensored. This is an uncensored coding llm optimized for code generation and technical tasks, not a clone of GPT or Claude. The model runs on our own GPU servers and is an open-weight model tuned for direct responses.
curl https://api.uncensoredcodingai.com/v1/chat/completions \
-H "Authorization: Bearer $API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "uncensored",
"messages": [{"role": "user", "content": "Write a blunt product review of a cheap VPN."}]
}'Replace YOUR_API_KEY with your actual key. The response will contain the generated text in the choices[0].message.content field. Ensure your input fits within the 100,000 token context window.
Python SDK Integration
Use the official openai Python package to interact with the API. Initialize the client with the custom base URL and your API key. This approach lets you leverage existing codebases that already support OpenAI-compatible endpoints.
from openai import OpenAI
client = OpenAI(base_url="https://api.uncensoredcodingai.com/v1", api_key="YOUR_KEY")
resp = client.chat.completions.create(
model="uncensored",
messages=[{"role": "user", "content": "Summarise this thread without softening it."}],
)
print(resp.choices[0].message.content)This script sends a prompt to the uncensored ai model and prints the response. The model ID is set to uncensored. You can adjust the temperature and max_tokens parameters to control creativity and output length. Remember that the uncensored llm does not have a memory of previous conversations unless you pass the full history in the messages array.
Node SDK Integration
For JavaScript and TypeScript projects, use the openai Node package. Configure the client with the baseUrl and apiKey environment variables or direct values.
import OpenAI from "openai";
const client = new OpenAI({ baseURL: "https://api.uncensoredcodingai.com/v1", apiKey: process.env.API_KEY });
const resp = await client.chat.completions.create({
model: "uncensored",
messages: [{ role: "user", content: "Draft a villain monologue for my game." }],
});
console.log(resp.choices[0].message.content);This example creates a chat completion request using the uncensored model. The response object contains the generated text. You can handle errors and status codes directly within your Node.js application logic. This method is ideal for backend services or serverless functions that need to generate code or text on demand.
Streaming Responses
Enable streaming by setting stream: true in your request. The API returns a Server-Sent Events (SSE) stream, allowing you to display output as it is generated. This is useful for large code blocks or lengthy explanations.
stream = client.chat.completions.create(
model="uncensored",
messages=[{"role": "user", "content": "Tell the story in second person."}],
stream=True,
)
for chunk in stream:
if chunk.choices and chunk.choices[0].delta.content:
print(chunk.choices[0].delta.content, end="", flush=True)Process the stream chunks as they arrive. Each chunk contains partial text that you can append to your UI or log. Streaming reduces the perceived latency for the user, as they see progress immediately. The total response time still depends on the model's inference speed and the length of the output.
Limits, Errors & Context
Your API key is limited to 300 requests per minute. If you exceed this, you will receive a 429 Too Many Requests error. The request body size is capped at 8 MB. Authentication errors return a 401 status if the key is invalid. Billing errors return a 402 status if your prepaid credit is exhausted; top up from $10 to restore service.
The model supports a 100k context window (prompt + completion). Ensure your inputs fit within this limit. If you need more context, you may need to manage truncation in your application logic. The uncensored ai model for coding is designed for reliability, so monitor your usage to avoid interruptions.
Under the hood: specs
Use this table to decide whether the API fits your project before you buy credit.
| Item | Value |
|---|---|
| Compatibility | OpenAI Chat Completions schema; official openai SDKs work unchanged |
| Authentication | Authorization: Bearer YOUR_KEY |
| Base URL | https://api.uncensoredcodingai.com/v1 |
| Methods | POST /v1/chat/completions · GET /v1/models |
| Model | uncensored |
| Sampling parameters | temperature, top_p, stop, seed, presence_penalty, frequency_penalty |
| Streaming | Supported (stream: true), usage included at the end |
| JSON mode | response_format: {"type": "json_object"} |
| Function calling | Yes — tools, tool_choice; replies carry tool_calls, also when streaming; send results back as role: tool |
| Context window | 100,000 tokens (prompt + completion together) |
| Completion length | 16,000 tokens max; 2,048 if max_tokens is not set |
| Response headers | X-Request-Id, X-Balance-USD, X-RateLimit-Limit-Requests, X-RateLimit-Limit-Concurrency |
| Rate limit | 300/min per key |
| Parallel requests | 8 requests at the same time per key |
| Request size | 8 MB request body |
| Top-up | USDT (TRC20) or USDC (Base), any whole amount from $10 to $500 |
| Volume bonus | +5% on $50+, +10% on $100+ |
| Price | input $0.25 / 1M tokens, output $1.00 / 1M tokens |
| How you pay | pay as you go from prepaid credit; nothing is charged for failed or refused requests |
| Trial credit | $0.50 of credit valid 7 days, no card needed |
| Subscription | paid credit never expires, no subscription |
| Content | adult content allowed; sexual content involving minors is refused |
| Sign-in | Google or e-mail and password |
| Keys | one key per account, regenerate any time (the old one stops working) |
Error reference
Errors come back as JSON with a stable type; failed and refused requests are not billed.
| HTTP | Type | What to do |
|---|---|---|
400 | bad_request | invalid JSON, empty messages, bad parameter, or prompt + max_tokens over the window — fix and resend |
401 | missing_key · invalid_key · key_revoked | check the Authorization header or use your current key |
402 | no_credit | balance is empty — top up, requests resume at once |
403 | content_blocked | sexual content involving minors — refused, not billed |
404 | not_found | unknown endpoint |
413 | request_too_large | request body larger than 8 MB |
429 | rate_limited · concurrency | slow down: rate or parallel limit reached |
503 | upstream_busy | model busy — retry in a few seconds |
Questions and answers
Is this model the same as GPT or Claude?
No. The uncensored ai model for coding is an open-weight model run on our own servers. It is tuned for direct responses without the guardrails typical of GPT, Claude, or Gemini. It is not a resell of another vendor's model.
What happens if I run out of credit?
Requests will return a 402 error. You can top up your account with credit from $10 upwards using crypto (USDT or USDC). Bonus credit is available for larger top-ups, and your credit never expires.
Does the model filter all content?
No, it is uncensored for lawful adult, controversial, and technical topics. However, a hard content limit always applies: sexual content involving minors is blocked. All other topics are processed without refusal.
Your key is one form away
Create an account, copy the key, change the base URL. That is the whole setup.