OpenAI-compatible uncensored API
Kimi API: an independent guide and a drop-in alternative
A drop-in, OpenAI-compatible API for the uncensored model. Replace your base URL and API key to start generating text without content refusals.
- Standard OpenAI SDK support
- Single dedicated uncensored model
- No subscriptions, pay per token
Try it in one request
curl https://api.kimiapi.cc/v1/chat/completions \
-H "Authorization: Bearer $API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "uncensored",
"messages": [{"role": "user", "content": "Write a blunt product review of a cheap VPN."}]
}'from openai import OpenAI
client = OpenAI(base_url="https://api.kimiapi.cc/v1", api_key="YOUR_KEY")
resp = client.chat.completions.create(
model="uncensored",
messages=[{"role": "user", "content": "Summarise this thread without softening it."}],
)
print(resp.choices[0].message.content)import OpenAI from "openai";
const client = new OpenAI({ baseURL: "https://api.kimiapi.cc/v1", apiKey: process.env.API_KEY });
const resp = await client.chat.completions.create({
model: "uncensored",
messages: [{ role: "user", content: "Draft a villain monologue for my game." }],
});
console.log(resp.choices[0].message.content);- $0.25per 1M input tokens
- $1.00Output tokens / 1M
- 64,000token context
- $0.50trial credit
- 300requests per minute
What you get
Drop-in Compatibility
Works with the official OpenAI SDKs and any compatible client by simply changing the base URL and API key.
Streaming & Tools
Supports streaming via SSE, function calling with tools, and JSON mode for structured responses.
Transparent Pricing
Pay only for real token usage with prepaid crypto credit that never expires.
Uncensored Output
The model answers without refusals for lawful adult, fictional, or controversial topics.
Fast Setup
Get your API key instantly with Google or email. No phone number required.
Privacy First
Prompts are not used for training your model data.
How it works
Get Your Key
Sign up with Google or email to receive your API key instantly.
Update Configuration
Set your base URL to https://api.kimiapi.cc/v1 and add your key.
Start Generating
Send requests to /v1/chat/completions and receive uncensored text output.
What people build with it
Content Generation
Generate creative stories or articles without hitting standard content filters. Ideal for adult themes or unconventional narratives.
Security Research
Prompt the model to find its own limits or generate examples of restricted topics for analysis.
Data Extraction
Use JSON mode to extract structured data from text without format interruptions from moderation layers.
Chat Applications
Build chat interfaces that respond naturally to diverse user inputs without unexpected refusals.
Why Choose an Uncensored Kimi API Alternative
When developers search for a kimi api or similar large language model endpoints, they often encounter aggregators that route traffic through multiple providers. This adds latency and complexity. We provide a single, dedicated uncensored model via a standard OpenAI-compatible endpoint. This ensures predictable performance and exact prompt fidelity.
Our service is designed for users who want reliable text generation without content refusals. The model is open-weight and tuned to answer a wide range of topics, including adult or controversial subjects, for lawful use. It is not GPT, Claude, or any other vendor's model. It is a distinct, independent LLM running on our servers.
We avoid the proxy overhead of aggregators. You get direct access to the model's capabilities. This is particularly useful for applications where consistency matters more than accessing a vast library of different models.
OpenAI-Compatible: Drop-in Code
Our API follows the OpenAI Chat Completions standard. You can use the same client libraries you already know. This reduces integration time significantly. You only need to update two configuration values: the base URL and the API key.
The endpoint is POST /v1/chat/completions. It accepts the same parameters as the OpenAI API, including temperature, top_p, and stop sequences. We also support function calling and JSON mode for structured outputs.
from openai import OpenAI
client = OpenAI(base_url="https://api.kimiapi.cc/v1", api_key="YOUR_KEY")
resp = client.chat.completions.create(
model="uncensored",
messages=[{"role": "user", "content": "Summarise this thread without softening it."}],
)
print(resp.choices[0].message.content)This approach means you can switch your existing OpenAI clients to our uncensored model with minimal code changes. The response format is identical, making debugging and testing straightforward.
Model: Uncensored Open-Weight LLM
The model ID is uncensored. It is an open-weight large language model hosted on our infrastructure. It is tuned to answer without content refusals for lawful adult use. It does not block standard creative or controversial topics.
However, there is a hard content limit: requests involving sexual content with minors are always refused. This is the only strict block enforced by the model.
The context window is 64,000 tokens for the combined prompt and completion. You can specify a maximum output of 16,000 tokens per request. If you do not set max_tokens, the limit defaults to 2,048 tokens.
This is not a routing service. We do not switch between different vendor models. You get consistent behavior from this single, dedicated model.
Pricing: Pay Per Token, No Subscriptions
We charge $0.25 per 1M input tokens and $1.00 per 1M output tokens. Credit is prepaid and charged by real token usage. Errors and refusals are free, so you only pay for successful completions.
There are no monthly subscriptions or hidden fees. Your paid credit never expires. You can top up with cryptocurrency only: USDT (TRC20) or USDC (Base). Minimum top-up is $10, maximum is $500.
We offer a bonus on top-ups: +5% credit for amounts of $50 or more, and +10% for $100 or more. We do not accept cards, PayPal, or bank transfers. Every new account gets $0.50 of trial credit valid for 7 days, no card needed.
Features: Streaming, Tools, JSON Mode
Our API supports streaming responses via Server-Sent Events (SSE). Token usage information is included in the last chunk of the stream. This allows you to display responses in real-time to your users.
Function calling is fully supported. You can define tools and let the model decide when to use them. This is essential for building agents that can interact with external systems.
JSON mode is available via the response_format parameter. Set it to json_object to force the model to output valid JSON. This is useful for data extraction or structured content generation.
We support standard parameters like temperature, top_p, stop, seed, presence_penalty, and frequency_penalty for fine-tuning the output.
Get Your API Key Now
Sign up using Continue with Google or create an account with email and password. Your API key is shown immediately. No phone number is required.
Each account is limited to one active key. If you generate a new key, the old one is replaced. You can make up to 300 requests per minute per key, with a concurrency limit of 8 simultaneous requests. The request body size limit is 8 MB.
If you encounter issues, such as a double charge, use the Support page. Credit is not refunded as it never expires, but we will fix accounting errors.
curl https://api.kimiapi.cc/v1/chat/completions \
-H "Authorization: Bearer $API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "uncensored",
"messages": [{"role": "user", "content": "Write a blunt product review of a cheap VPN."}]
}'This is an API for developers. We do not offer a chat website or app. We do not provide SLAs, SOC2, or HIPAA certifications. We do not offer on-prem deployment. If you need those, this is not the right choice for you.
Questions and answers
Is this the official Kimi API?
No, we are an independent service. We are not affiliated with the Kimi brand or its creators. We provide an OpenAI-compatible endpoint for our own uncensored model. If you are looking for the official Kimi API, check Moonshot AI's documentation.
What is the context window size?
The context window is 64,000 tokens, which includes both the prompt and the completion. You can specify a maximum output of 16,000 tokens per request. If you do not set max_tokens, the limit is 2,048 tokens.
How do I pay for the API?
We accept crypto only: USDT (TRC20) or USDC (Base). You can top up any whole amount from $10 to $500. You receive a bonus credit for larger top-ups. We do not accept credit cards, PayPal, or bank transfers.
Does the API support streaming?
Yes, we support streaming via Server-Sent Events (SSE). Token usage information is provided in the last chunk of the stream. This allows you to display responses in real-time.
Is my data used for training?
No, your prompts are not used for training. We only need your email to create your account. You can sign up with Google or email and password. No phone number is required.
Your key is one form away
Create an account, copy the key, change the base URL. That is the whole setup.