Get API key

OpenAI-compatible uncensored API

Kimi API: an independent guide and a drop-in alternative

A drop-in, OpenAI-compatible API for the uncensored model. Replace your base URL and API key to start generating text without content refusals.

  • Standard OpenAI SDK support
  • Single dedicated uncensored model
  • No subscriptions, pay per token

Get API keyRead the docs

Try it in one request

curl https://api.kimiapi.cc/v1/chat/completions \
  -H "Authorization: Bearer $API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "uncensored",
    "messages": [{"role": "user", "content": "Write a blunt product review of a cheap VPN."}]
  }'

What you get

Drop-in Compatibility

Works with the official OpenAI SDKs and any compatible client by simply changing the base URL and API key.

Streaming & Tools

Supports streaming via SSE, function calling with tools, and JSON mode for structured responses.

Transparent Pricing

Pay only for real token usage with prepaid crypto credit that never expires.

Uncensored Output

The model answers without refusals for lawful adult, fictional, or controversial topics.

Fast Setup

Get your API key instantly with Google or email. No phone number required.

Privacy First

Prompts are not used for training your model data.

How it works

Get Your Key

Sign up with Google or email to receive your API key instantly.

Update Configuration

Set your base URL to https://api.kimiapi.cc/v1 and add your key.

Start Generating

Send requests to /v1/chat/completions and receive uncensored text output.

What people build with it

Content Generation

Generate creative stories or articles without hitting standard content filters. Ideal for adult themes or unconventional narratives.

Security Research

Prompt the model to find its own limits or generate examples of restricted topics for analysis.

Data Extraction

Use JSON mode to extract structured data from text without format interruptions from moderation layers.

Chat Applications

Build chat interfaces that respond naturally to diverse user inputs without unexpected refusals.

Why Choose an Uncensored Kimi API Alternative

When developers search for a kimi api or similar large language model endpoints, they often encounter aggregators that route traffic through multiple providers. This adds latency and complexity. We provide a single, dedicated uncensored model via a standard OpenAI-compatible endpoint. This ensures predictable performance and exact prompt fidelity.

Our service is designed for users who want reliable text generation without content refusals. The model is open-weight and tuned to answer a wide range of topics, including adult or controversial subjects, for lawful use. It is not GPT, Claude, or any other vendor's model. It is a distinct, independent LLM running on our servers.

We avoid the proxy overhead of aggregators. You get direct access to the model's capabilities. This is particularly useful for applications where consistency matters more than accessing a vast library of different models.

OpenAI-Compatible: Drop-in Code

Our API follows the OpenAI Chat Completions standard. You can use the same client libraries you already know. This reduces integration time significantly. You only need to update two configuration values: the base URL and the API key.

The endpoint is POST /v1/chat/completions. It accepts the same parameters as the OpenAI API, including temperature, top_p, and stop sequences. We also support function calling and JSON mode for structured outputs.

from openai import OpenAI

client = OpenAI(base_url="https://api.kimiapi.cc/v1", api_key="YOUR_KEY")

resp = client.chat.completions.create(
    model="uncensored",
    messages=[{"role": "user", "content": "Summarise this thread without softening it."}],
)
print(resp.choices[0].message.content)

This approach means you can switch your existing OpenAI clients to our uncensored model with minimal code changes. The response format is identical, making debugging and testing straightforward.

Model: Uncensored Open-Weight LLM

The model ID is uncensored. It is an open-weight large language model hosted on our infrastructure. It is tuned to answer without content refusals for lawful adult use. It does not block standard creative or controversial topics.

However, there is a hard content limit: requests involving sexual content with minors are always refused. This is the only strict block enforced by the model.

The context window is 64,000 tokens for the combined prompt and completion. You can specify a maximum output of 16,000 tokens per request. If you do not set max_tokens, the limit defaults to 2,048 tokens.

This is not a routing service. We do not switch between different vendor models. You get consistent behavior from this single, dedicated model.

Pricing: Pay Per Token, No Subscriptions

We charge $0.25 per 1M input tokens and $1.00 per 1M output tokens. Credit is prepaid and charged by real token usage. Errors and refusals are free, so you only pay for successful completions.

There are no monthly subscriptions or hidden fees. Your paid credit never expires. You can top up with cryptocurrency only: USDT (TRC20) or USDC (Base). Minimum top-up is $10, maximum is $500.

We offer a bonus on top-ups: +5% credit for amounts of $50 or more, and +10% for $100 or more. We do not accept cards, PayPal, or bank transfers. Every new account gets $0.50 of trial credit valid for 7 days, no card needed.

Features: Streaming, Tools, JSON Mode

Our API supports streaming responses via Server-Sent Events (SSE). Token usage information is included in the last chunk of the stream. This allows you to display responses in real-time to your users.

Function calling is fully supported. You can define tools and let the model decide when to use them. This is essential for building agents that can interact with external systems.

JSON mode is available via the response_format parameter. Set it to json_object to force the model to output valid JSON. This is useful for data extraction or structured content generation.

We support standard parameters like temperature, top_p, stop, seed, presence_penalty, and frequency_penalty for fine-tuning the output.

Get Your API Key Now

Sign up using Continue with Google or create an account with email and password. Your API key is shown immediately. No phone number is required.

Each account is limited to one active key. If you generate a new key, the old one is replaced. You can make up to 300 requests per minute per key, with a concurrency limit of 8 simultaneous requests. The request body size limit is 8 MB.

If you encounter issues, such as a double charge, use the Support page. Credit is not refunded as it never expires, but we will fix accounting errors.

curl https://api.kimiapi.cc/v1/chat/completions \
  -H "Authorization: Bearer $API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "uncensored",
    "messages": [{"role": "user", "content": "Write a blunt product review of a cheap VPN."}]
  }'

This is an API for developers. We do not offer a chat website or app. We do not provide SLAs, SOC2, or HIPAA certifications. We do not offer on-prem deployment. If you need those, this is not the right choice for you.

Questions and answers

Is this the official Kimi API?

No, we are an independent service. We are not affiliated with the Kimi brand or its creators. We provide an OpenAI-compatible endpoint for our own uncensored model. If you are looking for the official Kimi API, check Moonshot AI's documentation.

What is the context window size?

The context window is 64,000 tokens, which includes both the prompt and the completion. You can specify a maximum output of 16,000 tokens per request. If you do not set max_tokens, the limit is 2,048 tokens.

How do I pay for the API?

We accept crypto only: USDT (TRC20) or USDC (Base). You can top up any whole amount from $10 to $500. You receive a bonus credit for larger top-ups. We do not accept credit cards, PayPal, or bank transfers.

Does the API support streaming?

Yes, we support streaming via Server-Sent Events (SSE). Token usage information is provided in the last chunk of the stream. This allows you to display responses in real-time.

Is my data used for training?

No, your prompts are not used for training. We only need your email to create your account. You can sign up with Google or email and password. No phone number is required.

Your key is one form away

Create an account, copy the key, change the base URL. That is the whole setup.

Get API key