EN ▾

Native compatibility, extreme stability

Unfiltered LLM API Gateway

A single-model API gateway built for developers, providing uncensored LLM services. Based on self-built GPU servers, ensuring low latency and high availability, making integration easier.

  • OpenAI-compatible interface
  • 100k context window
  • Pay as you go, no monthly fee

Get API KeyRead Docs

Try it in one click

curl https://api.llmzhongzhuan.com/v1/chat/completions \
  -H "Authorization: Bearer $API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "uncensored",
    "messages": [{"role": "user", "content": "Write a blunt product review of a cheap VPN."}]
  }'
Input token price per 1M
$0.25
Output token / 1M
$1.00
Token context
100,000
Free trial credit
$0.50
Requests per minute
300

Features

  1. Single model, no aggregation latency

    We do not aggregate multiple models; we focus on a single high-performance uncensored model. This eliminates routing unpredictability and ensures consistent, low-latency responses for every request.

  2. Native OpenAI compatibility

    Fully compatible with the /v1/chat/completions interface, supporting streaming and tool calling. No code changes needed; just replace the base_url to connect to existing OpenAI SDKs.

  3. Transparent pay as you go

    Input $0.25/1M tokens, output $1.00/1M tokens. No subscription fees, no hidden costs. Prepaid credit never expires and is used up as consumed.

  4. 100k long context support

    Supports a 100,000 token context window (including prompt and completion), meeting the needs of long document processing and complex logical reasoning.

  5. Clear privacy and limits

    Prompts are not used for training. The only hard limit is child sexual abuse material. All other legal adult, fictional, or controversial topics are uncensored.

  6. Minimalist developer experience

    Register with just an email, no card required. Get $0.50 free trial credit upon registration, valid for 7 days. API key is generated instantly and can be reset at any time.

Usage process

  1. Register account

    Register on the "Get API Key" page using your email and password. No credit card binding required.

  2. Get Key

    Immediately receive your API key after successful registration, ready for code configuration or third-party tools.

  3. Make a request

    Point the base_url to https://api.llmzhongzhuan.com/v1,使用 model id "uncensored" to send a request.

Application scenarios

  1. LLM application integration

    Developers can seamlessly integrate this API into existing LLM applications, leveraging its uncensored nature for more free-form answers. Compatible with standard OpenAI protocols, with very low switching costs.

  2. Content generation and creation

    Suitable for scenarios requiring adult content, controversial topics, or breaking common AI hallucinations. A single model ensures consistent output style, ideal for batch content generation tasks.

  3. Long text context processing

    Use the 100k context window to process long documents, codebases, or complex conversation histories. Ideal for RAG or summarization tasks requiring full context understanding.

  4. Development and testing environments

    New users get $0.50 free trial credit, perfect for quickly verifying API integration logic. Test streaming and tool calling without any upfront payment.

Why choose our API gateway service

Many API aggregation services offer diversity by routing multiple models, but this often leads to latency fluctuations and unpredictable output quality. Our API gateway takes the opposite approach: focusing on a single, fine-tuned uncensored LLM. This means you don't need to switch between models; instead, you get consistent, low-latency responses. For developers needing stability, this single-model architecture eliminates the complexity of an aggregation layer.

We self-host the model on our own GPU servers, ensuring efficient data transmission. Whether generating creative content or performing logical reasoning, you get the raw, unfiltered model output. This transparency lets you know exactly what you are calling, rather than relying on black-box routing.

Native OpenAI-compatible interface

Our API fully follows the OpenAI /v1/chat/completions standard. This means you can use any OpenAI-compatible SDK (such as Python, Node.js, or curl) to connect directly. Just point the base_url to https://api.llmzhongzhuan.com/v1 and set the model ID to uncensored.

Supports streaming (SSE) and function calling to meet the needs of modern AI applications. Note that we provide text generation only; we do not support images, audio, video, or embedding vectors. This is a pure, efficient text API proxy.

Flexible pay as you go model

We ditched expensive monthly subscriptions for a transparent pay as you go model. Input tokens are $0.25/1M, output tokens are $1.00/1M. Prepaid credit never expires. The top up threshold is low, starting at $10 with cryptocurrency (USDT, USDC) payments; top up $50 for a 5% bonus, $100 for a 10% bonus.

New users get $0.50 free trial credit upon registration, no card required, valid for 7 days. This allows you to test API stability and response time at zero cost. All credit is used to offset token consumption with no hidden fees.

Who it is for, who it is not for

Our uncensored AI API is for developers who need free output, do not mind a single model, and value latency stability. If you need multi-model routing, image generation, or embedding vectors, this is not for you. We are also not for scenarios requiring enterprise-grade SLA guarantees (such as a 99.99% availability commitment) or specific industry compliance certifications (such as HIPAA). We are an independent service, not affiliated with any major tech company.

Additionally, we support text generation only; we do not support OCR, search, or fine-tuning. If your application relies on these features, evaluate whether you need to combine them with other specialized APIs. We focus on providing pure, unrestricted text generation capabilities.

Frequently asked questions

Is this API a proxy for GPT or Claude?

No. We run a self-hosted open-weight model, fine-tuned to be uncensored. It is not a GPT, Claude, Gemini, or any third-party vendor model. We provide an independent API proxy service that does not rely on any external vendor model instances.

What are the API rate limits?

Each API key is limited to 300 requests per minute, with a request body size of no more than 8 MB. Each account can have only one key, but it can be reset at any time, and the old key will become invalid immediately. This is sufficient for most developers and small to medium-sized applications.

How do I top up after the free trial credit is used up?

You can top up starting from $10 via cryptocurrency (USDT, USDC). The top-up amount is credited to your account and never expires. Top up $50 to get a 5% bonus, top up $100 to get a 10% bonus. No monthly subscription fees.

Will my conversation data be used for training?

No. We only collect necessary account information (email and password) to manage your account. Your prompts and responses are not used to train the model. We respect user privacy and do not use data for other purposes.

What does 'uncensored' mean? What are the exceptions?

'Uncensored' means the model will not reject legal adult content, fictional stories, controversial topics, or safety research. The only hard limit is child sexual abuse material, which will be blocked. We encourage legal use but do not restrict other adult content.

Just fill out the form to get your key

Create an account, copy the key, and modify the Base URL. Configuration is that simple.

Get API key