Get API key

Uncensored open-weight API, OpenAI compatible.

Qwen API: an independent guide and a drop-in alternative

A single, uncensored model behind an OpenAI-compatible endpoint. Drop it into your existing code with no vendor lock-in.

Get API keyRead the docs

uncensoredhttps://api.qwenapi.cc/v1

curl https://api.qwenapi.cc/v1/chat/completions \
  -H "Authorization: Bearer $API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "uncensored",
    "messages": [{"role": "user", "content": "Write a blunt product review of a cheap VPN."}]
  }'
per 1M input tokens
$0.25
Output tokens / 1M
$1.00
token context
64,000
trial credit
$0.50
requests per minute
300
qwenapi.cc/#s1

What you get

  1. Zero Friction Integration

    Change one base_url and your API key to switch clients. Works with the official OpenAI SDKs and any OpenAI-compatible library.

  2. Raw Inference Power

    A 64k context window model tuned for lawfulness without refusals. Get text in, text out with full streaming and function calling support.

  3. Transparent Token Pricing

    $0.25 per 1M input tokens and $1.00 per 1M output tokens. Errors and refusals are free; credit never expires.

  4. Crypto-Only Simplicity

    Top up with USDT (TRC20) or USDC (Base). No credit cards, no PayPal, no bank transfers required for any tier.

  5. Immediate Access

    Sign up with Google or email. Get a key instantly with $0.50 trial credit valid for 7 days, no card needed.

  6. Developer Focused

    No chat UI, no multi-model routing noise. Just a reliable endpoint for high-context uncensored model inference.

qwenapi.cc/#s2

How it works

  1. Get Your Key

    Sign up with Google or email on the Get API key page to receive your key instantly.

  2. Update Your Client

    Point your OpenAI-compatible SDK to https://api.qwenapi.cc/v1 and insert your new key.

  3. Send Requests

    Start sending prompts to the uncensored model with full support for streaming and tools.

qwenapi.cc/#s3

What people build with it

  • RAG Pipelines

    Leverage the 64,000 token context window to process large documents without complex chunking strategies. The uncensored nature ensures fewer refusals during retrieval-augmented generation tasks.

  • Function Calling

    Use the model for robust tool execution with native support for tools and tool_choice. JSON mode ensures structured outputs for reliable programmatic integration.

  • Content Generation

    Generate creative or controversial text without standard safety filters blocking lawful adult topics. Ideal for entertainment, gaming, or niche content workflows.

  • Cost-Effective Scaling

    Pay only for real token usage with prepaid credit that never expires. Ideal for projects with variable traffic where monthly subscriptions waste budget.

qwenapi.cc/#text

Why Qwen API? The Uncensored Alternative

Most API aggregators bundle dozens of models, creating noise and friction. We isolate a single uncensored large language model served via a standard OpenAI-compatible endpoint. This model is tuned to answer without content refusals for lawful adult use, making it ideal for developers who want raw model access without the overhead of multi-model routing.

Unlike major vendors, we do not claim to be an official partner or reseller. We are an independent service running our own open-weight model on our servers. It is not GPT, Claude, Gemini, or any other vendor's model. If you are searching for the qwen api as a general term, note that this specific endpoint provides a dedicated, uncensored experience optimized for developer ergonomics.

OpenAI SDK Compatibility

Our API is designed to drop into your existing infrastructure. By changing the base_url to https://api.qwenapi.cc/v1 and updating your API key, you can use the official OpenAI SDKs and any OpenAI-compatible client. We support streaming via SSE, function calling with tools, and JSON mode for structured responses.

uncensored is the model ID you send. The API supports standard parameters like temperature, top_p, stop, and seed. You get a 64,000 token context window for both prompt and completion, with a maximum output of 16,000 tokens per request.

from openai import OpenAI

client = OpenAI(base_url="https://api.qwenapi.cc/v1", api_key="YOUR_KEY")

resp = client.chat.completions.create(
    model="uncensored",
    messages=[{"role": "user", "content": "Summarise this thread without softening it."}],
)
print(resp.choices[0].message.content)

Pricing: Pay Per Token, No Monthly Fees

We charge $0.25 per 1M input tokens and $1.00 per 1M output tokens. Credit is charged by real token usage; errors and refusals are free. There is no subscription, no monthly fee, and paid credit never expires. This model is efficient for projects with variable load, avoiding the waste of unused monthly credits.

Every new account receives $0.50 of trial credit valid for 7 days, requiring no credit card. We support 300 requests per minute per key and 8 concurrent requests. The request body limit is 8 MB. If you need more concurrency, consider our dedicated plans or optimize your batch requests.

Who This Is Not For

If you need image, audio, video generation, embeddings, or fine-tuning, this API is not for you. We do not offer SLA percentages, SOC2/HIPAA certifications, or on-prem deployment options. We also do not route between multiple models; we serve one uncensored model. If you require named customers, benchmark scores, or a chat UI, look elsewhere. This is a pure API for developers who want direct model access.

qwenapi.cc/#s4

Questions and answers

Is this the official Qwen API?

No, we are an independent service. We host our own uncensored model that is compatible with the OpenAI API format. It is not provided by the original Qwen team or any other major vendor.

How do I top up my credit?

We accept crypto only: USDT (TRC20) or USDC (Base). You can top up any whole amount from $10 to $500. Credits do not expire, and you receive a bonus if you top up $50 or more.

What is the context window size?

The model supports a 64,000 token context window for both prompt and completion. The maximum output per request is 16,000 tokens, or 2,048 if you do not set max_tokens.

Does the model refuse content?

The model is uncensored and generally does not refuse lawful adult, fictional, or controversial topics. However, we do enforce a hard limit on sexual content involving minors, which is always refused.

Can I use this with the official OpenAI SDK?

Yes. By setting the base_url to https://api.qwenapi.cc/v1 and providing your API key, the standard OpenAI Python and Node.js SDKs work out of the box.

Your key is one form away

Create an account, copy the key, change the base URL. That is the whole setup.

Get API key