Get API key

Code-First Uncensored API

NoFilter AI: The Uncensored LLM API for Code-First Integration

Drop our OpenAI-compatible endpoint into your code and generate uncensored text immediately. No content filters, no complex setup, just direct integration.

  • OpenAI-compatible endpoint
  • Zero content filters
  • Transparent token pricing

Try it in one request

curl https://api.nofilterai.cc/v1/chat/completions \
  -H "Authorization: Bearer $API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "uncensored",
    "messages": [{"role": "user", "content": "Write a blunt product review of a cheap VPN."}]
  }'

What you get

  1. Drop-in Compatibility

    Works with the official OpenAI SDKs and any OpenAI-compatible client by simply changing the base URL and API key.

  2. Advanced Control

    Supports streaming via SSE, JSON mode, function calling, and standard parameters like temperature, top_p, and seed.

  3. Generous Context

    Handles a 64,000-token context window with up to 16,000 tokens of output per request.

  4. Transparent Pricing

    Pay only for what you use: $0.25 per 1M input tokens and $1.00 per 1M output tokens.

  5. Crypto-First Top-ups

    Top up via USDT (TRC20) or USDC (Base) with no cards, no PayPal, and no monthly subscriptions.

  6. Instant Access

    Get your API key immediately after signing up with Google or email; no phone number required.

How it works

  1. Get Your Key

    Sign up with Google or email and retrieve your unique API key instantly.

  2. Configure Your Client

    Point your SDK to https://api.nofilterai.cc/v1 and inject your key.

  3. Start Generating

    Send requests to the uncensored model and receive text without content refusals.

What people build with it

  1. Adult-Ready Chatbots

    Build character-driven apps that handle romantic, mature, or explicit dialogue without unexpected content blocks or polite refusals.

  2. Creative Writing Tools

    Generate uncensored fiction, roleplay scenarios, or controversial topics where standard models might apply moral filters or tone down the narrative.

  3. Data Extraction

    Use JSON mode to reliably extract structured data from text without the model refusing to output due to minor content nuances.

  4. Security Research

    Test LLM boundaries and jailbreaks on a model tuned to answer without refusals for lawful adult use and research purposes.

Why Uncensored Matters for Developers

Standard LLMs are fine-tuned to be helpful, harmless, and honest, which often means they refuse lawful adult content, controversial opinions, or creative edge cases. NoFilter AI provides a single, uncensored endpoint optimized for developers who need their models to answer questions without unnecessary moralizing or content filters. This is not a chat interface; it is a raw API for direct integration. You control the context and the output, ensuring your application behaves exactly as you define it without hidden guardrails.

Our model is an open-weight model run on our own servers, tuned specifically to avoid refusals for lawful adult use. It is not GPT, Claude, Gemini, Grok, DeepSeek, Qwen, Llama, or any other vendor's model. When you need a text-in, text-out API that respects your application's logic over a generic safety profile, this is the tool for you.

OpenAI-Compatible Quick Start

Integration is straightforward because we follow the OpenAI chat-completions standard. You only need two endpoints: POST /v1/chat/completions for generation and GET /v1/models to verify availability. There are no embeddings, no image generation, no audio, and no video. Just text. This simplicity means you can drop our SDK into your existing codebase and start generating content immediately.

  • Base URL: https://api.nofilterai.cc/v1
  • Model ID: "uncensored"
  • SDKs: Works with the official OpenAI SDKs and any OpenAI-compatible client.

curl https://api.nofilterai.cc/v1/chat/completions \
  -H "Authorization: Bearer $API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "uncensored",
    "messages": [{"role": "user", "content": "Write a blunt product review of a cheap VPN."}]
  }'

The model accepts standard parameters like temperature, top_p, stop, and seed, giving you full control over the output style and randomness.

Streaming, JSON, and Function Calling

We support modern LLM features that are essential for robust application development. Streaming is enabled via Server-Sent Events (SSE), with token usage details provided in the last chunk. This allows your frontend to render text in real-time without waiting for the full response.

For structured data, use the response_format parameter set to json_object to enforce JSON output. Function calling (tools) is fully supported, allowing your LLM to interact with your backend services. You can specify tool choices and define schemas just as you would with other major providers. Note that we do not offer fine-tuning or model routing; this is a single, optimized uncensored model.

stream = client.chat.completions.create(
    model="uncensored",
    messages=[{"role": "user", "content": "Tell the story in second person."}],
    stream=True,
)
for chunk in stream:
    if chunk.choices and chunk.choices[0].delta.content:
        print(chunk.choices[0].delta.content, end="", flush=True)

Simple Token Pricing

No subscriptions, no monthly fees, and no hidden tiers. You pay only for the tokens you actually use. Errors and refusals are free, so you only spend when you get value. This transparent pricing model ensures you can predict costs accurately based on your token consumption.

  • Input Tokens: $0.25 per 1M tokens
  • Output Tokens: $1.00 per 1M tokens

Credit is prepaid and never expires. If you make a mistake, such as a double charge, contact Support for a fix; credit is not refunded but remains available for future use. This pay-as-you-go model is ideal for developers who want to scale costs directly with usage without committing to a monthly budget.

Questions and answers

What is the context window size?

The model supports a 64,000-token context window, which includes both the prompt and the completion. You can generate up to 16,000 tokens per request, or 2,048 tokens if you do not set the max_tokens parameter.

Is this the same as OpenAI's GPT model?

No. Our model is an open-weight model run on our own servers. It is not GPT, Claude, Gemini, Grok, DeepSeek, Qwen, or Llama. It is a distinct model tuned for uncensored output and is compatible with the OpenAI API format.

How do I pay for the API?

We accept crypto only: USDT (TRC20) or USDC (Base). You can top up any whole amount from $10 to $500. Credits do not expire, and you receive a 5% bonus for $50+ top-ups and a 10% bonus for $100+ top-ups.

Do you filter NSFW content?

We are uncensored and do not refuse lawful adult, fictional, or controversial topics. However, we do have a hard limit: sexual content involving minors is always refused. We do not encourage illegal use, but we do not apply broad moral filters.

What are the rate limits?

You are limited to 300 requests per minute per key and 8 concurrent requests at the same time per key. The maximum request body size is 8 MB. Each account can have one active key; generating a new key replaces the old one.

Your key is one form away

Create an account, copy the key, change the base URL. That is the whole setup.