Complete Guide: Integrating Unrestricted AI into Your Application
An unrestricted AI removes artificial safety barriers that limit creativity and response depth, allowing LLM models to generate adult content, explore controversial nuances, and execute complex roleplays without premature refusals. This technical guide explains how to integrate an OpenAI-compatible API that offers this consistent behavior, focusing on practical implementation and technical limit management.
Updated
Key points
- The 'uncensored' model is open, runs on our own servers, and ignores refusals for lawful adult content, except involving minors.
- The API follows the OpenAI standard, requiring only a change to the base_url and authentication token for immediate integration.
- The 100,000-token context allows long conversations but requires active management to avoid truncation or excessive costs.
- The price is transparent ($0.25/1M input tokens and $1.00/1M output) with prepaid credit that never expires.
Why Use an Unrestricted AI?
Many standard language models are trained with a 'safety' bias that makes them refuse adult, sexual, or politically sensitive topics even when they are perfectly valid for the intended use. An unrestricted AI eliminates this arbitrary filter, allowing the model to respond based on open-weight training, focusing on fidelity to the user prompt rather than a proprietary moderation layer.
This is crucial for immersive roleplay applications, where the character may have bolder personality traits, or for adult text generation where the model should not skip narrative steps to 'clean' the response. Unlike generic platforms that apply aggressive filters, our model is specifically tuned to respond to everything that is lawful, maintaining context coherence without frequent 'inappropriate content' interruptions.
Advantages of the OpenAI-Compatible API
The biggest technical advantage of using an API that follows the OpenAI standard is portability. You do not need to learn a new request syntax or write a custom client. Just point the base_url to https://api.apiiasemcensura.com/v1 and replace your API key. This allows the use of official libraries (SDKs) for Python, Node.js, and other languages.
- Immediate Integration: Existing code using
openaiworks with minimal configuration adjustments. - Standardization: The
/v1/chat/completionsendpoint is the industry standard, ensuring compatibility with third-party tools that accept OpenAI-compatible APIs. - Simplicity: No complex routing between multiple models. You have direct access to the 'uncensored' model hosted on our dedicated GPU servers.
Initial API Key Setup
The initial setup is designed to be fast and frictionless. You create an account by providing only an email and a password. No credit card is required to get the free trial credit, and the API key is displayed immediately after registration.
To get started, you need to obtain your API key on the account creation page. Remember that each account has a single active API key. If needed, you can regenerate it at any time, which instantly invalidates the previous API key. This API key must be sent in the header Authorization: Bearer YOUR_API_KEY in all requests.
Chat Completions Endpoint Implementation
The core of our API is the POST /v1/chat/completions endpoint. It accepts a list of messages (system, user, assistant) and returns a complete text response. To use the uncensored model, you must specify the model as uncensored.
from openai import OpenAI
client = OpenAI(base_url="https://api.apiiasemcensura.com/v1", api_key="YOUR_KEY")
resp = client.chat.completions.create(
model="uncensored",
messages=[{"role": "user", "content": "Summarise this thread without softening it."}],
)
print(resp.choices[0].message.content)
Enabling Streaming for Fast Responses
For applications requiring low perceived latency, such as real-time chatbots, enabling streaming via Server-Sent Events (SSE) is essential. By setting stream: true, the server sends token chunks as they are generated, allowing the client to display them immediately without waiting for the full calculation to finish.
This significantly improves the user experience, especially for long responses. The 'uncensored' model supports this functionality natively through the same chat completions endpoint.
Using Tool Calling for Advanced Features
In addition to text generation, the API supports tool calling. This allows the model to perform external actions, such as fetching data or performing calculations, by defining function schemas in the tools field of the request.
The model will identify when a tool should be used and return a tool_calls response instead of plain text. You then execute the function on your backend and send the result back to the API so the model can generate a final response based on the result. This is powerful for creating autonomous agents that operate without the tool restrictions typical of closed models.
Managing the 100k Token Context
Our API supports a context window of 100,000 tokens combining prompt and completion. This allows loading large amounts of conversation history or reference documents. However, the context is finite.
You must actively manage the conversation size. If the history exceeds the limit, the API may truncate older messages or reject the request if the body exceeds 8 MB. There is no fixed truncation rule at the beginning or end defined by the API; the responsibility to manage the context window (e.g., by removing old messages or summarizing previous conversations) falls on the developer.
Usage Monitoring and Limits
To ensure service stability, there are clear usage limits. Each API key is limited to 300 requests per minute. The request body cannot exceed 8 MB.
If you exceed the rate limit, you will receive an HTTP 429 (Too Many Requests) error. It is important to implement exponential backoff retries in your code. Since there is no monthly subscription, consumption is measured by processed tokens. You can monitor your credit balance on the account page, where prepaid credit never expires.
Conclusion and Next Steps
Integrating an unrestricted AI through an OpenAI-standard API offers the ideal balance between total control over generated content and technical implementation ease. By using the 'uncensored' model, you ensure that your roleplay, adult text, or research applications are not interrupted by artificial filters.
To get started, access our platform, create an account, and obtain your API key. Test the model with your $0.50 free trial credit and, when necessary, add credits to continue using the service. With clear technical documentation and transparent pricing, integration is direct and efficient.
Frequently asked questions
What does 'uncensored' exactly mean for this model?
It means the model lacks aggressive content filters that block adult, sexual, or controversial topics when they are lawful. It responds to the user prompt without refusals based on 'sensitivity', allowing greater creative freedom and fidelity to the context. The only strict and universal limit is sexual content involving minors, which is always blocked.
What are the prices and how does payment work?
The model is $0.25 per 1 million input tokens and $1.00 per 1 million output tokens. There is no monthly subscription or usage fees. You pay with prepaid credit, which never expires. The minimum top up is $10, and credit bonuses apply to larger top ups (+5% from $50, +10% from $100).
Can I use this model for image or audio generation?
No. This API offers only the language model (LLM) for text. There is no support for embeddings, image generation, audio, or video. The main endpoint is /v1/chat/completions, focused purely on text interaction.
Are my prompt data used to train the model?
No. Privacy is guaranteed: your prompts and responses are not used to train the model. The model is an open-weight model, running on our servers, and usage is solely to fulfill your specific request.
Your key is one form away
Create an account, copy the key, and change the base URL. That's all the configuration.