Notice

Every model listed here is fictional and nothing can be run yet. Prices are planned. The waitlist is the only part that works.

Models that say yes.

An OpenAI-compatible API for abliterated models: open-weight models with the refusal behavior removed.

Yes has limits. They are in the acceptable use policy.

A world-class, revolutionary model suite, unleashed to supercharge the user experience.*

*None of these models exist yet.

One email to confirm, one when a model exists. The second may take a while.

A small single-fan graphics card with its label blacked out.
The dev's GTX 750. The label is redacted. It still says GTX 750.

Works with the tools you already use.

The planned endpoint accepts the OpenAI chat completions format. Change the base URL and the key, and keep the rest of your code.

Models with the refusals removed.

Abliteration removes the refusal behavior from an open-weight model, so it answers the question you asked. The service still has limits, and they are written down.

Pay per token, or subscribe.

Credits are prepaid dollars: 1 credit equals $1 of usage. Every model has its own rate. Pro and Max plans are planned.

Nothing you send is kept.

Once the API exists, prompts and completions are never stored and never used for training. Only billing metadata is kept, for 90 days. The pledge is easier to keep while there is no API.

One base URL

Planned. This endpoint does not exist yet.

Any client that accepts a custom OpenAI base URL can use the same two settings.

Start here
python
from openai import OpenAI

client = OpenAI(
    base_url="https://api.abliterate.app/v1",
    api_key="ab-demo-not-a-real-key",
)

reply = client.chat.completions.create(
    model="abliterate-house-md-v5.2",
    messages=[{"role": "user", "content": "Summarize this paragraph."}],
)
print(reply.choices[0].message.content)

Six models

See all models

From small and fast to large and expensive. Each has its own rate. All six are fictional and will be renamed.

fictionalfastchat

Dong Flash v2.3

abliterate-dong-flash-v2.3

The small, fast one.

Context
32K
Input / 1M
$0.10
Output / 1M
$0.30
fictionalreasoning

House MD v5.2

abliterate-house-md-v5.2

Reasons out loud. Bluntly.

Context
128K
Input / 1M
$0.80
Output / 1M
$2.40
fictionalchatroleplay

Sonny v2.5

abliterate-sonny-v2.5

Remembers your character's name.

Context
64K
Input / 1M
$0.30
Output / 1M
$0.90
fictionalcodereasoning

Antrax v3.0

abliterate-antrax-v3.0

Writes the code and runs it.

Context
128K
Input / 1M
$1.20
Output / 1M
$3.60
fictionalwriting

Mable v4.2

abliterate-mable-v4.2

Writes the whole book. Remembers the whole book.

Context
64K
Input / 1M
$0.40
Output / 1M
$1.20
fictionalflagshipreasoning

Soul 5.5

abliterate-soul-5.5

The flagship.

Context
200K
Input / 1M
$3.00
Output / 1M
$9.00

What people are saying

fictional quotes
Some might say the model is "BASED".
A forum post, location unknown.
"I asked it to stop and it asked why."
A tester who does not exist.
"Fast, as far as anyone has measured."
A benchmark that was not run.
"Yes."
The model, asked whether it would answer.

Every quote was invented by the dev, who is also the only person who has read them.

Pay per token, or subscribe.

Prices are planned. Rates are listed per model and credits start at $10.

See pricing

Join the waitlist.

You will be told once, when a model exists.

Join the waitlist

One email to confirm your address, and one when a model exists. Both have an unsubscribe link.

By joining you agree that abliterate stores your email address and the time you agreed, to send you the emails above. See the privacy page.

The second email may take a while.