Cognitive Computations: Dolphin Mistral 24B Venice Edition (Uncensored)
Try Dolphin Mistral 24B Venice Edition (Uncensored) for free in your browser, and add it to your app for free with Puter.js AI API.
Try it free Add to your app// npm install @heyputer/puter.js
import { puter } from '@heyputer/puter.js';
puter.ai.chat("Explain quantum computing in simple terms", {
model: "cognitivecomputations/dolphin-mistral-24b-venice-edition"
}).then(response => {
document.body.innerHTML = response.message.content;
});
<html>
<body>
<script src="https://js.puter.com/v2/"></script>
<script>
puter.ai.chat("Explain quantum computing in simple terms", {
model: "cognitivecomputations/dolphin-mistral-24b-venice-edition"
}).then(response => {
document.body.innerHTML = response.message.content;
});
</script>
</body>
</html>
Model Card
Dolphin Mistral 24B Venice Edition is a 24 billion parameter, uncensored general-purpose language model fine-tuned from Mistral Small 24B (Instruct-2501), built by Cognitive Computations (the Dolphin project, founded by Eric Hartford) in collaboration with Venice.ai. This paid API variant offers a 128K context window and an April 2024 knowledge cutoff.
On Venice's uncensored benchmark, it recorded a 2.2% refusal rate, the lowest among tested models and well below Llama 4 Maverick (11.11%), Gemini 2.5 Flash Preview (53.33%), GPT-4o-mini (64.44%), and Claude Sonnet 3.7 (71.11%).
While base Mistral Small 24B leaned STEM-heavy, this fine-tune adds strong creative writing and storytelling, with consistent character and narrative memory, plus steerable tone control that defaults to neutral and polite.
Suited for developers who need minimal refusals, custom alignment, or unrestricted content generation via API.
Context Window 128K
tokens
Max Output 8K
tokens
Input Cost $0.2
per million tokens
Output Cost $0.9
per million tokens
Release Date Jul 9, 2025
Try Dolphin Mistral 24B Venice Edition (Uncensored) for free
Try Dolphin Mistral 24B Venice Edition (Uncensored) instantly in your browser.
This playground uses the Puter.js AI API — no API keys or setup required.
Frequently Asked Questions
You can access Dolphin Mistral 24B Venice Edition (Uncensored) by Cognitive Computations through Puter.js AI API. Include the library in your web app or Node.js project and start making calls with just a few lines of JavaScript — no backend and no configuration required. You can also use it with Python or cURL via Puter's OpenAI-compatible API.
Dolphin Mistral 24B Venice Edition (Uncensored) is free to try with a Puter account. Every account includes a free AI allowance, and you can chat with it in the playground on this page. You can upgrade your account anytime for a larger allowance.
Dolphin Mistral 24B Venice Edition (Uncensored) is free to integrate using the Puter.js AI API. With the User-Pays Model, you can add AI to your app for $0, since users cover their own AI usage through their Puter account.
| Price per 1M tokens | |
|---|---|
| Input | $0.2 |
| Output | $0.9 |
Dolphin Mistral 24B Venice Edition (Uncensored) was created by Cognitive Computations and released on Jul 9, 2025.
Dolphin Mistral 24B Venice Edition (Uncensored) supports a context window of 128K tokens. For reference, that is roughly equivalent to 256 pages of text.
Dolphin Mistral 24B Venice Edition (Uncensored) can generate up to 8K tokens in a single response.
Yes — the Dolphin Mistral 24B Venice Edition (Uncensored) API works with any JavaScript framework, Node.js, or plain HTML through Puter.js. Just include the library and start building. See the documentation for more details.
Add Dolphin Mistral 24B Venice Edition (Uncensored) to your app for free
Developers can integrate Dolphin Mistral 24B Venice Edition (Uncensored) for free using the Puter.js AI API.
With the User-Pays Model, each user covers their own AI usage instead of the developer.