Ship a Full-Stack App with One Prompt

Copy this prompt into your AI coding agent, or open it in one below.

Give this to your AI Create a to-do list app using Puter.js

Coding manually? see the guide

Pruna AI

Pruna AI API

Access Pruna AI instantly with Puter.js, and add AI to any app in a few lines of code without backend or API keys.

// npm install @heyputer/puter.js
import { puter } from '@heyputer/puter.js';

puter.ai.txt2img("A beautiful sunset", {
    model: "prunaai/z-image-turbo"
}).then(imageElement => {
    document.body.appendChild(imageElement);
});
<html>
<body>
    <script src="https://js.puter.com/v2/"></script>
    <script>
        puter.ai.txt2img("A beautiful sunset", {
            model: "prunaai/z-image-turbo"
        }).then(imageElement => {
            document.body.appendChild(imageElement);
        });
    </script>
</body>
</html>

List of Pruna AI Models

Image

FLUX Fast

prunaai/flux-fast

FLUX Fast is Pruna AI's speed-optimized version of Black Forest Labs' FLUX.1 [dev] image generation model. Pruna applies its own compression stack to the base model, combining flux caching (reusing outputs from earlier diffusion steps) with its diffusers2 compiler. On FLUX.1 [dev], Pruna reports this combination gives a 2.7x speedup over the original model, while keeping output quality close to the source on standard metrics such as LPIPS, SSIM, and PSNR. Pruna markets FLUX Fast as the fastest FLUX endpoint available, aimed at cases where turnaround time matters, such as interactive apps and high-throughput generation pipelines. It suits developers who want FLUX.1 [dev] quality without the latency of the standard model.

Image

P-Image

prunaai/p-image

P-Image is Pruna AI's own text-to-image model, generating images in about one second at $0.005 each. Pruna describes it as building on prior open image models from Black Forest Labs (the FLUX series) and Alibaba (Qwen Image, Z-Image) rather than a single from-scratch architecture, though the company hasn't published full architectural details. Pruna positions P-Image as matching the quality of state-of-the-art image models while running about 30 times cheaper and 50 times faster, per its own figures. It's built for exact prompt adherence and reliable, controllable text rendering inside generated images. Use cases include real-time or high-volume applications, such as advertising and gaming assets, where both cost and latency matter.

Image

P-Image LoRA

prunaai/p-image-lora

P-Image LoRA is Pruna AI's inference endpoint for running custom-trained LoRA (Low-Rank Adaptation) adapters on top of P-Image. Rather than a fine-tuned model in itself, it's the serving side of a pair, users train a LoRA with Pruna's separate P-Image Trainer, then deploy it through P-Image LoRA to generate images with that custom style, subject, or character baked in. Trained adapters can be shared through Pruna's P-Image LoRA collection on Hugging Face. It carries the same fast, Pruna-optimized generation as the base P-Image model. It suits developers who want P-Image's speed and pricing combined with a personalized model, without retraining or hosting a full image model themselves.

Image

Wan 2.2 Image

prunaai/wan-2.2-image

Wan 2.2 Image is Pruna AI's optimized version of Alibaba's Wan 2.2 model running in image-generation mode, rather than the video mode Wan is best known for. Pruna applies its compression engine to the base model, generating 2-megapixel images in 3-4 seconds at $0.02 each. Pruna reports its version running 2.4 times faster than SeedDream, 1.8 times faster than FLUX1.1 [pro], and 1.1 times faster than its own earlier Wan 2.1 Image model. The model is tuned toward cinematic, aesthetically strong output rather than speed alone. It suits text-to-image workloads where a photographic, cinematic look matters, such as creative and marketing content, drawing on Wan's video-model lineage for composition.

Image

Z-Image Turbo

prunaai/z-image-turbo

Z-Image Turbo is Pruna AI's optimized version of Z-Image, a 6B-parameter open-source text-to-image model from Alibaba's Tongyi Lab, the team behind Qwen, released under Apache 2.0 in November 2025. The base model uses a Scalable Single-Stream DiT architecture and is distilled with Decoupled-DMD to just 8 sampling steps, which is what "Turbo" refers to. Pruna layers its own caching, compilation, and quantization on top for further speed, with generation reported at under one second on enterprise hardware. Z-Image ranks first among open-source image models on the Artificial Analysis leaderboard and supports native Chinese and English text rendering. Pricing is tiered by output resolution, from 0.25 cents at 0.5 megapixels up to 2 cents at 4 megapixels.

Frequently Asked Questions

What is this Pruna AI API about?

The Pruna AI API gives you access to models for AI image generation. Through Puter.js, you can start using Pruna AI models instantly with zero setup or configuration.

Which Pruna AI models can I use?

Puter.js supports a variety of Pruna AI models, including FLUX Fast, P-Image, P-Image LoRA, and more. Find all AI models supported by Puter.js in the AI model list.

How much does it cost?

With the User-Pays model, users cover their own AI costs through their Puter account. This means you can build apps without worrying about infrastructure expenses.

What is Puter.js?

Puter.js is a JavaScript library that provides access to AI, storage, and other cloud services directly from a single API. It handles authentication, infrastructure, and scaling so you can focus on building your app.

Does this work with React / Vue / Vanilla JS / Node / etc.?

Yes — the Pruna AI API through Puter.js works with any JavaScript framework, Node.js, or plain HTML. Just include the library and start building. See the documentation for more details.