Ship a Full-Stack App with One Prompt

Copy this prompt into your AI coding agent, or open it in one below.

Give this to your AI Create a to-do list app using Puter.js

Coding manually? see the guide

Thinking Machines Lab

Thinking Machines Lab: Inkling Small

thinkingmachines/inkling-small

Access Inkling Small from Thinking Machines Lab using Puter.js AI API.

Get Started
// npm install @heyputer/puter.js
import { puter } from '@heyputer/puter.js';

puter.ai.chat("Explain quantum computing in simple terms", {
    model: "thinkingmachines/inkling-small"
}).then(response => {
    document.body.innerHTML = response.message.content;
});
<html>
<body>
    <script src="https://js.puter.com/v2/"></script>
    <script>
        puter.ai.chat("Explain quantum computing in simple terms", {
            model: "thinkingmachines/inkling-small"
        }).then(response => {
            document.body.innerHTML = response.message.content;
        });
    </script>
</body>
</html>
# pip install openai
from openai import OpenAI

client = OpenAI(
    base_url="https://api.puter.com/puterai/openai/v1/",
    api_key="YOUR_PUTER_AUTH_TOKEN",
)

response = client.chat.completions.create(
    model="thinkingmachines/inkling-small",
    messages=[
        {"role": "user", "content": "Explain quantum computing in simple terms"}
    ],
)

print(response.choices[0].message.content)
curl https://api.puter.com/puterai/openai/v1/chat/completions \
  -H "Content-Type: application/json" \
  -H "Authorization: Bearer YOUR_PUTER_AUTH_TOKEN" \
  -d '{
    "model": "thinkingmachines/inkling-small",
    "messages": [
      {"role": "user", "content": "Explain quantum computing in simple terms"}
    ]
  }'

Model Card

Inkling Small is a mixture-of-experts model from Thinking Machines Lab, the AI research company co-founded by former OpenAI CTO Mira Murati. Released on July 30, 2026, about two weeks after the flagship Inkling FP4, it has 276 billion total parameters with 12 billion active per token, roughly a quarter of Inkling FP4's active-parameter count, and its weights are released under an Apache 2.0 license.

Like its larger sibling, the model is multimodal, accepting text, image, and audio inputs and producing text output, and it supports controllable reasoning effort.

Thinking Machines Lab reports 80.2% on SWE-bench Verified, 89.5% on GPQA Diamond, and 82.2% on IFBench. On the third-party Artificial Analysis Intelligence Index it scores 40, within a point of Inkling FP4's 41, and the company says it beats Inkling FP4 on several reasoning and coding benchmarks while trailing it on factual-knowledge evaluations. It is available through Together AI's inference platform.

Context Window 524K

tokens

Max Output 524K

tokens

Input Cost $0.5

per million tokens

Output Cost $1.2

per million tokens

Release Date Jul 30, 2026

 

Model Playground

Try Inkling Small instantly in your browser.
This playground uses the Puter.js AI API — no API keys or setup required.

Chat thinkingmachines/inkling-small
Thinking Machines Lab
Chat with Inkling Small
Powered by Puter.js

Frequently Asked Questions

How do I use Inkling Small?

You can access Inkling Small by Thinking Machines Lab through Puter.js AI API. Include the library in your web app or Node.js project and start making calls with just a few lines of JavaScript — no backend and no configuration required. You can also use it with Python or cURL via Puter's OpenAI-compatible API.

Is Inkling Small free?

Yes, it is free if you're using it through Puter.js. With the User-Pays Model, you can add Inkling Small to your app at no cost — your users pay for their own AI usage directly, making it completely free for you as a developer.

What is the pricing for Inkling Small?
Inkling Small costs $0.5 per 1M input tokens and $1.2 per 1M output tokens.
Price per 1M tokens
Input$0.5
Output$1.2
Who created Inkling Small?

Inkling Small was created by Thinking Machines Lab and released on Jul 30, 2026.

What is the context window of Inkling Small?

Inkling Small supports a context window of 524K tokens. For reference, that is roughly equivalent to 1,049 pages of text.

What is the max output length of Inkling Small?

Inkling Small can generate up to 524K tokens in a single response.

Does it work with React / Vue / Vanilla JS / Node / etc.?

Yes — the Inkling Small API works with any JavaScript framework, Node.js, or plain HTML through Puter.js. Just include the library and start building. See the documentation for more details.

Get started with Puter.js

Add Inkling Small to your app without worrying about API keys or setup.

Read the Docs View Tutorials