Google: Gemma 3n 2B
This model is no longer available.Add AI to your application with Puter.js.
Explore Other ModelsModel Card
Gemma 3n E2B Instruct (Free) is Google's mobile-first open model with an effective 2B parameter memory footprint using Per-Layer Embeddings. It's optimized for on-device AI with audio, text, image, and video understanding.
Context Window 8K
tokens
Max Output 2K
tokens
Input Cost $0
per million tokens
Output Cost $0
per million tokens
Release Date Jun 25, 2025
Code Example
Add AI to your app with the Puter.js AI API, no API keys or setup required.
// npm install @heyputer/puter.js
import { puter } from '@heyputer/puter.js';
puter.ai.chat("Explain quantum computing in simple terms").then(response => {
document.body.innerHTML = response.message.content;
});
<html>
<body>
<script src="https://js.puter.com/v2/"></script>
<script>
puter.ai.chat("Explain quantum computing in simple terms").then(response => {
document.body.innerHTML = response.message.content;
});
</script>
</body>
</html>
More AI Models From Google
Gemini 3.8 Flash
Gemini 3.8 Flash is Google's workhorse Flash-tier model, released September 2, 2026, three weeks after Gemini 3.7 Flash. It's built for long-horizon software engineering, agentic workflows, and multi-step reasoning in professional domains. Google reports it outperforms 3.7 Flash and other frontier models on DeepSWE v1.1 for autonomous engineering tasks, on Vals Finance Agent V2, and on Harvey's Legal Agent Benchmark. It scores 54.9% on HLE-Verified, and completes more than three times as many tasks as 3.7 Flash in Google's long-running, document-heavy workflow evaluations. It accepts text, image, video, audio, and PDF input with a 1M token context window and 64K token output limit, supports function calling and iterative tool use, and has a March 2026 knowledge cutoff. Priced at $0.75 per million input tokens and $3.75 per million output tokens through 2026, it targets teams running coding agents or document-heavy enterprise workflows.
ChatGemini 3.7 Flash
Gemini 3.7 Flash is Google's workhorse Flash-tier model, released August 13, 2026, three weeks after Gemini 3.6 Flash. It's built for coding and agentic workflows, targeting software engineering, web development, and knowledge-dense domains like finance and law. Google reports gains over Gemini 3.6 Flash on several benchmarks. DeepSWE v1.1 rose from 49.0% to 65.3%, FrontierCode 1.1 from 34.4% to 43.6%, and AutomationBench from 17.0% to 30.4%. On FrontierCode 1.1 it scores above Claude Sonnet 5 (42.7%) and GPT-5.6 Terra (41.3%), though GPT-5.6 Terra edges it out on Terminal-bench 2.1 (87.4% vs 85.8%). It accepts text, image, video, audio, and PDF input with a 1M token context window and 64K token output limit. It supports function calling, search as a tool, and computer use, and has a March 2026 knowledge cutoff. It's priced at roughly half of Gemini 3.6 Flash's rate, fitting teams running coding agents or high-volume document processing.
ChatGemini Robotics ER 2 Preview
Gemini Robotics ER 2 Preview is Google's embodied reasoning model for robotics, available through the Gemini API and Google AI Studio. It takes video, images, audio and text, reasons about a physical scene, plans multi-step tasks, and hands actions off to a vision-language-action model, a robotics API, or developer-defined tools through function calling. Google says it can plan its next step while a robot is moving, works with the Gemini Live API, and can monitor a task in video, detect failures and retry individual steps. It also supports multi-robot coordination. In Google's reported tests it reached 91.3% accuracy at identifying when a key event occurred in a video (mean error 0.96 seconds) and 57.4% on progress classification. It is meant for developers building robot planning, success detection and scene understanding on top of the Gemini API.
Frequently Asked Questions
Gemma 3n 2B is no longer available through Puter.js. Explore other AI models for alternatives.
| Price per 1M tokens | |
|---|---|
| Input | $0 |
| Output | $0 |
Gemma 3n 2B was created by Google and released on Jun 25, 2025.
Gemma 3n 2B supports a context window of 8K tokens. For reference, that is roughly equivalent to 16 pages of text.
Gemma 3n 2B can generate up to 2K tokens in a single response.
Gemma 3n 2B scores 4.8 on the Artificial Analysis Intelligence Index, outperforming 0% of tracked models. On math, it scores 10.3 (outperforms 13% of models).
Yes — the Gemma 3n 2B API works with any JavaScript framework, Node.js, or plain HTML through Puter.js. Just include the library and start building. See the documentation for more details.
Get started with Puter.js
Add AI to your application without worrying about API keys or setup.
Explore Models View Tutorials