Ship a Full-Stack App with One Prompt

Copy this prompt into your AI coding agent, or open it in one below.

Give this to your AI Create a to-do list app using Puter.js

Coding manually? see the guide

Google: Veo 3 with Audio

This model is no longer available.

Add AI to your application with Puter.js.

Explore Other Models

Model Card

Google Veo 3 with Audio is the audio-enabled configuration of Veo 3 that generates synchronized sound effects, dialogue, ambient noise, and music natively alongside video content. It produces complete audiovisual experiences from text prompts, eliminating the need for separate audio post-production.

Max Duration 8s

seconds

Frame Rate 24

fps

Aspect Ratio 16:9, 9:16

supported

Release Date May 20, 2025

 

Code Example

Add AI to your app with the Puter.js AI API, no API keys or setup required.

// npm install @heyputer/puter.js
import { puter } from '@heyputer/puter.js';

puter.ai.txt2vid("A cat playing with a ball of yarn").then(video => {
    document.body.appendChild(video);
});
<html>
<body>
    <script src="https://js.puter.com/v2/"></script>
    <script>
        puter.ai.txt2vid("A cat playing with a ball of yarn").then(video => {
            document.body.appendChild(video);
        });
    </script>
</body>
</html>

More AI Models From Google

Find other Google models →

Chat

Gemini 3.8 Flash

Gemini 3.8 Flash is Google's workhorse Flash-tier model, released September 2, 2026, three weeks after Gemini 3.7 Flash. It's built for long-horizon software engineering, agentic workflows, and multi-step reasoning in professional domains. Google reports it outperforms 3.7 Flash and other frontier models on DeepSWE v1.1 for autonomous engineering tasks, on Vals Finance Agent V2, and on Harvey's Legal Agent Benchmark. It scores 54.9% on HLE-Verified, and completes more than three times as many tasks as 3.7 Flash in Google's long-running, document-heavy workflow evaluations. It accepts text, image, video, audio, and PDF input with a 1M token context window and 64K token output limit, supports function calling and iterative tool use, and has a March 2026 knowledge cutoff. Priced at $0.75 per million input tokens and $3.75 per million output tokens through 2026, it targets teams running coding agents or document-heavy enterprise workflows.

Chat

Gemini 3.7 Flash

Gemini 3.7 Flash is Google's workhorse Flash-tier model, released August 13, 2026, three weeks after Gemini 3.6 Flash. It's built for coding and agentic workflows, targeting software engineering, web development, and knowledge-dense domains like finance and law. Google reports gains over Gemini 3.6 Flash on several benchmarks. DeepSWE v1.1 rose from 49.0% to 65.3%, FrontierCode 1.1 from 34.4% to 43.6%, and AutomationBench from 17.0% to 30.4%. On FrontierCode 1.1 it scores above Claude Sonnet 5 (42.7%) and GPT-5.6 Terra (41.3%), though GPT-5.6 Terra edges it out on Terminal-bench 2.1 (87.4% vs 85.8%). It accepts text, image, video, audio, and PDF input with a 1M token context window and 64K token output limit. It supports function calling, search as a tool, and computer use, and has a March 2026 knowledge cutoff. It's priced at roughly half of Gemini 3.6 Flash's rate, fitting teams running coding agents or high-volume document processing.

Chat

Gemini Robotics ER 2 Preview

Gemini Robotics ER 2 Preview is Google's embodied reasoning model for robotics, available through the Gemini API and Google AI Studio. It takes video, images, audio and text, reasons about a physical scene, plans multi-step tasks, and hands actions off to a vision-language-action model, a robotics API, or developer-defined tools through function calling. Google says it can plan its next step while a robot is moving, works with the Gemini Live API, and can monitor a task in video, detect failures and retry individual steps. It also supports multi-robot coordination. In Google's reported tests it reached 91.3% accuracy at identifying when a key event occurred in a video (mean error 0.96 seconds) and 57.4% on progress classification. It is meant for developers building robot planning, success detection and scene understanding on top of the Gemini API.

Frequently Asked Questions

How do I use Veo 3 with Audio?

Veo 3 with Audio is no longer available through Puter.js. Explore other AI models for alternatives.

Who created Veo 3 with Audio?

Veo 3 with Audio was created by Google and released on May 20, 2025.

Does it work with React / Vue / Vanilla JS / Node / etc.?

Yes — the Veo 3 with Audio API works with any JavaScript framework, Node.js, or plain HTML through Puter.js. Just include the library and start building. See the documentation for more details.

Get started with Puter.js

Add AI to your application without worrying about API keys or setup.

Explore Models View Tutorials