Google: Veo 3 Fast with Audio
This model is no longer available.Add AI to your application with Puter.js.
Explore Other ModelsModel Card
Google Veo 3 Fast with Audio is the audio-enabled version of the speed-optimized Veo 3 Fast model, combining faster generation times and lower costs with native synchronized audio generation. It delivers sound effects, dialogue, and ambient audio while optimizing for speed and affordability in production workflows.
Max Duration 8s
seconds
Frame Rate 24
fps
Aspect Ratio 16:9, 9:16
supported
Release Date May 20, 2025
Code Example
Add AI to your app with the Puter.js AI API, no API keys or setup required.
// npm install @heyputer/puter.js
import { puter } from '@heyputer/puter.js';
puter.ai.txt2vid("A cat playing with a ball of yarn").then(video => {
document.body.appendChild(video);
});
<html>
<body>
<script src="https://js.puter.com/v2/"></script>
<script>
puter.ai.txt2vid("A cat playing with a ball of yarn").then(video => {
document.body.appendChild(video);
});
</script>
</body>
</html>
More AI Models From Google
Gemini 3.8 Flash
Gemini 3.8 Flash is Google's workhorse Flash-tier model, released September 2, 2026, three weeks after Gemini 3.7 Flash. It's built for long-horizon software engineering, agentic workflows, and multi-step reasoning in professional domains. Google reports it outperforms 3.7 Flash and other frontier models on DeepSWE v1.1 for autonomous engineering tasks, on Vals Finance Agent V2, and on Harvey's Legal Agent Benchmark. It scores 54.9% on HLE-Verified, and completes more than three times as many tasks as 3.7 Flash in Google's long-running, document-heavy workflow evaluations. It accepts text, image, video, audio, and PDF input with a 1M token context window and 64K token output limit, supports function calling and iterative tool use, and has a March 2026 knowledge cutoff. Priced at $0.75 per million input tokens and $3.75 per million output tokens through 2026, it targets teams running coding agents or document-heavy enterprise workflows.
ChatGemini 3.7 Flash
Gemini 3.7 Flash is Google's workhorse Flash-tier model, released August 13, 2026, three weeks after Gemini 3.6 Flash. It's built for coding and agentic workflows, targeting software engineering, web development, and knowledge-dense domains like finance and law. Google reports gains over Gemini 3.6 Flash on several benchmarks. DeepSWE v1.1 rose from 49.0% to 65.3%, FrontierCode 1.1 from 34.4% to 43.6%, and AutomationBench from 17.0% to 30.4%. On FrontierCode 1.1 it scores above Claude Sonnet 5 (42.7%) and GPT-5.6 Terra (41.3%), though GPT-5.6 Terra edges it out on Terminal-bench 2.1 (87.4% vs 85.8%). It accepts text, image, video, audio, and PDF input with a 1M token context window and 64K token output limit. It supports function calling, search as a tool, and computer use, and has a March 2026 knowledge cutoff. It's priced at roughly half of Gemini 3.6 Flash's rate, fitting teams running coding agents or high-volume document processing.
ChatGemini Robotics ER 2 Preview
Gemini Robotics ER 2 Preview is Google's embodied reasoning model for robotics, available through the Gemini API and Google AI Studio. It takes video, images, audio and text, reasons about a physical scene, plans multi-step tasks, and hands actions off to a vision-language-action model, a robotics API, or developer-defined tools through function calling. Google says it can plan its next step while a robot is moving, works with the Gemini Live API, and can monitor a task in video, detect failures and retry individual steps. It also supports multi-robot coordination. In Google's reported tests it reached 91.3% accuracy at identifying when a key event occurred in a video (mean error 0.96 seconds) and 57.4% on progress classification. It is meant for developers building robot planning, success detection and scene understanding on top of the Gemini API.
Frequently Asked Questions
Veo 3 Fast with Audio is no longer available through Puter.js. Explore other AI models for alternatives.
Veo 3 Fast with Audio was created by Google and released on May 20, 2025.
Yes — the Veo 3 Fast with Audio API works with any JavaScript framework, Node.js, or plain HTML through Puter.js. Just include the library and start building. See the documentation for more details.
Get started with Puter.js
Add AI to your application without worrying about API keys or setup.
Explore Models View Tutorials