Overview
GeminiGenAI bundles three generation capabilities in one web app: AI video generation (text-to-video, image-to-video with reference-image support, and video extension) using Google Veo 2, Veo 3, and Veo 3 Fast-class models, with 16:9 and 9:16 aspect ratios and 720p/1080p output; AI image generation from text; and AI text-to-speech via Gemini TTS (Gemini 2.5 Flash and Gemini 2.5 Pro voices), including document-to-speech for formats such as DOCX, PDF, EPUB, and HTML.
The web editor includes a model selector, prompt editor, negative prompts, and a live credit-cost estimate before each generation. A developer API exposes the same controls — model, duration, aspect ratio, resolution, prompts, reference images — and the TTS API adds per-block voice assignment for multi-voice dialogue, 400+ voices across languages and accents, speed adjustment from 0.25x to 4x, and MP3/WAV export.
Billing is credit-based and pay-as-you-go. No official pricing figures were verifiable on the official site at research time — pricing and free-plan details are not publicly disclosed. GeminiGenAI is an independent third-party platform and is not affiliated with Google.