Loading...
Loading...
Google introduced Gemini Embedding 2, a multimodal embedding model, with new pricing for text, image, audio, and video inputs, replacing Gemini 3.1 Flash-Lite Preview.
New pricing model for multimodal inputs
Gemini 3.1 Flash-Lite Preview is now available. Try it in AI Studio . Our newest embeddings model, more stable and with higher rate limits than previous versions, available to developers on the free and paid tiers of the Gemini API. Last updated 2026-03-09 UTC. [[["Easy to understand","easyToUnderstand","thumb-up"],["Solved my problem","solvedMyProblem","thumb-up"],["Other","otherUp","thumb-up"]],[["Missing the information I need","missingTheInformationINeed","thumb-down"],["Too complicated / too many steps","tooComplicatedTooManySteps","thumb-down"],["Out of date","outOfDate","thumb-down"],["Samples / code issue","samplesCodeIssue","thumb-down"],["Other","otherDown","thumb-down"]],["Last updated 2026-03-09 UTC."],[],[]]
Announcing Gemini Embedding 2 , our first fully multimodal embedding model. Gemini Embedding 2 Preview gemini-embedding-2-preview Our first multimodal embedding model, mapping text, images, video, audio, and PDFs into a unified embedding space. Text input price Image input price $0.45 ($0.00012 per image) Audio input price $6.50 ($0.00016 per second) Video input price $12.00 ($0.00079 per fps)
Announcing Gemini Embedding 2 , our first fully multimodal embedding model.
Check your models and usage for free. See the relevant official changes, estimated cost, and a practical next step.