Google Foundation Models
Unlock the complete study guide + 1,040 practice questions across 16 full exams.
Bundled into the existing Generative AI Leader premium course — no separate purchase.
14-day money-back guarantee — no questions asked.
Included in this chapter:
- The four families and what each is for
- Gemini: the multimodal workhorse
- Gemma, Imagen, and Veo up close
- Choosing a model for a use case
Google's four generative foundation model families at a glance
| Dimension | Gemini | Gemma | Imagen | Veo |
|---|---|---|---|---|
| Primary output | Text plus multimodal reasoning | Text (open-weight LLM) | Images | Video with audio |
| Openness | Proprietary, managed | Open-weight, self-hostable | Proprietary, managed | Proprietary, managed |
| Where it runs | Google-managed API | Your infra, on-device, or Model Garden | Google-managed API | Google-managed API |
| Signature strength | Multimodal input, long context, agentic reasoning | Lightweight, private, customizable | Photorealism and text-in-image | Cinematic clips with native audio |
| Reach for it when | Chatbots, analysis, agents, summarization | Data must stay in your control, offline, or fine-tuned | Marketing, product, and packaging imagery | Video ads, storyboards, previsualization |
Decision tree
Cheat sheet
Unlock with Premium — includes all practice exams and the complete study guide.