Gemini logo

Gemini

Featured
Verified Company

Google’s multimodal AI assistant

Text Vision Audio Code +4 more
4.6 Rating
1.4K reviews
30K+ views

At a Glance

Developer Google DeepMind
Model Type Multimodal
Context Window 1M
Release Date Dec 6, 2023
Pricing Freemium

About Gemini

Google’s multimodal AI for text, images and code, integrated across Google apps.

What Gemini Can Do

Text

Understands and writes natural language — prompts in, prose out.

Vision

Reads images: photos, screenshots, diagrams and scanned documents.

Audio

Works with speech and sound as an input or an output, not just text.

Code

Writes, explains and refactors source code across common languages.

Function Calling

Calls your own functions and tools with structured arguments.

Web Search

Looks things up on the live web rather than answering from memory.

JSON Mode

Guarantees valid, schema-shaped JSON your application can parse.

Streaming

Streams the response token by token so output appears immediately.

Reviews

4.6/5
Based on 1.4K reviews

No written reviews yet

Be the first to share your experience with Gemini.

Write the first review
Share your experience
Log in to write a review for Gemini.
Log in to review

Ready to build with Gemini?

Head to the official site for docs and API keys, or compare Gemini against every other model in the directory.