Gemini 2.0 Flash-Lite

Retired · Language, Multimodal

Gemini 2.0 Flash-Lite is a multimodal language model from Google, released in public preview on 5 February 2025 and generally available in the Gemini API, through Google AI Studio and Vertex AI, from 25 February. It took text, image, audio and video input with a context window of 1,048,576 tokens, returned text, and was shut down on 1 June 2026. [1] [2] [3]Model Information [4]June 1, 2026 Primary source

Timeline of Gemini 2.0 Flash-Lite →

Claims and evidence

  • Released Public preview (gemini-2.0-flash-lite-preview-02-05); generally available on 2025-02-25. Primary source[1]2.0 Flash-Lite section [4]February 5, 2025 [5]Gemini 2.0 models: gemini-2.0-flash-lite-preview-02-05
  • Deprecated Date of the deprecation announcement in the Gemini API release notes. Primary source[4]February 18, 2026
  • Retired Primary source[4]June 1, 2026 [5]Gemini 2.0 models [6]Retired models
  • Generally available (gemini-2.0-flash-lite-001) Primary source[4]February 25, 2025 [2] [5]Gemini 2.0 models
  • Status Retired Primary source[5]Gemini 2.0 models [4]June 1, 2026 [7]warning banner [6]Retired models
  • Replaced by Gemini 3.1 Flash-LiteRecommended replacement for gemini-2.0-flash-lite on the Gemini API deprecations page and the Agent Platform model versions page when checked on 2026-10-01. The deprecations page names gemini-2.5-flash-lite as the replacement for the earlier preview IDs. Primary source[5]Gemini 2.0 models [7]warning banner [6]Retired models
    • Change · Other Google says it offers better quality than Gemini 1.5 Flash at the same speed and cost.Compared with Gemini 1.5 Flash Primary source[1]2.0 Flash-Lite section
    • Change · Other Uses a single price per input type, dropping the split between short and long context prices used for Gemini 1.5 Flash.Compared with Gemini 1.5 Flash Primary source[8]
    • Input text, image, audio, video Primary source[7]Supported data types [3]Model Information: Inputs
    • Output text Primary source[7]Supported data types [3]Model Information: Outputs
    • Feature Function calling Primary source[7]Capabilities
    • Context window 1,048,576 tokens Primary source[7]Token limits: input token limit [3]Model Information: Inputs
    • Parameters Not disclosedNot disclosed by Google. No claim madeNo source
    • Access APIGemini API in Google AI Studio and Vertex AI. Primary source[1]2.0 Flash-Lite section [2]

    Lineage

    Predecessors

    No known predecessor.

    Successors

    Based on

    Not derived from another model.

    Variants and derived

    None recorded.

    Siblings

    None recorded.

    All ancestors

    None.

    All descendants

    Variants

    No variants recorded in this record.

    Related AI Radar coverage

    AI Radar coverage starts in June 2026; no coverage linked yet.

    All model releases from Google on AI Radar →