whichllm — Browse and compare AI model specs and pricing

Venice AI

Gemini 3.5 Flash-Lite model ID, context window & pricing

gemini-flash-lite

Quick facts

Model ID gemini-3-5-flash-lite
Source Venice AI
Context Window 1000000
Pricing $0.38 input / $3.12 output per 1M tokens
Capabilities tool calling, reasoning, structured output, temperature control

Model overview

Gemini 3.5 Flash-Lite is an AI model from Venice AI with 1000000 token context window and text, image, audio, video input support.

Published pricing is $0.38 input and $3.12 output per 1M tokens.

  • Workloads that use text, image, audio, video inputs with text outputs.
  • Agent and tool workflows that need function calling.
  • Reasoning-heavy prompts where stepwise problem solving matters.
Model ID gemini-3-5-flash-lite
Provider Venice AI
Family gemini-flash-lite
Status -
Knowledge Cutoff 2026-03
Release Date 2026-07-09
Input Modalities text, image, audio, video
Output Modalities text
Context Window 1000000
Input Limit -
Output Limit 65536
Tool Calling Yes
Reasoning Yes
Structured Output Yes
Temperature Control Yes
Open Weights No
Input Cost / 1M tokens $0.38
Output Cost / 1M tokens $3.12
Reasoning Cost / 1M tokens -
Cache Read Cost / 1M tokens $0.04
Cache Write Cost / 1M tokens -