LLMgoogle

Google: Gemini 3.5 Flash Lite

Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.

Anyone in the Project can @-mention Google: Gemini 3.5 Flash Lite with the team's shared context - pooled credits, one chat, one memory.

All models

Starter is free forever - 1 Project, 100 credits/month, 1 MCP. No card.

Verdict

Gemini 3.5 Flash Lite trades reasoning depth for speed and cost, making it Google's cheapest multimodal option at $0.30/Mtok input. The 1M token context window handles large documents and video files, but expect weaker performance on complex reasoning compared to full Flash or Pro models. Reach for this when you need fast, cheap multimodal processing and can tolerate occasional accuracy drops on nuanced tasks.

Best for

  • High-volume content moderation pipelines
  • Quick image and video captioning
  • Cost-sensitive document extraction
  • Rapid prototyping with multimodal inputs
  • Simple classification across text and images

Strengths

At $0.30/Mtok input, this is Google's most affordable model with vision and audio support. The 1M token window lets you process hour-long video transcripts or multi-hundred-page PDFs in a single call. Multimodal capability at this price point makes it viable for high-throughput pipelines where per-request cost matters more than perfect accuracy. Latency is competitive with other Flash-tier models.

Trade-offs

No public benchmarks yet, but Flash Lite models historically sacrifice 10-15 percentage points on MMLU and reasoning evals compared to their full counterparts. Expect weaker chain-of-thought performance and more frequent hallucinations on ambiguous prompts. The $2.50/Mtok output price climbs quickly if you generate long responses. For tasks requiring careful analysis or multi-step logic, standard Gemini Flash or Pro will outperform despite higher cost.

Specifications

Provider
google
Category
llm
Context length
1,048,576 tokens
Max output
65,536 tokens
Modalities
text, image, video, file, audio
License
proprietary
Released
2026-07-21

Pricing

Input
$0.30/Mtok
Output
$2.50/Mtok
Model ID
google/gemini-3.5-flash-lite

Per-token prices show what the model costs upstream. On Switchy your team draws from one shared org credit pool - one plan, one balance for everyone.

Team cost calculator

Estimated monthly spend
$16.90
17.6M tokens / month
5 seats · 80 msgs/day

Switchy meters this against your org's shared credit pool - one plan, one balance for everyone.

Providers

Provider-level routing data is not available yet for this model.

Performance

Performance snapshots are collected daily. Check back after the next ingestion run.

Benchmarks

Public benchmark scores are not available yet for this model. Check back after the next ingestion run.

Works well with

Top MCPs

Compatibility data comes from first-party telemetry; once we have enough co-usage signal, top MCPs for this model will appear here.

How Switchy teams use it

Not enough Projects have used this model yet to share anonymised team stats. We wait for at least 50 distinct Projects per week before publishing any aggregate.

Starter prompts

Extract Invoice Line Items

Extract all line items from this invoice image as JSON. Include item description, quantity, unit price, and total. Return only valid JSON with no explanation.
Open in a Project →

Moderate User-Uploaded Images

Review this image for policy violations: nudity, violence, hate symbols, or spam. Respond with 'APPROVED' or 'FLAGGED: [reason]' in under 10 words.
Open in a Project →

Summarize Meeting Recording

Summarize this meeting recording in 5 bullet points. Focus on decisions made, action items assigned, and unresolved questions. Be concise.
Open in a Project →

Classify Support Tickets

Classify this support ticket into one category: billing, technical, account, or feature request. Add priority: low, medium, or high. Format: 'Category: X | Priority: Y'.
Open in a Project →

Generate Alt Text for Images

Write a concise alt text description for this image in one sentence. Focus on what's visible and relevant for screen readers. Max 15 words.
Open in a Project →

Compare with

More language models

See all language models

Data last verified 7 hours ago.Sources aggregated hourly to weekly. See docs/architecture/model-directory.md.