LLMgoogle

Google: Gemini 3.6 Flash

Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...

Anyone in the Project can @-mention Google: Gemini 3.6 Flash with the team's shared context - pooled credits, one chat, one memory.

All models

Starter is free forever - 1 Project, 100 credits/month, 1 MCP. No card.

Verdict

Gemini 3.6 Flash targets high-throughput teams that need multimodal reasoning without the latency or cost of frontier models. With a 1M token context window and $1.50/$7.50 per Mtok pricing, it handles long documents, video analysis, and audio transcription at roughly one-third the cost of GPT-4o. The trade-off: no public benchmarks yet means you're flying blind on reasoning quality relative to Claude or GPT-4 class models. Reach for this when speed and multimodal breadth matter more than proven performance on complex logic tasks.

Best for

  • High-volume document processing pipelines
  • Video content analysis and summarization
  • Audio transcription with contextual reasoning
  • Cost-sensitive multimodal prototyping
  • Long-context customer support workflows

Strengths

The 1M token context window lets you process entire codebases, multi-hour meeting recordings, or book-length documents in a single call. Multimodal support spans text, image, video, file, and audio — rare breadth at this price point. At $1.50 input and $7.50 output per million tokens, it undercuts GPT-4o by roughly 67% on input and 50% on output, making it viable for high-volume batch jobs where cost compounds quickly.

Trade-offs

No public benchmarks means you cannot compare reasoning quality, instruction-following accuracy, or coding performance against Claude Sonnet, GPT-4o, or Llama 3.3. Early Flash models historically lagged frontier models on complex multi-step reasoning and nuanced instruction adherence. The output pricing ($7.50/Mtok) climbs steeply for generation-heavy tasks like creative writing or long-form content, eroding the cost advantage. Without MMLU, HumanEval, or GPQA scores, you're testing blind.

Specifications

Provider
google
Category
llm
Context length
1,048,576 tokens
Max output
65,536 tokens
Modalities
text, image, video, file, audio
License
proprietary
Released
2026-07-21

Pricing

Input
$0.75/Mtok
Output
$3.75/Mtok
Model ID
google/gemini-3.6-flash

Per-token prices show what the model costs upstream. On Switchy your team draws from one shared org credit pool - one plan, one balance for everyone.

Team cost calculator

Estimated monthly spend
$29.04
17.6M tokens / month
5 seats · 80 msgs/day

Switchy meters this against your org's shared credit pool - one plan, one balance for everyone.

Providers

Provider-level routing data is not available yet for this model.

Performance

Performance snapshots are collected daily. Check back after the next ingestion run.

Benchmarks

Public benchmark scores are not available yet for this model. Check back after the next ingestion run.

Works well with

Top MCPs

Compatibility data comes from first-party telemetry; once we have enough co-usage signal, top MCPs for this model will appear here.

How Switchy teams use it

Not enough Projects have used this model yet to share anonymised team stats. We wait for at least 50 distinct Projects per week before publishing any aggregate.

Starter prompts

Summarize Long Video

Watch this video and provide a structured summary with timestamps for each major topic discussed. Highlight any action items or decisions made.
Open in a Project →

Analyze Support Transcript

Review this support chat transcript and identify: 1) the root cause of the issue, 2) which response resolved it, 3) any gaps in our help documentation.
Open in a Project →

Extract Data from PDFs

Extract all line items, dates, and totals from this invoice PDF. Return the data as a JSON array with fields: description, quantity, unit_price, total.
Open in a Project →

Transcribe and Tag Audio

Transcribe this meeting recording. Label each speaker, tag topics discussed, and list all action items with assigned owners.
Open in a Project →

Compare Document Versions

Compare these two contract versions and list every substantive change. Ignore formatting differences and focus on terms, clauses, and obligations.
Open in a Project →

Compare with

More language models

See all language models

Data last verified 7 hours ago.Sources aggregated hourly to weekly. See docs/architecture/model-directory.md.