Google’s fast multimodal model for coding and agents, released 13 August 2026, three weeks after 3.6 Flash. Google pitches it against Claude Sonnet 5 and GPT-5.6 Terra at a lower price.
Paste a server URL, pick a model, and watch it call your tools in a real conversation. No install.
No server of your own? Leave it blank and use one of the public mock servers.
MCP Evals
Can Gemini 3.7 Flash actually complete tasks with your server?
One chat shows the tool calls work. An eval writes real tasks from your tool schemas, has Gemini 3.7 Flash drive each one, and reports where it got the wrong answer even though every call succeeded. Pick Gemini 3.7 Flash as the driver.
Gemini 3.7 Flash is the Flash release between 3.6 Flash and 3.8 Flash, built for fast agent workflows, coding and multi-step reasoning. It has a 1.05M-token window, up to 64K tokens of output, and takes text, image, video, audio and file input. It is available through the Gemini API, AI Studio and Antigravity. Artificial Analysis scores it 56 on its Intelligence Index, four points above 3.6 Flash.
These are the vendor’s own claims, not measurements of ours. Run the model against your own server to find out whether they hold for your tools.
Run the same four-step task on 3.7 Flash and 3.8 Flash against the complex-schema mock: create_user_profile, process_order, analyze_data, then configure_workflow. If 3.7 Flash makes the same calls, you can stay on it; if it drops a later step, that is the gap 3.8 closes. All the mock servers →
| Model | Model ID |
|---|---|
| Gemini 3.7 Flash | google/gemini-3.7-flash |