Claude Sonnet 4.6
Claude Sonnet 4.6 is a verified AI model profile covering official specifications, pricing or access, capabilities, practical use…
Gemini 3.7 Flash is Google's stable 1M-context coding and agent model with 65K output, multimodal inputs and introductory $0.75/$3.75 pricing.
Gemini 3.7 Flash is Google's GA workhorse for coding and agents, with 1M multimodal context, 65K output and a broad built-in tool suite.
Gemini 3.7 Flash is a generally available Google model built as a fast production workhorse for coding, agents and reliable multi-step execution.
Gemini 3.7 Flash sits in an interesting position after the newer 3.8 Flash release. It is no longer Google’s most intelligent Flash model, but it remains a stable GA endpoint designed for scaled production. That distinction matters for teams that prefer a mature deployment target over immediately moving to the newest generation. Google highlighted software engineering, web development and agentic reliability as the main improvements when 3.7 became GA.
The model accepts text, images, video, audio and PDFs, with a 1,048,576-token input limit and 65,536-token output limit. Thinking can be set to low, medium or high. It supports function calling, structured outputs, code execution, file search, computer use, Search grounding, Maps grounding and URL context. This makes it more than a chat model: it is designed to sit inside workflows that combine reasoning with external information and actions.
For software engineering, a team can give the model a large repository, design screenshots and issue context, then use tools to inspect and modify code. In operations, it can analyze PDFs and recordings, call internal functions and return structured results. Search grounding can provide current information, while Maps grounding makes it useful for location-aware workflows. Batch and Flex inference also create options for separating interactive from asynchronous workloads.
Gemini 3.7 Flash is Stable/GA. Google released it on August 13, 2026 and still lists it as a production model after the launch of 3.8 Flash. There is no need to classify it as deprecated simply because a newer Flash exists. It remains a legitimate choice where existing evaluations, latency profiles or integration behavior favor 3.7.
Through December 31, 2026, standard paid pricing is $0.75/M input and $3.75/M output, with cheaper Batch and Flex rates. Google plans to raise standard pricing to $1.50/M input and $7.50/M output in 2027. Search and Maps grounding can add separate query charges after included monthly allowances, so a grounded agent’s true cost is more than token price alone.
Choose Gemini 3.7 Flash for stable production coding, agents and multimodal workflows when its quality meets your needs. Evaluate 3.8 Flash for new systems that need Google’s strongest Flash performance. Use Flash-Lite tiers when high-volume simple tasks are more important than sophisticated reasoning, and use specialized image/video/audio models when the output modality is not text.
Google does not state a knowledge-cutoff date on the current 3.7 model page, so a directory should not invent one. Computer use remains a preview capability, which means operational behavior and limits can evolve. Tool-based systems also need validation: a model can select the wrong function, misunderstand a document or produce plausible but incorrect structured data even when the JSON schema is valid.
In addition, its main capabilities include 1M context, 65K output, Thinking levels, Coding, Function calling and Structured outputs. For example, common use cases include Coding, Agents, Web development, Enterprise workflows and Multimodal document analysis.
In practice, this model may suit Coding agents, Web development, Enterprise automation, Multimodal analysis, Long-context workflows and High-volume reasoning. Also, notable strengths include GA production model, Large context, Strong agent/coding focus and Introductory low price. However, review trade-offs such as Knowledge cutoff not stated on current model page, Computer use is preview and Newer 3.8 Flash may be preferable for the hardest agentic work before adopting it.
Meanwhile, Introductory pricing through December 31, 2026: $0.75/M input and $3.75/M output. From January 1, 2027 Google lists $1.50/M input and $7.50/M output. Through the end of 2026, Gemini 3.7 Flash uses promotional $0.75/M input and $3.75/M output pricing, with cheaper Batch/Flex processing available.
Use the official model website, official documentation, pricing or release source and additional primary source to confirm current availability, limits and pricing. Product details can change after publication, so rely on primary documentation for final decisions.
Next, continue your research in the AI models directory, Google models and General Purpose Language Model models. Compare providers, pricing, modalities and practical limitations side by side to choose the right model for your workflow.
Audit this web application against the design screenshots and fix the mismatches.Use tools to complete this multi-step coding issue and verify the result.Analyze these PDFs, images and recordings and return a structured operational report.Introductory pricing through December 31, 2026: $0.75/M input and $3.75/M output. From January 1, 2027 Google lists $1.50/M input and $7.50/M output.
A strong production workhorse for coding and agents, especially where teams value stable GA status and Google’s built-in grounding/tool ecosystem.