Gemini 3.7 Flash: Benchmarks, Pricing, Specs

Gemini 3.7 Flash: Benchmarks, Pricing, Specs

Summary: Gemini 3.7 Flash in 30 Seconds Gemini 3.7 Flash landed on August 13, 2026, only three weeks after 3.6 Flash. Introductory pricing is $0.75 in / $3.75 out per 1M tokens: half of 3.6 Flash’s standard rate. The promo ends December 31, 2026. Coding jumped hard: FrontierCode 1.1 went from 34.4% to 43.6%, DeepSWE v1.1 from 48.6% to 65.3%. Context window is 1,048,576 tokens with a 65,536 token output cap. It takes text, images, video, audio and PDF. It is not the smartest model on the board: 56 on the Artificial Analysis index, one point behind GPT-5.6 Terra and Muse Spark 1.2, at roughly a third of their price. Google’s Flash line has never been about winning the intelligence crown. It is the model you run at volume: cheap enough to call thousands of times a day, fast enough to sit inside a product. Gemini 3.7 Flash stretches that definition. ...

August 13, 2026 ·  10 min ·  2024 words
DeepSeek V4 Pro 0813: Price and Benchmarks

DeepSeek V4 Pro 0813: Price and Benchmarks

Summary: DeepSeek V4 Pro in 30 Seconds DeepSeek V4 Pro 0813 went generally available on August 13, 2026. No press release, no blog post, just one line in a changelog. 1.7 trillion parameters, mixture-of-experts architecture, 1M token context, 384K token output ceiling. The weights are on Hugging Face under the MIT license. You can download them, modify them and ship them in a commercial product. Official scores are bold: 87.9 on Terminal Bench 2.1, 62.7 on DeepSWE. Independent testing is more restrained: 53 on the Artificial Analysis Intelligence Index. Pricing is $0.435 in and $0.87 out per million tokens. From August 16 the off-peak rate is half of that. The catch: its smaller sibling V4 Flash scores 52 on the same index at one third of the price. You are paying triple for one point. Model launches have settled into a ritual: a teaser video, a wall of benchmark charts, an excited founder post. DeepSeek skipped all of it. DeepSeek V4 Pro 0813 went generally available on August 13, 2026 with no announcement at all. One line landed in the API changelog and the weights appeared on Hugging Face. ...

August 13, 2026 ·  11 min ·  2150 words
Grok 4.6 Is Here! Benchmarks, Price and What Changed

Grok 4.6 Is Here! Benchmarks, Price and What Changed

Summary: Grok 4.6 in 30 Seconds Grok 4.6 is SpaceXAI’s new flagship model, announced on August 12, 2026. It is a post-training upgrade on the Grok 4.5 base, not a new foundation model. It scores 61 on the Artificial Analysis Intelligence Index: five points above Grok 4.5 from a month ago, level with GPT-5.6 Sol, two points behind Claude Opus 5. Pricing did not move: $2 per 1M input tokens and $6 per 1M output. Roughly half what rivals charge. A new reasoning tier arrived: xhigh. The model also checks and verifies its own work far more often. The context window is 500K tokens, but any request above 200K doubles the price of the whole call. In the AI race, a month is a long time. In early July we were writing about Grok 4.5. Five weeks later SpaceXAI announced Grok 4.6 and put the model straight back into the frontier conversation. ...

August 13, 2026 ·  9 min ·  1754 words
Gemini 3.6 Flash and 3.5 Flash-Lite Introduced

Gemini 3.6 Flash and 3.5 Flash-Lite Introduced

On July 21, 2026, Google announced three new Gemini models at once: Gemini 3.6 Flash, Gemini 3.5 Flash-Lite, and the defense-focused Gemini 3.5 Flash Cyber. The theme of the announcement is clear: not higher scores, but getting the same job done with fewer tokens. As agentic workflows move into production, the bill is driven less by how smart a model is and more by how many tokens that intelligence burns. Google aimed straight at that. Here are the details. ...

July 21, 2026 ·  9 min ·  1754 words
Grok 4.5 Is Here! Cheaper and 4.2x More Efficient AI

Grok 4.5 Is Here! Cheaper and 4.2x More Efficient AI

A powerful move has arrived on the price and efficiency front of the AI race! On July 8, 2026, xAI, Elon Musk’s AI company now operating under the SpaceX umbrella, announced Grok 4.5, its first major model release since going public. The model is also the fruit of a collaboration with the Cursor team, and its pitch is clear: not the highest benchmark score, but the most efficient way to get work done. ...

July 8, 2026 ·  4 min ·  719 words