What Is Qwen3.8-Max? The AI That Ran Alone for 125 Hours

What Is Qwen3.8-Max? The AI That Ran Alone for 125 Hours

Summary: Qwen3.8-Max in 30 Seconds Qwen3.8-Max is Alibaba’s new flagship model, made generally available on August 2, 2026. 2.4 trillion parameters, 95 billion active (MoE architecture). It reads text, images and video, and returns text. Context window is in the 1 million token class. On one task it ran 125 hours (about 5 days) with no human input, rebuilding an experiment from a machine learning paper from scratch, confirming its six findings, then inventing a method that beats the paper. It beats Claude Opus 4.8 on most coding and agent tests, trades blows with Claude Fable 5 and GPT-5.6 Sol, and falls behind on some. API pricing is $2 input / $6 output per million tokens. Repeated input costs $0.25. This is the first time Alibaba will open-weight a Max-class model. “Next week” is the word, but there is no date, license or model card yet. Ask an AI to “rebuild the experiment in this paper, then improve on it” and it normally stalls after a few turns, waiting for you to step in and steer. ...

August 3, 2026 ·  18 min ·  3622 words
Gemini 3.6 Flash and 3.5 Flash-Lite Introduced

Gemini 3.6 Flash and 3.5 Flash-Lite Introduced

On July 21, 2026, Google announced three new Gemini models at once: Gemini 3.6 Flash, Gemini 3.5 Flash-Lite, and the defense-focused Gemini 3.5 Flash Cyber. The theme of the announcement is clear: not higher scores, but getting the same job done with fewer tokens. As agentic workflows move into production, the bill is driven less by how smart a model is and more by how many tokens that intelligence burns. Google aimed straight at that. Here are the details. ...

July 21, 2026 ·  9 min ·  1753 words
What Is Kimi K3? The AI That Designed a Chip in 48 Hours

What Is Kimi K3? The AI That Designed a Chip in 48 Hours

TL;DR: Kimi K3 in 30 Seconds Kimi K3 is the new AI model Chinese company Moonshot AI announced on July 16, 2026; at 2.8 trillion parameters, it is the largest open source model released so far. In a single 48-hour run it designed a working chip, finished an astrophysics study in 2 hours instead of 2 weeks, and cut a teaser video from 56 raw clips. It reads 1 million tokens at once: roughly like reading the entire Lord of the Rings trilogy in one sitting and remembering all of it. On coding it plays in the same league as the strongest Claude and ChatGPT models, and beats them on some tests. It falls behind on hard knowledge questions. You can try it in the Kimi app and on kimi.com. The model files go public on July 27, 2026. Ask an AI to design a chip for you and what happens? It probably writes you a nice article about how chips are designed. ...

July 17, 2026 ·  10 min ·  2058 words
Grok 4.5 Is Here! Cheaper and 4.2x More Efficient AI

Grok 4.5 Is Here! Cheaper and 4.2x More Efficient AI

A powerful move has arrived on the price and efficiency front of the AI race! On July 8, 2026, xAI, Elon Musk’s AI company now operating under the SpaceX umbrella, announced Grok 4.5, its first major model release since going public. The model is also the fruit of a collaboration with the Cursor team, and its pitch is clear: not the highest benchmark score, but the most efficient way to get work done. ...

July 8, 2026 ·  4 min ·  719 words
Claude Sonnet 5: The Next-Generation Agentic AI

Claude Sonnet 5: The Next-Generation Agentic AI

The cards are being redealt in the world of artificial intelligence! Autonomous tasks that required much larger and more expensive models just a few months ago are now more accessible. Meet Anthropic’s most capable agent to date: Claude Sonnet 5. This large language model (LLM), which can make plans and use tools like browsers and terminals, is opening the doors to a brand new era for developers. If you have used Claude Sonnet 4.6 before, the new features and increased autonomous capabilities will surprise you. ...

July 1, 2026 ·  4 min ·  651 words