Qwen3.8-Omni-Flash beats Gemini 3.8 Flash on WildClawBench (71.0 vs 58.9) and AliMeeting (89.7 vs 37.1) at 4.2x lower video cost. Read our full teardown.
Within eight days in early September 2026, DeepSeek-V4.1-Flash and Gemini 3.8 Flash toppled previous-generation $90/M-token flagship models on DeepSWE v1.1. Here is the forensic engineering breakdown: 890-byte KV cache, CED topology, RLVR benchmaxxing, and real-world agent TCO.
China’s Flash AI models are not a retreat from frontier AI. They are a deployment strategy built around cheaper inference, domestic chips and agents.