DeepSeek’s ‘kill line’: why a model 85 times cheaper is resetting the default

DeepSeek V4 Flash is redrawing how models are judged. On OpenCode, a single AI coding tool, the model burned 8 trillion tokens in one day, more than the entire OpenRouter platform’s daily average of about 6.6 trillion across 400-plus models. The reason is price. DeepSeek charges about 2 yuan per million output tokens; Anthropic’s Claude … Read more

Alibaba’s QwenWork enters the AI-office war with Qwen3.8-Max underneath

On 3 August Alibaba opened public beta for QwenWork, its enterprise agent product, folding together QoderWork, MuleRun and Wukong into one entry for local desktop control, long cloud tasks and team collaboration. It ships alongside Qwen3.8-Max, Alibaba’s new flagship: 2.4 trillion total parameters, 95 billion active per inference, a 1-million-token context and vision. In a … Read more

OpenAI slashes GPT-5.6 API prices by up to 80 per cent

On 31 July, OpenAI cut API prices across the GPT-5.6 family. The steepest cut hits Luna, down 80 per cent, input to $0.2 per million tokens, output to $1.2. The balanced Terra model drops 20 per cent, input $2, output $12 per million. The flagship Sol holds its price but gains a Fast mode running … Read more

China Launches an IPv6 Push for Large AI Models, With Shenzhen in the Mix

On 28 July, the Cyberspace Administration of China, together with the cyberspace offices of Beijing, Shanghai, Zhejiang and Shenzhen and five leading large-model firms, launched the AI Large-Model IPv6 Capability Enhancement Initiative in Xiong’an. The goal is to push generative large-model applications to fully support IPv6, laying the network foundation for the intelligent era. As … Read more