@sky.wav: Their new V4 Flash 0731 scored 50 on the Intelligence Index (+10 points) — matching GPT-5.6 Luna and beating GLM 5.2, which was the previous SOTA just a month ago. But the insane part? It’s ~70% smaller than GLM 5.2… while outperforming it. Same ultra-efficient architecture: 284B total parameters but only 13B active, full 1 million token context. This means high-end intelligence on a single RTX 5090 is coming way sooner than the 18-month predictions. API pricing is still only $0.14 / $0.28 per million tokens — roughly 60% cheaper than OpenAI even after their big price cut, with 98% cache hit discount making repeated work almost free. Agentic performance? Elo rating jumped from 1189 → 1559. Frontend Coding Arena? 7st overall with 1586 points (+154 from preview), 3st among open models, and the best performance-per-dollar of any model. Token usage down. Hallucinations down. Meanwhile Claude Sonnet is looking expensive and mid Chinese models just delivered better intelligence + much smaller size + lower cost all at once. Local AI just got massively accelerated 🇨🇳 Western labs… are y’all okay?? 👀 #deepseekv4 #chatgpt #aimodel #bestai #aifyp