DeepSeek V4.1 Flash Benchmarks: Open-Weights Model Beats GPT-5.6 Sol at Agentic Coding
DeepSeek V4.1 Flash (released 10 September 2026) beats GPT-5.6 Sol and Claude Opus-5.0 on DeepSWE, AutomationBench, Agent's Last Exam and CyberGym. 552B open-weights MoE with 1M-token context and $0.30 per million peak input pricing. Full benchmark tables, pricing maths and what it means for Australian teams.