90.6%
Terminal-Bench 2.1
Test DeepSeek V4.1 Flash before paying flagship rates
Native vision, a 1M-token context window, MIT weights, and off-peak API rates of $0.15/M input and $0.60/M output make this the freshest value candidate for agentic coding.
Your move: Run it against your coding eval set, then keep a premium fallback for the failures that matter.