Post
-
Quick comparison of reported DeepSeek-V4-Flash-0731 benchmark results vs Claude vs GPT Looks like we will have GPT-5.6-Luna-ish at home (which just had a huge price cut) Likely to be priced competitively from other inference providers in the long run as well Compare @ArtificialAnlys Cost per Intelligence Index Task belowImage hidden -
I am not sure how accurate this is since 0731 weights should perform substantially better, increasing the denominator in (cost / intelligence), hence beating the preview weights However, ds4 flash and gpt 5.6 luna all being clustered at the frontier makes senseImage hidden