In brief
- CAISI’s evaluation ranked DeepSeek V4 Pro eight months behind the U.S. frontier, using an IRT-based scoring system across nine benchmarks including two private, unverifiable datasets.
- The cost comparison excluded all U.S. models deemed too expensive or too weak—leaving only GPT-5.4…
Read Full Article at Source