Big US labs have absolutely no competitive advantage in flash model range
- Luna, $1.2, AA score 52
- Sonnet 5, $10, AA score 55
- Haiku 4.5, $5, AA score 30
- Gemini 3.7 flash, $3.75, AA score 56
Mean while Chinese open models:
- GLM 5.3 flash, $0.5, AA score 57
- Qwen 3.8 27B, run locally, AA score 52
- Qwen 3.8 flash next, unknown yet, but will be super cheap and capable
- deepseek v4 flash, $0.66, AA score 51
The only way out for OpenAI, Anthropic, Google, and SpaceXAI is keep releasing top models and slim down profit margins with better usage. The competition is keeping them on their toes, and it’s net benefit for users